function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| git-goods/gitanimals GitAnimals is a gamification service that lets GitHub users raise virtual pets based on their commit and contribution activity, rendered as… | 70 | 1768 | active |
| pondorasti/emojis A web application that generates custom Slack emojis from text prompts using AI image generation (SDXL emoji model) via Replicate. It is a … | 57 | 1768 | active |
| GPUOpen-LibrariesAndSDKs/FidelityFX-SDK AMD's FidelityFX SDK (FSR SDK) is a collection of heavily optimized GPU techniques for DirectX 12 and Vulkan applications, including super … | 84 | 1766 | active |
| ankitdhall/lidar_camera_calibration A ROS package that computes the rigid-body transformation (rotation and translation) between a LiDAR and a camera using 3D-3D point corresp… | 44 | 1766 | active |
| Coyote-A/ultimate-upscale-for-automatic1111 An extension for the AUTOMATIC1111 Stable Diffusion web UI that upscales images to 2K/4K+ by processing them in tiled passes with diffusion… | 31 | 1766 | stable |
| leoding86/webextension-pixiv-toolkit Pixiv Toolkit Next is a WebExtension (Chrome/Firefox) that adds a download and content-management toolkit to Pixiv, letting users download … | 34 | 1763 | active |
| levihsu/OOTDiffusion Official implementation of OOTDiffusion, a latent diffusion model for controllable virtual try-on that generates images of a person wearing… | 26 | 6586 | maintenance |
| shuyu-labs/BigBanana-AI-Director BigBanana AI Director is a self-hosted, industrial-grade AI platform for generating short dramas and motion comics end-to-end, from script … | 80 | 1761 | active |
| sp4cerat/Fast-Quadric-Mesh-Simplification A fast, memory-efficient C++ library and CLI tool implementing quadric-based edge collapse mesh simplification to reduce triangle counts in… | 32 | 1759 | stable |
| xinntao/ESRGAN ESRGAN (Enhanced SRGAN) is a PyTorch-based image super-resolution model that won the PIRM 2018 Challenge on Perceptual Super-Resolution. Th… | 32 | 6568 | maintenance |
| nguyenq/tess4j Tess4J is a Java JNA wrapper for the Tesseract OCR API, enabling optical character recognition in Java applications. It supports TIFF, JPEG… | 91 | 1757 | stable |
| VAST-AI-Research/TripoSG TripoSG is an open-source image-to-3D generation foundation model that produces high-fidelity 3D meshes from single images using large-scal… | 27 | 1755 | active |
| giventofly/pixelit Pixel It is a JavaScript library that converts images into pixel art on an HTML canvas, with configurable pixel scale, color palettes, and … | 59 | 1751 | active |
| Stability-AI/StableCascade Official codebase for Stable Cascade, a text-to-image generation model built on the Würstchen architecture with a highly compressed latent … | 26 | 6540 | maintenance |
| mozmorris/react-webcam react-webcam is a React component that wraps the browser's getUserMedia API to display a live webcam video stream. It supports capturing sc… | 64 | 1750 | active |
| dmester/jdenticon Jdenticon is a JavaScript library for generating deterministic, highly recognizable identicon avatars from hash values, rendered via HTML5 … | 23 | 1748 | stable |
| CompVis/taming-transformers The official implementation of 'Taming Transformers for High-Resolution Image Synthesis' (CVPR 2021), combining a convolutional VQGAN codeb… | 32 | 6521 | maintenance |
| rougier/freetype-gl A small C library for rendering Unicode text in OpenGL using a single texture atlas and a single vertex buffer, built on top of FreeType. I… | 65 | 1745 | stable |
| TencentARC/BrushNet BrushNet is the official PyTorch implementation of an ECCV 2024 plug-and-play image inpainting model that embeds pixel-level masked image f… | 25 | 1745 | active |
| sindresorhus/pageres-cli A Node.js command-line tool that captures screenshots of websites at multiple resolutions using headless Chrome (Puppeteer). It is useful f… | 44 | 1743 | active |
| auduno/clmtrackr clmtrackr is a JavaScript library for fitting facial models to faces in videos or images using Constrained Local Models with regularized la… | 23 | 6500 | maintenance |
| kirevdokimov/Unity-UI-Rounded-Corners A Unity package providing components and shaders that add rounded corners to UGUI Image elements. It supports uniform or per-corner radii, … | 23 | 1741 | stable |
| openai/consistency_models Official PyTorch implementation of Consistency Models, a generative image model family from OpenAI supporting consistency distillation, con… | 10 | 6486 | maintenance |
| Xiaojiu-z/EasyControl EasyControl is the official implementation of an ICCV 2025 paper adding efficient and flexible conditional control to Diffusion Transformer… | 35 | 1737 | active |
| kristerkari/react-native-svg-transformer A Metro transformer for React Native that lets you import SVG files as React components, using SVGR under the hood. It enables sharing the … | 80 | 1736 | stable |
| SHI-Labs/OneFormer OneFormer is a universal image segmentation framework (CVPR 2023) that unifies semantic, instance, and panoptic segmentation in a single tr… | 32 | 1736 | stable |
| google/automl Google Brain's AutoML repository containing implementations of AutoML models and libraries such as EfficientNet, EfficientNetV2, and Effici… | 10 | 6474 | maintenance |
| NimaNzrii/comfyui-photoshop A Photoshop plugin that integrates ComfyUI's AI image generation directly into the Photoshop workspace. It lets users run Stable Diffusion … | 71 | 1734 | active |
| i12bp8/TagTinker TagTinker is a Flipper Zero application for researching infrared electronic shelf label (ESL) protocols, letting users transmit custom imag… | 50 | 1734 | active |
| kmonkeyhead/MORT MORT is a Windows desktop application that extracts on-screen text in real time using OCR and translates it via databases or machine transl… | 99 | 1728 | active |
| jacebrowning/memegen A free and open source REST API for programmatically generating meme images from URLs, built with Python, Sanic, and Pillow. It is stateles… | 77 | 1728 | active |
| duerrsimon/bioicons Bioicons is a free, open-source library of SVG vector icons for scientific illustrations in biology and chemistry, browsable through a web … | 67 | 1728 | active |
| mahmoodlab/CLAM CLAM is an open-source Python toolkit for data-efficient, weakly supervised classification of whole-slide images (WSIs) in computational pa… | 39 | 1728 | active |
| freestylefly/director_ai An AI-powered mobile application for creating comic-style short dramas, generating scripts, storyboards, and composed videos from a single … | 52 | 1726 | active |
| liuruoze/EasyPR EasyPR is an open-source C++ library built on OpenCV for recognizing Chinese license plates in unconstrained situations, outputting plate c… | 23 | 6429 | maintenance |
| xtyxtyx/sorry A web application for creating custom captioned GIF memes (Chinese 'sorry 为所欲为' style) from popular video templates, with subtitle effect c… | 32 | 6428 | maintenance |
| HITsz-TMG/VideoClaw VideoClaw is an AI director system that turns a one-line idea or story synopsis into a fully automated video production pipeline, covering … | 67 | 1724 | active |
| MultimediaTechLab/YOLO Official MIT-licensed implementation of the YOLOv9, YOLOv7, and YOLO-RD real-time object detection models, including pre-trained weights, t… | 56 | 1723 | active |
| facebookresearch/ConvNeXt Official PyTorch implementation of ConvNeXt, a pure convolutional neural network architecture from the CVPR 2022 paper 'A ConvNet for the 2… | 10 | 6416 | maintenance |
| yoshi389111/github-profile-3d-contrib A GitHub Action that generates a 3D visualization of a user's GitHub contribution calendar as an SVG image. It commits the generated image … | 85 | 1722 | active |
| jau123/MeiGen-AI-Design-MCP An open-source MCP server that adds AI image and video generation capabilities to AI coding tools like Claude Code, Cursor, and Codex. It s… | 80 | 1719 | active |
| sbrin/lopaka Lopaka is a browser-based graphics editor for designing pixel-perfect UIs for embedded displays, with code generation for libraries like U8… | 72 | 1719 | active |
| One-2-3-45/One-2-3-45 One-2-3-45 is the official PyTorch implementation of a NeurIPS 2023 paper that converts any single image into a full 360-degree 3D textured… | 29 | 1718 | stable |
| verlab/accelerated_features XFeat is a lightweight, fast learned keypoint detector and descriptor for local feature extraction and image matching, supporting both spar… | 16 | 1716 | active |
| cocodataset/cocoapi Official API for the COCO (Common Objects in Context) dataset, providing Matlab, Python, and Lua interfaces to load, parse, and visualize C… | 32 | 6384 | maintenance |
| Donaldcwl/browser-image-compression A JavaScript library that compresses jpeg, png, webp, and bmp images directly in the web browser by reducing resolution or storage size. It… | 23 | 1714 | stable |
| sambecker/exif-photo-blog A self-hostable Next.js photo blog application that extracts and displays EXIF camera details (aperture, shutter speed, ISO, lens, film sim… | 72 | 1713 | active |
| liip/LiipImagineBundle LiipImagineBundle is a Symfony bundle providing an image manipulation abstraction toolkit built on the Imagine library. It lets developers … | 94 | 1712 | active |
| kha-white/mokuro mokuro is a Python tool that performs text detection and OCR on Japanese manga pages and generates overlay files (.mokuro or HTML) enabling… | 86 | 1712 | active |
| imageio/imageio Imageio is a mature Python library for reading and writing image and video data, including animated images, volumetric data, and scientific… | 92 | 1711 | stable |
| lukechilds/merge-images A small JavaScript library that composites multiple images into one, abstracting away canvas boilerplate into a single promise-based functi… | 69 | 1710 | stable |
| skydoves/ColorPickerView ColorPickerView is an Android UI library that lets users pick colors by tapping on an HSV color wheel, custom drawable palettes, or gallery… | 61 | 1710 | stable |
| 0beqz/realism-effects A collection of postprocessing effects for three.js that enhance scene realism, including screen-space global illumination (SSGI), motion b… | 32 | 1710 | active |
| electerious/Lychee Lychee is a self-hosted, open-source photo-management web application written in PHP that lets users upload, organize, search, and share ph… | 23 | 6357 | maintenance |
| omerbt/TokenFlow TokenFlow is the official PyTorch implementation of an ICLR 2024 paper for text-driven, temporally consistent video editing using a pre-tra… | 30 | 1708 | stable |
| salgum1114/react-design-editor A React library providing a canvas-based design editor built on Fabric.js and Ant Design, supporting image editing, diagram composition, an… | 99 | 1707 | active |
| blendi-remade/sprite-sheet-creator A Next.js web application that generates 2D pixel art sprite sheets and game maps from text prompts or uploaded images using fal.ai image m… | 56 | 1706 | active |
| ggchivalrous/yiyin Yiyin (壹印) is a free, open-source desktop application for adding watermark frames to photos, generating styled image borders with customiza… | 83 | 1703 | active |
| iMoonLab/yolov13 Official PyTorch implementation of YOLOv13, a real-time object detection model family (Nano to X-Large) featuring Hypergraph-based Adaptive… | 32 | 1702 | active |
| fangfufu/Linux-Fake-Background-Webcam A Python application that creates a virtual webcam on GNU/Linux with fake backgrounds, including background replacement, blurring, animated… | 63 | 1700 | active |
| sihyun-yu/REPA REPA is the official PyTorch implementation of the ICLR 2025 paper 'Representation Alignment for Generation', a regularization technique th… | 27 | 1700 | active |
| kingsic/SGQRCode SGQRCode is an easy-to-use iOS library for scanning barcodes and QR codes, generating QR codes, and recognizing QR codes from images. It pr… | 23 | 1697 | stable |
| divriots/jampack Jampack is a post-processing CLI tool that optimizes the output of static site generators for best user experience and Core Web Vitals scor… | 70 | 1696 | active |
| stephansturges/WALDO WALDO is an open-source object detection model based on a YOLOv8 backbone, trained with a synthetic data pipeline to detect people, vehicle… | 32 | 1695 | active |
| baskerville/plato Plato is a document reader application for Kobo e-ink e-readers, written in Rust. It supports PDF, EPUB, DJVU, CBZ, FB2, MOBI, XPS and TXT … | 65 | 1694 | active |
| NVlabs/InstantSplat InstantSplat is a research framework for photorealistic 3D scene reconstruction from extremely sparse image views using Gaussian Splatting,… | 33 | 1694 | active |
| dnfield/flutter_svg A Dart/Flutter library that parses SVG files and renders them as Flutter widgets. Originally created by Dan Field, it is now maintained by … | 23 | 1691 | active |
| GreycLab/CImg CImg is a small, open-source, header-only C++ template library for image processing. It provides a single image class supporting up to 4-di… | 76 | 1690 | stable |
| xCss/bing A web service that serves Bing's daily wallpaper images with a browsable gallery and a previously offered API for fetching wallpapers by da… | 56 | 1690 | active |
| adamlyttleapps/claude-skill-aso-appstore-screenshots A Claude Code skill that generates high-converting App Store screenshots for iOS apps by analyzing the codebase to identify core benefits a… | 48 | 1687 | active |
| kijai/ComfyUI-FramePackWrapper A ComfyUI custom node wrapper for FramePack, enabling next-frame video generation from images using HunyuanVideo-based diffusion models. It… | 47 | 1686 | active |
| sedthh/pyxelate Pyxelate is a Python library and CLI tool that converts images into 8-bit pixel art by downsampling and learning a reduced color palette. I… | 51 | 1685 | active |
| koush/ion Ion is an Android library for asynchronous networking and image loading, built on NIO and AndroidAsync. It provides a fluent API for downlo… | 76 | 6243 | maintenance |
| stefanhaustein/TerminalImageViewer Terminal Image Viewer (tiv) is a small C++ command-line program that displays images directly in modern terminals using RGB ANSI codes and … | 66 | 1683 | active |
| cubewhy/skid-homework A browser-based, AI-powered homework solver built with Next.js that sends images and PDFs of homework problems to a Gemini or OpenAI-compat… | 61 | 1683 | active |
| williamyang1991/DualStyleGAN Official PyTorch implementation of DualStyleGAN, a CVPR 2022 model for exemplar-based high-resolution (1024px) portrait style transfer. It … | 32 | 1683 | stable |
| CatimaLoyalty/Android Catima is a free, open-source Android app for storing loyalty cards and tickets with a built-in barcode scanner. It works fully offline, co… | 98 | 1678 | active |
| mebjas/html5-qrcode A lightweight, zero-dependency JavaScript/TypeScript library for scanning QR codes and barcodes in the browser using the device camera or l… | 48 | 6213 | maintenance |
| fire-keeper/BlindWatermark A Python library and CLI/GUI tool that embeds invisible blind watermarks into images using discrete wavelet transforms, protecting creators… | 23 | 1675 | active |
| franciszzj/Leffa Leffa is a diffusion-based framework for controllable person image generation, supporting virtual try-on and pose transfer via a regulariza… | 40 | 1672 | active |
| naomiaro/waveform-playlist A multitrack web audio editor and player library built on the Web Audio API, Tone.js, and React (with a Web Components alternative), featur… | 94 | 1671 | active |
| LiberatedPixelCup/Universal-LPC-Spritesheet-Character-Generator A web-based character generator that assembles Liberated Pixel Cup (LPC) pixel art assets into customizable, game-ready character spriteshe… | 77 | 1671 | active |
| dgmjs/dgmjs DGM.js is a TypeScript library providing an infinite canvas with programmable 'smart shapes' for building diagramming, whiteboarding, and s… | 62 | 1671 | active |
| ZrrSkywalker/Personalize-SAM PerSAM is the official implementation of 'Personalize Segment Anything Model with One Shot', which customizes the Segment Anything Model (S… | 29 | 1671 | active |
| tkarras/progressive_growing_of_gans Official TensorFlow implementation of the ICLR 2018 NVIDIA paper 'Progressive Growing of GANs', which trains generators and discriminators … | 32 | 6179 | maintenance |
| foliojs/fontkit fontkit is an advanced font engine library for Node.js and the browser, used by PDFKit. It parses many font formats and provides glyph shap… | 23 | 1668 | stable |
| opendatalab/labelU LabelU is an open-source multimodal data annotation platform supporting images, video, and audio with tools like bounding boxes, segmentati… | 96 | 1665 | active |
| asny/three-d three-d is a Rust library providing an OpenGL/WebGL/OpenGL ES renderer for drawing 2D and 3D graphics across desktop, web, and mobile platf… | 63 | 1663 | active |
| bamlab/react-native-image-resizer A React Native library for resizing and compressing local images on iOS and Android. It supports JPEG/PNG/WEBP formats, rotation, quality c… | 67 | 1661 | active |
| ermig1979/AntiDupl AntiDupl.NET is a free open-source Windows desktop application that finds duplicate and similar images on disk by comparing file contents, … | 70 | 1659 | active |
| huggingface/gsplat.js gsplat.js is an open-source JavaScript/TypeScript library for 3D Gaussian Splatting, offering scene, camera, loader, and WebGL renderer com… | 67 | 1659 | active |
| goldvideo/h265player A complete web-based H.265 (HEVC) video player solution built in JavaScript. It uses JS-based stream demuxing, WebAssembly (FFmpeg) video d… | 70 | 1656 | active |
| davidbyttow/govips govips is a Go library that wraps the libvips image processing library, exposing fast image operations like resizing, format conversion, an… | 79 | 1654 | active |
| mhamilton723/FeatUp FeatUp is a model-agnostic framework that upsamples the spatial resolution of deep neural network features by 16-32x without changing their… | 16 | 1654 | active |
| taki0112/UGATIT Official TensorFlow implementation of U-GAT-IT, an unsupervised image-to-image translation model using attention modules and adaptive layer… | 32 | 6116 | maintenance |
| cubiq/ComfyUI_IPAdapter_plus A ComfyUI custom node implementation of IPAdapter models for image-to-image conditioning in Stable Diffusion workflows. It transfers the su… | 35 | 6110 | maintenance |
| Stability-AI/stable-virtual-camera Stable Virtual Camera (SEVA) is a generalist diffusion model for novel view synthesis that generates 3D-consistent views of a scene from an… | 54 | 1652 | active |
| joaomoreno/gifcap gifcap is a browser-based screen recorder that captures your screen or a single window and encodes it into an optimized animated GIF entire… | 60 | 1651 | active |
| LiangliangNan/Easy3D Easy3D is a lightweight C++ library (with Python bindings) for processing and rendering 3D data such as point clouds, polygonal surface mes… | 74 | 1650 | active |
| fluttercandies/flutter_wechat_assets_picker A Flutter package providing an image, video, and audio picker with a UI modeled on WeChat's asset picker. It is built on photo_manager and … | 96 | 1649 | active |