domain: image-processing
1843 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| cbh123/stickerbaker StickerBaker is an open-source web application that generates AI stickers from text prompts or face uploads, powered by Replicate models (f… | 25 | 1100 | active |
| eszdman/PhotonCamera PhotonCamera is an open-source Android camera app that applies enhanced computational photography image processing to captured photos. It u… | 68 | 1099 | active |
| HswAI2026/JuZhou-V1 JuZhou 1.0 is an ultra-lightweight 0.387B-parameter text-to-image foundation model designed for fully offline, on-device execution on mobil… | 54 | 1097 | active |
| boona13/image-extender An open-source Next.js web app for AI image outpainting that extends images in any direction using Gemini models via OpenRouter, with Poiss… | 52 | 1096 | active |
| WangLibo1995/GeoSeg GeoSeg is an open-source PyTorch-based semantic segmentation toolbox focused on Vision Transformers for remote sensing imagery, featuring t… | 32 | 1096 | active |
| LujiaJin/One-Pot_Multi-Frame_Denoising Official PyTorch implementation of the One-Pot Multi-frame Denoising (OPD) method published at BMVC 2022 and extended in IJCV. It provides … | 60 | 1094 | stable |
| benhowdle89/grade Grade is a small JavaScript library that generates complementary gradient backgrounds from the top two dominant colors of supplied images, … | 32 | 3759 | maintenance |
| Eyeline-Labs/Go-with-the-Flow Official implementation of the CVPR 2025 Oral paper 'Go-with-the-Flow', which controls motion in video diffusion models by replacing i.i.d.… | 41 | 1093 | active |
| welltop-cn/ComfyUI-TeaCache A ComfyUI plugin integrating TeaCache, a training-free caching method that accelerates diffusion model inference by exploiting output diffe… | 35 | 1092 | active |
| jarun/imgp imgp is a fast command-line batch image resizer and rotator for JPEG and PNG images, powered by multiprocessing and Pillow-SIMD. It support… | 87 | 1091 | active |
| image-js/image-js ImageJS is a JavaScript/TypeScript library for image processing and manipulation, offering features like resizing, cropping, filtering, col… | 91 | 1090 | stable |
| mlc-ai/web-stable-diffusion A project that compiles and runs Stable Diffusion text-to-image models entirely inside web browsers using WebGPU and WebAssembly, with no s… | 30 | 3721 | maintenance |
| SimpleITK/SimpleITK SimpleITK is a simplified C++ interface to the Insight Toolkit (ITK) for multi-dimensional image analysis, including filtering, segmentatio… | 98 | 1084 | stable |
| Alpha-VLLM/Lumina-mGPT-2.0 Lumina-mGPT 2.0 is a stand-alone decoder-only autoregressive model trained from scratch that unifies a broad range of image generation task… | 42 | 1084 | active |
| openai/glide-text2im Official codebase for GLIDE, a diffusion-based text-conditional image synthesis model from OpenAI. It provides pretrained models and notebo… | 10 | 3684 | maintenance |
| lolishinshi/imsearch A Rust-based large-scale similar image search tool that uses feature point matching (ORB features with a FAISS-style index) to find full im… | 94 | 1074 | active |
| yeates/PromptFix PromptFix is a PyTorch implementation of a diffusion-model-based image restoration model that follows natural language instructions to fix … | 24 | 1070 | active |
| ototadana/sd-face-editor A Stable Diffusion Web UI extension that detects and regenerates faces in generated images to fix broken faces, change facial expressions, … | 23 | 1070 | active |
| jeffbass/imagezmq imageZMQ is a set of Python classes that transport OpenCV images between computers using PyZMQ messaging. It enables distributed computer v… | 62 | 1069 | stable |
| RL-VIG/LibFewShot LibFewShot is a comprehensive PyTorch library for few-shot learning, implementing many fine-tuning, meta-learning, and metric-learning meth… | 54 | 1069 | active |
| tomayac/SVGcode SVGcode is a Progressive Web App that converts raster images (JPG, PNG, GIF, WebP, AVIF, etc.) into SVG vector graphics. It runs in the bro… | 70 | 1066 | active |
| hujie-frank/SENet Official Caffe/CUDA implementation of Squeeze-and-Excitation Networks (SENet), channel-attention building blocks for convolutional neural n… | 32 | 3646 | maintenance |
| fossasia/magic-epaper-app Magic ePaper is an open-source Flutter mobile app for designing content and transferring it to battery-free NFC ePaper badges. It offers dr… | 77 | 1063 | active |
| rust-cv/cv Rust CV is a mono-repo of pure-Rust computer vision crates aiming to encapsulate capabilities of OpenCV, OpenMVG, and vSLAM frameworks in c… | 47 | 1063 | active |
| AIrjen/OneButtonPrompt OneButtonPrompt is a Python tool that generates complete Stable Diffusion prompts from scratch with a single click, for beginners and users… | 44 | 1062 | active |
| vijishmadhavan/ArtLine ArtLine is a deep learning project that converts portrait photos into line art portraits, with a ControlNet-based variant that adjusts styl… | 32 | 3630 | maintenance |
| clovaai/stargan-v2 The official PyTorch implementation of StarGAN v2, a CVPR 2020 paper on diverse image-to-image translation across multiple domains using a … | 32 | 3617 | maintenance |
| sb-ai-lab/EmotiEffLib EmotiEffLib (formerly HSEmotion) is a lightweight library for facial emotion and engagement recognition in photos and videos, available in … | 65 | 1057 | active |
| yoyo-nb/Thin-Plate-Spline-Motion-Model The official PyTorch implementation of the CVPR 2022 paper 'Thin-Plate Spline Motion Model for Image Animation'. It animates a source image… | 32 | 3604 | maintenance |
| hnvn/flutter_image_cropper A Flutter plugin that provides image cropping and rotation on Android, iOS, and Web by wrapping native libraries (uCrop, TOCropViewControll… | 75 | 1055 | active |
| williamyang1991/VToonify Official PyTorch implementation of VToonify, a SIGGRAPH Asia 2022 framework for controllable high-resolution portrait video style transfer … | 32 | 3584 | maintenance |
| psoho/fast-poster fastposter is a self-hostable poster/image generation service with a drag-and-drop web editor for composing text, images, QR codes, and ava… | 32 | 1050 | active |
| showlab/MotionDirector MotionDirector is a research library for customizing text-to-video diffusion models to generate videos with desired motions from a small se… | 27 | 1050 | active |
| qqlu/Entity EntitySeg is an open-source PyTorch toolbox for open-world, high-quality image segmentation, built on Detectron2. It aggregates multiple re… | 32 | 1048 | active |
| anuragxel/salt SALT is a Python-based image labeling tool built on Meta AI's Segment Anything Model, providing a barebones GUI for annotating images with … | 30 | 1048 | active |
| addyosmani/squish Squish is a browser-based batch image compression tool that uses WebAssembly codecs to compress and convert images entirely client-side. It… | 22 | 1048 | active |
| zhbhun/idify Idify is a browser-based application for creating ID, passport, and visa photos with all processing done locally in the browser. It require… | 56 | 1047 | active |
| kikoso/android-stackblur An Android library that applies a StackBlur (gaussian-like) blur effect to Bitmaps with configurable radius or gradient. It offers Java, ND… | 32 | 3566 | maintenance |
| nv-tlabs/PiD PiD is a plug-and-play pixel diffusion decoder from NVIDIA that replaces VAE/RAE decoders, decoding latent representations directly into hi… | 56 | 1045 | active |
| liuyuan-pal/SyncDreamer SyncDreamer is a synchronized multiview diffusion model that generates multiview-consistent images from a single-view image, released with … | 50 | 1045 | active |
| Tencent-Hunyuan/InstantCharacter InstantCharacter is a tuning-free framework built on diffusion transformers that generates character-consistent images from a single refere… | 29 | 1045 | active |
| muapi CLI muapi-cli and its Generative Media Skills provide a schema-driven CLI, skill library, and MCP server that let AI agents (Claude Code, Curso… | 93 | 1043 | active |
| iptag/jimeng-api A self-hosted API service that reverse-engineers Jimeng AI (China) and Dreamina (international) to expose free AI image and video generatio… | 10 | 1041 | active |
| foolwood/SiamMask Official PyTorch implementation of SiamMask, a deep learning framework for fast online visual object tracking and video object segmentation… | 35 | 3547 | maintenance |
| Nutlope/blinkshot BlinkShot is an open-source web application that generates AI images in real time as you type, powered by the Flux Schnell model via Togeth… | 67 | 1040 | active |
| ibireme/YYWebImage YYWebImage is an asynchronous image loading framework for iOS, part of YYKit, offering remote/local image loading, animated WebP/APNG/GIF d… | 23 | 3526 | maintenance |
| NVlabs/DiffusionNFT DiffusionNFT is a research library implementing an online reinforcement learning paradigm for diffusion models that optimizes policy direct… | 47 | 1034 | active |
| antimatter15/ocrad.js Ocrad.js is a pure-JavaScript port of the Ocrad OCR engine, compiled to JavaScript via Emscripten, that converts scanned images of text bac… | 32 | 3517 | maintenance |
| TTPlanetPig/Comfyui_TTP_Toolset A collection of ComfyUI custom nodes for tiled image processing, including object-aware Smart Tile 2.0 workflows for detail img2img upscali… | 66 | 1033 | active |
| HarborYuan/ovsam Official PyTorch implementation of Open-Vocabulary SAM (ECCV 2024), a model that unifies SAM's interactive segmentation with CLIP's open-vo… | 42 | 1033 | active |
| nodeca/probe-image-size A small JavaScript library that reads image dimensions (width, height, type, mime, orientation) from URLs, streams, or buffers without down… | 76 | 1032 | active |
| podgorskiy/ALAE Official PyTorch implementation of Adversarial Latent Autoencoders (ALAE/StyleALAE), a CVPR 2020 paper combining autoencoders with GAN trai… | 32 | 3511 | maintenance |
| t3mujinpack/t3mujinpack A collection of film emulation presets for the open-source RAW photo developer Darktable, emulating classic films like Fuji Velvia, Kodak P… | 29 | 1031 | active |
| MemeCrafters/meme-generator A Python meme generator library that produces various humorous meme images from user-provided avatars and text using built-in templates. It… | 73 | 1030 | active |
| awentzonline/image-analogies A Python library implementing neural image analogies using VGG16 feature maps with PatchMatch-based matching and blending, based on the 'Im… | 23 | 3502 | maintenance |
| continue-revolution/sd-webui-segment-anything A Stable Diffusion WebUI extension that integrates Segment Anything and GroundingDINO to generate segmentation masks from clicks or text pr… | 30 | 3499 | maintenance |
| kuprel/min-dalle min(DALL·E) is a fast, minimal PyTorch port of DALL·E Mini/Mega stripped down for text-to-image inference, with only numpy, requests, pillo… | 31 | 3494 | maintenance |
| lovasoa/dezoomify-rs dezoomify-rs is a desktop application and CLI tool written in Rust that downloads high-resolution zoomable (tiled) images from websites and… | 100 | 1026 | active |
| zhyever/PatchFusion PatchFusion is a CVPR 2024 end-to-end tile-based framework for high-resolution monocular metric depth estimation from single images. It fus… | 57 | 1026 | active |
| JackAILab/ConsistentID ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using … | 52 | 1026 | active |
| chrisgoringe/cg-use-everywhere A ComfyUI custom node plugin that provides 'Anything Everywhere' nodes which broadcast data (like MODEL, CLIP, VAE) to all nodes that need … | 69 | 1024 | active |
| DingXiaoH/RepVGG RepVGG is a PyTorch implementation of the VGG-style ConvNet architecture from the CVPR 2021 paper, achieving over 84% top-1 ImageNet accura… | 32 | 3478 | maintenance |
| steffest/DPaint-js DPaint.js is a browser-based image editor modeled after Deluxe Paint, with strong support for retro Amiga file formats like IFF ILBM images… | 77 | 1021 | active |
| TencentARC/SEED-Voken SEED-Voken is a collection of visual tokenizers (Open-MAGVIT2 and IBQ) that convert images and videos into discrete tokens for autoregressi… | 48 | 1021 | active |
| yangxy/PASD PASD (Pixel-Aware Stable Diffusion) is a Python research codebase implementing an ECCV 2024 method for realistic image super-resolution and… | 28 | 1021 | active |
| JiahuiYu/generative_inpainting An open-source implementation of DeepFill v1/v2 generative image inpainting models, featuring Contextual Attention (CVPR 2018) and Gated Co… | 32 | 3466 | maintenance |
| tensorlayer/SRGAN Reference implementation of SRGAN, a generative adversarial network for photo-realistic single image super-resolution, built on TensorLayer… | 23 | 3466 | maintenance |
| ocropus-archive/DUP-ocropy OCRopy is a collection of Python-based tools for document analysis and OCR, covering binarization, page layout analysis, and text line reco… | 10 | 3465 | maintenance |
| richzhang/colorization A Python library implementing automatic colorization of grayscale photos using deep neural networks from the ECCV 2016 'Colorful Image Colo… | 32 | 3461 | maintenance |
| zai-org/GLM-Image GLM-Image is an open-source image generation model combining a 9B autoregressive generator with a 7B diffusion decoder, excelling at text r… | 48 | 1019 | active |
| google/prompt-to-prompt Google's official implementation of the Prompt-to-Prompt paper, which enables text-driven image editing in Latent Diffusion and Stable Diff… | 10 | 3456 | maintenance |
| evanw/glfx.js glfx.js is a JavaScript library for applying real-time image effects and photo adjustments in the browser using WebGL. It leverages the GPU… | 32 | 3451 | maintenance |
| Sxela/WarpFusion WarpFusion is a Stable Diffusion-based video-to-video style transfer tool distributed as a Jupyter/Colab notebook. It applies AI animation … | 31 | 1016 | active |
| eliukblau/pixterm PIXterm is a Go CLI tool that renders images directly in ANSI terminals using true color escape codes and unicode half-block characters. It… | 85 | 1015 | active |
| Alpha-VLLM/Lumina-DiMOO Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outpu… | 55 | 1015 | active |
| alessandrofrancesconi/gimp-plugin-bimp BIMP is a GIMP plugin that applies a set of image manipulations (resize, crop, rotate, watermark, color correction, format conversion, etc.… | 23 | 1015 | active |
| fallenshock/FlowEdit Official PyTorch implementation of FlowEdit, an ICCV 2025 method for inversion-free, text-based editing of real images using pre-trained fl… | 66 | 1014 | active |
| Gregwar/Image A PHP library providing a simple object-oriented API for image handling, resizing, cropping, and applying filters, with built-in caching of… | 32 | 1014 | stable |
| addyosmani/bg-remove A React + Vite web application that removes image backgrounds entirely in the browser using Transformers.js with the RMBG-1.4 model (and op… | 22 | 1014 | active |
| hezarai/hezar Hezar is an all-in-one Python AI library for the Persian language, covering NLP, speech recognition, OCR, and image captioning through a ta… | 78 | 1013 | active |
| fiji/fiji Fiji is a batteries-included distribution of ImageJ, bundling thousands of plugins for scientific image processing and analysis into a port… | 74 | 1013 | active |
| Alpha-VLLM/Lumina-Image-2.0 Lumina-Image 2.0 is an open-source 2.6B-parameter text-to-image generation framework built on a unified Next-DiT architecture with a unifie… | 58 | 1013 | active |
| eragonruan/text-detection-ctpn A TensorFlow implementation of the Connectionist Text Proposal Network (CTPN) for detecting horizontal scene text in images. It includes pr… | 23 | 3429 | maintenance |
| chrissimpkins/Crunch Crunch is a lossy PNG image optimization tool that combines bit depth, color type, and palette reduction with zopfli DEFLATE compression vi… | 23 | 3425 | maintenance |
| sail-sg/EditAnything Edit Anything is a Python application for text-guided image editing and generation, combining Segment Anything, ControlNet, BLIP2, and Stab… | 33 | 3422 | maintenance |
| RupertAvery/DiffusionToolkit Diffusion Toolkit is a Windows desktop application that indexes and views metadata (prompts, models, settings) embedded in AI-generated ima… | 65 | 1009 | active |
| waifu2x (nunif) waifu2x is an image super-resolution and noise-reduction tool for anime-style art (and photos) using deep convolutional neural networks, or… | 78 | 3418 | maintenance |
| shaoanlu/faceswap-GAN A Jupyter Notebook-based implementation of face swapping using a denoising autoencoder architecture enhanced with adversarial losses, VGGFa… | 32 | 3416 | maintenance |
| facebookresearch/Mask2Former Mask2Former is the official PyTorch implementation of the CVPR 2022 paper 'Masked-attention Mask Transformer for Universal Image Segmentati… | 10 | 3416 | maintenance |
| makegirlsmoe/makegirlsmoe_web The React front-end for MakeGirlsMoe, a web app that generates anime character portraits using a GAN model. It provides the interactive UI … | 32 | 3414 | maintenance |
| MeiGen-AI/PosterCraft PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight… | 48 | 1007 | active |
| libtv-labs/libtv-skills A collection of AI agent skill packages that expose LibLib.tv's AIGC capabilities (AI image and video generation) via its OpenAPI. It follo… | 47 | 1007 | active |
| jhfmat/ISP-pipeline-hdrplus A C/C++ image processing library (Matlib) implementing a fast ISP pipeline with HDR+ multi-frame denoising, super-low-light processing, and… | 32 | 1005 | active |
| Kosinkadink/ComfyUI-Advanced-ControlNet A set of ComfyUI custom nodes providing advanced ControlNet scheduling, weighting, and masking for Stable Diffusion workflows. It supports … | 71 | 1004 | active |
| dlbeer/quirc Quirc is a small, dependency-free C library for extracting and decoding QR codes from images, fast enough for realtime video. It handles ro… | 42 | 1004 | active |
| clovaai/CRAFT-pytorch Official PyTorch implementation of CRAFT (Character Region Awareness for Text Detection), a scene text detector that localizes text by pred… | 32 | 3398 | maintenance |
| AkiraBit/PicSharp PicSharp is a modern, cross-platform desktop application for high-performance image compression, built with Tauri and TypeScript. It suppor… | 59 | 1003 | active |
| mit-han-lab/efficientvit A collection of efficient vision foundation models from MIT Han Lab, including EfficientViT backbones for perception, EfficientViT-SAM for … | 48 | 3354 | maintenance |
| eladrich/pixel2style2pixel Official PyTorch implementation of pixel2style2pixel (pSp), a StyleGAN encoder from CVPR 2021 that maps real images directly into the W+ la… | 32 | 3350 | maintenance |
| tamarott/SinGAN Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat… | 32 | 3344 | maintenance |