domain: image-processing
1843 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| kean/Nuke Nuke is a Swift image loading and caching framework for Apple platforms, providing an ImagePipeline for fetching, processing, and displayin… | 99 | 8656 | stable |
| lllyasviel/IC-Light IC-Light is a Python tool for manipulating the illumination of images using diffusion models, offering text-conditioned and background-cond… | 27 | 8508 | active |
| Nutlope/logocreator An open-source AI logo generator web app that creates brand-ready logos using FLUX models on Together AI, with logo editing, style presets,… | 64 | 8435 | active |
| photopea/photopea Photopea is a free online image editor for raster and vector graphics that runs entirely in the browser, supporting PSD, AI, Sketch, and do… | 45 | 8405 | active |
| CASIA-LMC-Lab/FastSAM FastSAM is a CNN-based Segment Anything Model trained on only 2% of the SA-1B dataset, achieving comparable segmentation performance to SAM… | 19 | 8401 | active |
| XPixelGroup/BasicSR BasicSR is an open-source PyTorch toolbox for image and video restoration tasks such as super-resolution, denoising, deblurring, and JPEG a… | 23 | 8367 | stable |
| bytedeco/javacv JavaCV is a Java library that wraps OpenCV, FFmpeg, and other computer vision and multimedia libraries via JavaCPP Presets, with utility cl… | 86 | 8335 | active |
| vietnh1009/ASCII-generator A Python tool that converts images and videos into ASCII art, outputting text files or image/video files in grayscale or color. It supports… | 32 | 8318 | stable |
| exadel-inc/CompreFace Exadel CompreFace is a free, open-source face recognition system that provides REST APIs for face recognition, verification, detection, lan… | 23 | 8273 | stable |
| QwenLM/Qwen-Image Qwen-Image is a 20B MMDiT image generation foundation model from the Qwen team, with strong complex text rendering (especially Chinese) and… | 48 | 8265 | active |
| lltcggie/waifu2x-caffe A Windows GUI/CLI application that reimplements the waifu2x image upscaling and noise-reduction tool using the Caffe deep learning framewor… | 56 | 8221 | stable |
| carson-katri/dream-textures A Blender add-on that integrates Stable Diffusion for generating textures, concept art, and background assets directly inside Blender. It s… | 23 | 8195 | active |
| brycedrennan/imaginAIry A Python library and CLI tool (imaginairy/aimg) for generating images and videos with Stable Diffusion and Stable Video Diffusion models. I… | 63 | 8179 | active |
| zumerlab/snapdom SnapDOM is a high-performance, dependency-free browser library that captures DOM elements as self-contained SVG representations and exports… | 86 | 8040 | active |
| SixLabors/ImageSharp ImageSharp is a fully managed, high-performance 2D graphics and image processing library for .NET 8+. It provides loading, resizing, format… | 95 | 8034 | stable |
| nadermx/backgroundremover A free, open-source command line tool that removes backgrounds from images and videos using the U2Net AI model built on PyTorch. It is inst… | 93 | 8020 | active |
| bingoogolapple/BGAQRCode-Android An Android library for scanning and generating QR codes and barcodes, offering both ZXing and ZBar engines behind customizable scan views. … | 64 | 8008 | stable |
| PixiEditor/PixiEditor PixiEditor is a free, open-source, cross-platform desktop 2D graphics editor built in C# with Avalonia. It combines pixel art, painting, an… | 99 | 7999 | active |
| MochiDiffusion/MochiDiffusion A native macOS app built with SwiftUI for running Stable Diffusion and FLUX.2 Klein image generation locally on Apple Silicon Macs. It uses… | 78 | 7926 | active |
| TheLastBen/fast-stable-diffusion A collection of Google Colab notebooks for quickly running Stable Diffusion UIs (AUTOMATIC1111, ComfyUI) and training DreamBooth models for… | 57 | 7910 | active |
| TencentARC/GFPGAN GFPGAN is a Python library built on PyTorch that restores and enhances real-world degraded face photos using GAN-based priors. It provides … | 23 | 37657 | maintenance |
| geekyutao/Inpaint-Anything Inpaint Anything combines Segment Anything (SAM) with inpainting models like LaMa and Stable Diffusion to remove, fill, or replace objects … | 65 | 7703 | active |
| lllyasviel/Omost Omost is a Python application that converts LLM coding capability into image composition by having pretrained LLMs (based on Llama3 and Phi… | 24 | 7610 | active |
| RapidAI/RapidOCR RapidOCR is an open-source, multi-language OCR toolkit that performs text detection and recognition using models converted to run on ONNX R… | 97 | 7599 | active |
| 1adrianb/face-alignment A Python library built on PyTorch that detects 2D and 3D facial landmarks in images using the FAN deep learning face alignment network. It … | 70 | 7538 | active |
| phoboslab/qoi QOI is the 'Quite OK Image Format', a fast, lossless image compression format with a single-file MIT-licensed C/C++ reference implementatio… | 70 | 7522 | stable |
| hybridgroup/gocv GoCV is a Go language binding for the OpenCV 4 computer vision library, supporting Linux, macOS, Windows, and Docker. It includes support f… | 76 | 7491 | active |
| EutropicAI/Final2x Final2x is a cross-platform desktop application for image super-resolution (upscaling) built with Electron, Vue3, and a PyTorch-based Pytho… | 74 | 7323 | active |
| vladmandic/sdnext SD.Next is an open-source, self-hosted WebUI server application for AI generative image and video creation built on Stable Diffusion and Di… | 76 | 7320 | active |
| AbdullahAlfaraj/Auto-Photoshop-StableDiffusion-Plugin A Photoshop plugin (built on Adobe UXP) that lets users generate Stable Diffusion images directly inside Photoshop, using Automatic1111 Web… | 22 | 7288 | active |
| imgly/background-removal-js An npm package (browser and Node.js variants) that removes image backgrounds using ONNX-based image segmentation/matting models running ent… | 43 | 7287 | active |
| liuliu/ccv ccv is a modern, minimalist computer vision library written in C/C++ with an application-driven set of state-of-the-art algorithms includin… | 77 | 7243 | active |
| civitai/civitai Civitai is a web platform for sharing, discovering, and discussing Stable Diffusion models, textual inversions, LoRAs, VAEs, and other gene… | 77 | 7238 | active |
| zetbaitsu/Compressor Compressor is a lightweight Android image compression library written in Kotlin. It lets developers shrink large photos into smaller files … | 54 | 7227 | stable |
| bubkoo/html-to-image A TypeScript library that generates images (PNG, JPEG, SVG, blob, canvas, or pixel data) from DOM nodes using HTML5 canvas and SVG foreignO… | 72 | 7222 | stable |
| kohya-ss/sd-scripts A collection of Python training, generation, and utility scripts for Stable Diffusion and other image generation models, most widely used f… | 89 | 7210 | active |
| PeterL1n/BackgroundMattingV2 Official PyTorch implementation of the CVPR 2021 paper 'Real-Time High-Resolution Background Matting'. It produces state-of-the-art alpha m… | 23 | 7189 | stable |
| ControlNet ControlNet is a neural network architecture that adds conditional control (edges, poses, depth, etc.) to pretrained text-to-image diffusion… | 31 | 34091 | maintenance |
| zxing/zxing ZXing ('Zebra Crossing') is an open-source, multi-format 1D/2D barcode image processing library implemented in Java, with ports to other la… | 73 | 34077 | maintenance |
| xushengfeng/eSearch eSearch is a cross-platform desktop application (Electron) combining screenshot capture, offline OCR based on PaddleOCR, screen search, tra… | 98 | 7036 | active |
| latentcat/qrbtf QRBTF is an AI and parametric QR code generator that creates stylized, scannable QR codes, available as a hosted web app at qrbtf.com with … | 30 | 6996 | active |
| mapbox/pixelmatch A tiny, dependency-free JavaScript library for pixel-level image comparison, originally built for comparing screenshots in tests. It works … | 89 | 6931 | stable |
| sczhou/ProPainter ProPainter is a PyTorch-based video inpainting model from ICCV 2023 that combines dual-domain propagation with a mask-guided sparse video T… | 22 | 6916 | stable |
| leejet/stable-diffusion.cpp A pure C/C++ inference engine for diffusion models (Stable Diffusion, FLUX, Wan, Qwen Image, Z-Image, and more) built on ggml in the style … | 91 | 6846 | active |
| Dooy/chatgpt-web-midjourney-proxy A unified web/desktop UI for ChatGPT plus AI image, music, and video generation services like Midjourney, Suno, Luma, Runway, and Flux. It … | 76 | 6785 | active |
| visioncortex/vtracer VTracer is an open-source tool and library that converts raster images (JPG, PNG) into vector graphics (SVG), handling both colored images … | 98 | 6745 | active |
| tencent-ailab/IP-Adapter IP-Adapter is a lightweight (22M parameter) adapter that adds image prompt capability to pretrained text-to-image diffusion models like Sta… | 28 | 6677 | stable |
| halide/Halide Halide is an embedded DSL (in C++ and Python) for writing high-performance, data-parallel image and array processing pipelines. It separate… | 70 | 6590 | stable |
| 11cafe/jaaz Jaaz is an open-source multimodal creative assistant application that serves as a privacy-focused, locally usable alternative to Canva and … | 53 | 6589 | active |
| scikit-image/scikit-image scikit-image is a Python library providing a collection of peer-reviewed image processing algorithms built on NumPy and SciPy. It offers ro… | 82 | 6577 | stable |
| iperov/DeepFaceLive DeepFaceLive is a real-time face-swap application for PC streaming and video calls, using trained face models (DFM) applied to webcam or vi… | 10 | 31011 | maintenance |
| HVision-NKU/StoryDiffusion StoryDiffusion is the official implementation of a NeurIPS 2024 Spotlight paper introducing Consistent Self-Attention for character-consist… | 24 | 6452 | active |
| sz3/libcimbar libcimbar is an optimized C++ implementation of the cimbar (color icon matrix) high-density 2D barcode format for air-gapped data transfer … | 95 | 6426 | active |
| OpenDroneMap/ODM OpenDroneMap (ODM) is an open source command line toolkit that processes aerial drone, balloon, or kite imagery into classified point cloud… | 90 | 6417 | active |
| madmaze/pytesseract Python-tesseract is a Python wrapper for Google's Tesseract-OCR engine that recognizes and extracts text embedded in images. It supports al… | 64 | 6383 | stable |
| GNOME/gimp GIMP (GNU Image Manipulation Program) is a free, open-source raster graphics editor for photo retouching, image composition, and authoring.… | 77 | 6378 | active |
| mindee/doctr docTR is a Python OCR library that extracts text from documents and images using a two-stage deep learning approach: text detection followe… | 90 | 6315 | active |
| Lymphatus/caesium-image-compressor Caesium Image Compressor is a free, open-source desktop application for compressing JPG, PNG, WebP, and TIFF images while preserving visual… | 68 | 6266 | active |
| szad670401/HyperLPR HyperLPR3 is a high-performance open-source framework for recognizing Chinese license plates, built with deep learning and available as a P… | 27 | 6255 | active |
| RangiLyu/nanodet NanoDet-Plus is a super fast, lightweight anchor-free object detection model implemented in PyTorch, with model sizes as small as 980KB (IN… | 23 | 6252 | stable |
| Akegarasu/lora-scripts SD-Trainer is a GUI application and set of scripts for training LoRA and Dreambooth fine-tunes of Stable Diffusion diffusion models, wrappi… | 66 | 6110 | active |
| h2non/imaginary Imaginary is a fast HTTP microservice written in Go for high-level image processing, backed by bimg and libvips. It exposes image operation… | 46 | 6079 | stable |
| shimat/opencvsharp OpenCvSharp is a cross-platform .NET wrapper for the OpenCV computer vision library, published as NuGet packages with bundled native binari… | 98 | 6072 | active |
| basketikun/chatgpt2api A self-hosted reverse-engineered implementation of ChatGPT's official web interfaces, exposing OpenAI-compatible API endpoints for text gen… | 77 | 6012 | active |
| Doubiiu/ToonCrafter ToonCrafter is a generative model that interpolates two cartoon images into a short animation by leveraging pre-trained image-to-video diff… | 29 | 6003 | stable |
| chaiNNer-org/chaiNNer chaiNNer is a free, open-source, node-based desktop application for building image processing pipelines by connecting nodes on a canvas. Or… | 81 | 5994 | active |
| lxfater/inpaint-web A free, open-source browser-based tool for image inpainting (object removal) and image upscaling (super-resolution), built with WebGPU and … | 53 | 5912 | active |
| image-rs/image A Rust library providing encoding and decoding for many common image formats (PNG, JPEG, GIF, WebP, AVIF, TIFF, and more) plus basic image … | 77 | 5862 | stable |
| ChaoningZhang/MobileSAM MobileSAM is the official implementation of a lightweight version of Meta's Segment Anything Model (SAM), replacing the heavyweight image e… | 65 | 5858 | stable |
| PaddlePaddle/PaddleClas PaddleClas is a Python library and toolkit for image classification, recognition, and retrieval built on the PaddlePaddle deep learning fra… | 66 | 5838 | active |
| fengyuanchen/compressorjs Compressor.js is a JavaScript library that compresses images in the browser using the native HTMLCanvasElement.toBlob() method. It is typic… | 81 | 5764 | stable |
| disintegration/imaging Imaging is a Go library providing basic image processing functions such as resize, rotate, crop, and brightness/contrast adjustments. It wo… | 23 | 5756 | stable |
| kornelski/pngquant pngquant is a command-line lossy PNG compressor that converts images to efficient 8-bit palette PNGs with alpha support, often reducing fil… | 72 | 5742 | stable |
| imagemin/imagemin A Node.js library that minifies images (JPEG, PNG, GIF, SVG) through a pluggable architecture, accepting file globs or buffers and returnin… | 27 | 5722 | active |
| mozilla/mozjpeg MozJPEG is an improved JPEG encoder built as a patch on libjpeg-turbo, producing smaller files at higher visual quality while remaining ful… | 35 | 5716 | stable |
| zhongerxin/Cowart Cowart is a Codex plugin that provides a native infinite canvas widget built on tldraw for brainstorming, annotating images, and AI-driven … | 58 | 5683 | active |
| OpenSenseNova/SenseNova-U1 SenseNova-U is a series of open-weight unified multimodal models (e.g., SenseNova-U1.5-8B-MoT) built on the NEO-unify architecture that com… | 59 | 5668 | active |
| idealo/imagededup imagededup is a Python library for finding exact and near-duplicate images in a collection using perceptual hashing algorithms (PHash, DHas… | 48 | 5666 | stable |
| Fanghua-Yu/SUPIR SUPIR is a Python-based photo-realistic image restoration system built on SDXL diffusion priors and LLaVA captioning, presented at CVPR 202… | 36 | 5649 | active |
| basketikun/infinite-canvas An open-source infinite canvas workbench for AI-driven visual creation, combining canvas orchestration, AI image generation, reference-imag… | 80 | 5629 | active |
| ATH-MaaS/ComfyUI-Copilot ComfyUI-Copilot is an AI-powered custom node for ComfyUI that acts as an intelligent assistant for building, debugging, and optimizing imag… | 53 | 5489 | active |
| coobird/thumbnailator Thumbnailator is a Java library for generating high-quality image thumbnails with a simple fluent API. It wraps the Java Image I/O and Java… | 63 | 5425 | active |
| mayocream/koharu Koharu is a local-first desktop application that automates manga translation using machine learning, combining text/bubble detection, OCR, … | 82 | 5410 | active |
| GargantuaX/gemini-watermark-remover A 100% client-side tool that removes Gemini AI watermarks from images and videos using a mathematically precise Reverse Alpha Blending algo… | 82 | 5409 | active |
| Hillobar/Rope Rope is a GUI-focused desktop application for face swapping that implements the insightface inswapper_128 model. It offers batch swapping, … | 79 | 5367 | active |
| novicezk/midjourney-proxy A Java-based proxy service that wraps MidJourney's Discord channel into a REST API, enabling programmatic AI image generation. It supports … | 45 | 5347 | active |
| wiltodelta/remove-ai-watermarks A Python library and CLI for removing AI watermarks and provenance metadata from images and video the user generated themselves. It handles… | 77 | 5278 | active |
| timesler/facenet-pytorch A PyTorch library providing pretrained face detection (MTCNN) and facial recognition (Inception ResNet V1) models, ported from the TensorFl… | 42 | 5162 | stable |
| yisol/IDM-VTON Official implementation of IDM-VTON, an ECCV 2024 paper that improves diffusion models for high-fidelity virtual try-on, swapping garments … | 30 | 5156 | active |
| aigc-apps/sd-webui-EasyPhoto EasyPhoto is a Stable Diffusion WebUI plugin for generating AI portraits by training a personal 'digital doppelganger' from 5-20 user photo… | 28 | 5155 | active |
| philz1337x/clarity-upscaler Clarity AI is a free and open-source AI image upscaler and enhancer built on Stable Diffusion, positioned as an alternative to Magnific. It… | 30 | 5115 | active |
| ai-dawang/PlugNPlay-Modules A curated collection of plug-and-play deep learning modules (convolutions, attention mechanisms, downsampling, and feature fusion blocks) i… | 38 | 5105 | active |
| dmMaze/BallonsTranslator A desktop GUI application that uses deep learning to automatically translate comics and manga, combining text detection, OCR, inpainting, a… | 99 | 5065 | active |
| RandyGaul/cute_headers A collection of cross-platform, single-file C/C++ header-only libraries with no dependencies, primarily aimed at game development. It inclu… | 75 | 5052 | active |
| soruly/trace.moe trace.moe is an anime scene search engine that identifies which anime, episode, and exact timestamp a screenshot comes from. This repositor… | 76 | 5022 | active |
| pollinations/pollinations Pollinations is an open-source generative AI platform offering free APIs for text, image, and vision model inference without requiring API … | 77 | 4996 | active |
| pkuliyi2015/multidiffusion-upscaler-for-automatic1111 An Automatic1111 Stable Diffusion WebUI extension that enables generating and upscaling ultra-large images (2K+) on GPUs with limited VRAM … | 30 | 4994 | active |
| TheJoeFin/Text-Grab Text Grab is a Windows OCR utility that extracts text from anywhere on screen — screenshots, images, videos, PDFs, or app windows — entirel… | 94 | 4978 | active |
| TimOliver/TOCropViewController TOCropViewController is an open-source iOS view controller for cropping UIImage objects and performing basic rotations, designed to feel li… | 92 | 4942 | stable |
| wuyoscar/GPT-Image2-Skill A curated prompt gallery and library for OpenAI's GPT Image 2 model, packaged as an agentic skill and Python CLI for image generation and e… | 74 | 4923 | active |