domain: image-processing
1843 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| zuruoke/watermark-removal A machine learning tool that removes watermarks from images using deep learning image inpainting, based on Contextual Attention and Gated C… | 83 | 5139 | maintenance |
| Tom94/tev tev is a high dynamic range (HDR/EDR) image viewer written in C++ for people who care about accurate colors. It supports many formats (EXR,… | 99 | 1426 | active |
| mdbloice/Augmentor Augmentor is a standalone Python library for image augmentation in machine learning, providing a pipeline of stochastic operations like rot… | 32 | 5133 | maintenance |
| apple/ml-aim Apple's official repository for AIM (Autoregressive Image Models), providing code and pretrained checkpoints for AIMv1 and AIMv2 large visi… | 42 | 1424 | active |
| mix1009/sdwebuiapi A Python client library for the AUTOMATIC1111 stable-diffusion-webui REST API, wrapping endpoints like txt2img, img2img, and upscaling into… | 32 | 1423 | active |
| affinelayer/pix2pix-tensorflow A TensorFlow implementation of pix2pix, a conditional GAN that learns a mapping from input images to output images. It is a faithful port o… | 32 | 5081 | maintenance |
| CSAILVision/semantic-segmentation-pytorch A PyTorch implementation of semantic segmentation (scene parsing) models for the MIT ADE20K dataset, including pretrained model zoo and tra… | 32 | 5078 | maintenance |
| khanamiryan/php-qrcode-detector-decoder A pure PHP library for detecting and decoding QR codes from images, ported from the ZXing library. It works without any PHP extensions beyo… | 37 | 1412 | active |
| SmileyChris/easy-thumbnails A Django application for generating and managing image thumbnails, supporting predefined aliases, template tags, and Python API usage. It s… | 60 | 1410 | active |
| IDEA-Research/DINO-X-API DINO-X API is a Python client library and examples for accessing DINO-X, a hosted unified vision model for open-world object detection and … | 36 | 1410 | active |
| gnobitab/InstaFlow InstaFlow is a one-step text-to-image generation model based on Rectified Flow, enabling ultra-fast Stable Diffusion inference without iter… | 28 | 1409 | active |
| willmiao/ComfyUI-Lora-Manager A ComfyUI extension that provides a web-based manager for organizing, previewing, downloading, and applying LoRA models with metadata and r… | 83 | 1404 | active |
| nachifur/MulimgViewer MulimgViewer is a Python-based multi-image viewer that displays many images in a single interface for side-by-side comparison, parallel sel… | 66 | 1400 | active |
| imagej/imagej2 ImageJ2 is an open-source Java framework and application for processing and analyzing N-dimensional scientific image data, built on the Img… | 75 | 1399 | stable |
| yfeng95/PRNet PRNet is a Python/TensorFlow implementation of the ECCV 2018 Position Map Regression Network for joint 3D face reconstruction and dense ali… | 32 | 5013 | maintenance |
| liyue-aigc/female-portrait-director A modular Codex Skill that directs and expands detailed AI female portrait prompts into coherent, camera-ready scenes, or generates images … | 73 | 1396 | active |
| jexom/sd-webui-depth-lib A depth map library and poser plugin for the Automatic1111 stable-diffusion-webui, designed to supply depth maps to the ControlNet extensio… | 10 | 1395 | stable |
| logtd/ComfyUI-Fluxtapoz A set of ComfyUI custom nodes for editing and stylizing images with Flux models, implementing techniques like RF-Inversion, RF-Edit, Firefl… | 23 | 1392 | active |
| exif-js/exif-js Exif.js is a JavaScript library for reading EXIF and IPTC metadata from JPEG and TIFF images in the browser. It works with image elements o… | 23 | 4979 | maintenance |
| claviska/SimpleImage A single-file PHP library that wraps the GD extension with a fluent, chainable API for loading, manipulating, and saving images. It support… | 53 | 1388 | stable |
| cszn/BSRGAN BSRGAN is a PyTorch implementation of a practical degradation model for deep blind image super-resolution, presented at ICCV 2021. It provi… | 32 | 1386 | stable |
| 14790897/handwriting-web A self-hostable web application that converts typed text into images (or PDFs) simulating handwriting, using uploaded fonts, background ima… | 92 | 1385 | active |
| OpenGVLab/DragGAN An unofficial full-featured Python implementation of DragGAN, the interactive point-based image manipulation method built on StyleGAN gener… | 29 | 4946 | maintenance |
| sdsykes/fastimage FastImage is a Ruby library that finds the size or type of an image given its URI by fetching only the minimal bytes needed. It supports ma… | 75 | 1380 | stable |
| zhixuhao/unet A Keras implementation of the U-Net convolutional network architecture for image segmentation, based on the original biomedical segmentatio… | 66 | 4941 | maintenance |
| keyu-tian/SparK SparK is the official PyTorch implementation of an ICLR 2023 Spotlight paper that applies BERT/MAE-style masked image modeling to convoluti… | 22 | 1376 | stable |
| qubvel/segmentation_models A Python library providing neural network architectures for image segmentation (Unet, FPN, Linknet, PSPNet) built on Keras and TensorFlow K… | 23 | 4923 | maintenance |
| ali-vilab/TeaCache TeaCache is a training-free caching approach that accelerates inference for video diffusion models by estimating output differences across … | 33 | 1369 | active |
| spatie/image A PHP library for manipulating images with an expressive, fluent API, supporting drivers like Imagick and GD. It provides operations such a… | 94 | 1364 | stable |
| ImprintLab/MedSegDiff MedSegDiff is a diffusion probabilistic model framework for segmenting and reconstructing organs and tissues from medical images, with a tr… | 50 | 1363 | active |
| bytedance/UNO UNO is a research framework from ByteDance for subject-driven image generation with diffusion transformers, supporting both single- and mul… | 38 | 1362 | active |
| scraed/LanPaint LanPaint is a training-free inpainting sampler for stable diffusion models, implemented as a ComfyUI custom node. It uses iterative 'think … | 85 | 1361 | active |
| mzucker/noteshrink A Python command-line script that cleans up scans and photos of handwritten notes by separating background from ink, quantizing colors, and… | 32 | 4841 | maintenance |
| lxtGH/OMG-Seg Official research codebase for OMG-Seg (CVPR 2024) and OMG-LLaVA (NeurIPS 2024), unified models for image-level, object-level, and pixel-le… | 47 | 1354 | active |
| zmrlft/GreenWall A Wails-based desktop application that customizes your GitHub contribution graph by converting uploaded images into contribution heatmap pa… | 68 | 1352 | active |
| wyhuai/DDNM DDNM is a Python research codebase implementing the Denoising Diffusion Null-Space Model for zero-shot image restoration, published as an I… | 32 | 1349 | stable |
| jhc13/taggui TagGUI is a cross-platform desktop application for quickly adding and editing image tags and captions, aimed at creators of image datasets … | 53 | 1346 | active |
| webtoon/psd A fast, zero-dependency TypeScript parser for Adobe Photoshop PSD/PSB files that runs in web browsers and Node.js. It extracts layer inform… | 23 | 1345 | active |
| jbilcke-hf/ai-comic-factory AI Comic Factory is a web application that generates comic panels and pages from a single text prompt by combining an LLM for story/dialogu… | 10 | 1345 | active |
| imgbot/Imgbot Imgbot is a GitHub App backed by an Azure Functions solution that crawls a repository's image files and losslessly compresses them, then op… | 24 | 1344 | active |
| muzishen/IMAGDressing IMAGDressing-v1 is a diffusion-based framework for customizable virtual dressing that generates human images with fixed garments and contro… | 44 | 1343 | active |
| zanllp/infinite-image-browsing Infinite Image Browsing (IIB) is a full-featured image and video management application with fast thumbnail-based browsing, AI-generation m… | 93 | 1341 | active |
| FireRedTeam/FireRed-Image-Edit FireRed-Image-Edit is an open-source image editing foundation model built on diffusion models, released as PyTorch model weights with infer… | 49 | 1341 | active |
| receyuki/stable-diffusion-prompt-reader A standalone desktop application (with GUI and CLI) for reading generation prompts and parameters embedded in Stable Diffusion images, with… | 21 | 1339 | stable |
| senguptaumd/Background-Matting Official research code for 'Background Matting: The World is Your Green Screen' (CVPR 2020), a deep network that extracts per-pixel alpha m… | 32 | 4769 | maintenance |
| christophschuhmann/improved-aesthetic-predictor A CLIP+MLP neural network that predicts how much people on average like an image, trained on AVA dataset ratings. It is widely used for fil… | 32 | 1338 | stable |
| ByteDance-Seed/SeedVR SeedVR/SeedVR2 are diffusion-transformer based models for generic real-world and AIGC video and image restoration, with SeedVR2 using adver… | 47 | 1334 | active |
| jtydhr88/ComfyUI-qwenmultiangle A ComfyUI custom node providing an interactive Three.js 3D viewport for controlling camera azimuth, elevation, and zoom. It outputs formatt… | 56 | 1332 | active |
| ldqk/ImageSearch A .NET 10 desktop demo application that performs reverse image search (search by image) over local hard drives with tens of millions of ima… | 93 | 1329 | active |
| Gourieff/ComfyUI-ReActor ComfyUI-ReActor is a fast and simple face swap extension node for ComfyUI, based on the ReActor face-swapping engine. It includes a nudity … | 65 | 1327 | active |
| numandev1/react-native-compressor A React Native library that compresses images, videos, and audio with WhatsApp-like quality, plus background upload, file download, and vid… | 96 | 1325 | active |
| cvzone/cvzone CVZone is a Python computer vision helper library that wraps OpenCV and MediaPipe to simplify image processing and AI functions like hand t… | 32 | 1325 | active |
| nateraw/stable-diffusion-videos A Python library for generating videos with Stable Diffusion by walking the latent space and morphing between text prompts. It supports bea… | 56 | 4705 | maintenance |
| thephpleague/color-extractor A PHP library that extracts the most representative colors from an image, building a color palette sorted by pixel count. It handles transp… | 88 | 1323 | stable |
| meta-pytorch/segment-anything-fast A fast, batched offline inference-oriented fork of Meta's Segment Anything (SAM) image segmentation model. It applies optimizations like bf… | 45 | 1321 | active |
| Decimation/SmartImage SmartImage is a reverse image search tool that queries multiple engines (SauceNao, IQDB, Ascii2D, trace.moe, TinEye, Yandex, and more) and … | 97 | 1320 | active |
| nroduit/Weasis Weasis is an open-source DICOM viewer for medical imaging that runs standalone or embedded in web applications. It integrates with PACS, VN… | 95 | 1317 | active |
| Capsize-Games/airunner AI Runner is a privacy-focused desktop application for running local AI models offline, combining an AI chat companion with voice conversat… | 79 | 1314 | active |
| ali-vilab/MimicBrush MimicBrush is the official implementation of a zero-shot image editing method that lets users mask a region in a source image and provide a… | 24 | 1311 | active |
| flozz/StackBlur StackBlur.js is a JavaScript library implementing the fast, almost-Gaussian StackBlur algorithm for images and canvas elements. It works in… | 73 | 1310 | stable |
| spatie/laravel-image-optimizer A Laravel package that optimizes PNG, JPG, SVG, and GIF images by running them through a chain of installed optimization binaries. It is th… | 77 | 1309 | stable |
| derrian-distro/LoRA_Easy_Training_Scripts A PySide6 desktop GUI that wraps Kohya's sd-scripts to simplify training LoRA, LoCon, and other LoRA-type models for Stable Diffusion. It s… | 33 | 1307 | active |
| seetaface/SeetaFaceEngine SeetaFace Engine is an open-source C++ face recognition engine comprising face detection, face alignment, and face identification modules. … | 32 | 4636 | maintenance |
| Vincentqyw/image-matching-webui A Gradio-based web UI that matches keypoints between two images using many state-of-the-art image matching algorithms (LoFTR, SuperGlue, Li… | 91 | 1302 | active |
| luosiallen/latent-consistency-model Official implementation of Latent Consistency Models (LCM), a diffusion-based approach for synthesizing high-resolution images with few-ste… | 27 | 4615 | maintenance |
| rsmbl/Resemble.js Resemble.js is a JavaScript library for analyzing and comparing images using HTML5 canvas, producing pixel-level diff images and analysis d… | 32 | 4612 | maintenance |
| sighook/pixload pixload is a set of Perl CLI tools for creating and injecting payloads into image files (BMP, GIF, JPG, PNG, WebP). It is used in offensive… | 23 | 1300 | active |
| Uminosachi/sd-webui-inpaint-anything A Stable Diffusion Web UI extension that performs inpainting and outpainting using masks generated by Segment Anything models (SAM 2, SAM-H… | 30 | 1298 | active |
| PyImageSearch/imutils A Python library of convenience functions that simplify common OpenCV image processing tasks such as translation, rotation, resizing, skele… | 32 | 4590 | maintenance |
| ARM-software/astc-encoder The Arm ASTC Encoder (astcenc) is a command-line tool and codec library for compressing and decompressing images in the Adaptive Scalable T… | 92 | 1295 | stable |
| W2GenAI-Lab/LucidFlux LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a… | 55 | 1293 | active |
| Parskatt/RoMa RoMa (romatch) is a Python library for robust dense feature matching between image pairs, estimating pixel-dense warps and reliable certain… | 52 | 1293 | active |
| ClownsharkBatwing/RES4LYF RES4LYF is a ComfyUI custom node collection providing advanced diffusion samplers (RES samplers) that achieve high-quality image generation… | 67 | 1289 | active |
| t3dotgg/quickpic QuickPic is a free, open-source web app by Theo that converts SVGs to high-resolution PNGs and offers other quick image utilities like squa… | 22 | 1289 | active |
| azagaya/laigter Laigter is an open-source desktop tool that automatically generates normal, specular, parallax, and occlusion maps from 2D textures, design… | 90 | 1288 | active |
| RickdeJager/stegseek Stegseek is a lightning-fast command-line cracker for steghide steganography, built as a fork of the original steghide project that can tes… | 23 | 1288 | stable |
| bytedance/Bernini Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer perf… | 57 | 1287 | active |
| lzhgus/Capso Capso is a free, open-source native macOS app for screenshots and screen recording, built with Swift 6.0 and SwiftUI as an alternative to C… | 81 | 1286 | active |
| huanngzh/MV-Adapter MV-Adapter is a plug-and-play adapter that turns pre-trained text-to-image diffusion models (e.g., SDXL, SD2.1) into multi-view consistent … | 34 | 1285 | active |
| zju3dv/MatchAnything MatchAnything is a deep learning model for universal cross-modality image matching, released as research code accompanying a TPAMI 2026 pap… | 64 | 1279 | active |
| plemeri/transparent-background A Python tool and CLI that removes backgrounds from images and videos using the InSPyReNet deep learning model (ACCV 2022). It supports ima… | 63 | 1278 | active |
| Tencent-Hunyuan/SRPO SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion… | 54 | 1278 | active |
| NVlabs/stylegan2-ada-pytorch Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi… | 32 | 4487 | maintenance |
| manycore-maas/Painter Painter is a JSON-driven canvas drawing library for WeChat mini programs (also usable in Node and HTML5) that renders images from a declara… | 23 | 4475 | maintenance |
| vipshop/cache-dit Cache-DiT is a PyTorch-native inference engine that accelerates Diffusion Transformer (DiT) models with hybrid caching, parallelism, quanti… | 82 | 1267 | active |
| DimitarPetrov/stegify stegify is a Go command line tool and library for LSB (Least Significant Bit) steganography that hides any file inside images such as PNG a… | 23 | 1267 | stable |
| blueimp/JavaScript-Load-Image A JavaScript library that loads images from File/Blob objects or URLs and returns optionally scaled, cropped, or rotated HTML img or canvas… | 32 | 4456 | maintenance |
| brendan-duncan/image A pure-Dart library for decoding, encoding, and manipulating images in many formats (PNG, JPEG, GIF, WebP, TIFF, BMP, and more). It works w… | 76 | 1263 | active |
| bryandlee/animegan2-pytorch A PyTorch implementation of AnimeGANv2, a GAN-based image-to-image style transfer model that converts photos into anime-style images. It pr… | 32 | 4452 | maintenance |
| GraphiteEditor/Graphite Graphite is a free, open source 2D graphics editor built in Rust that combines layer-based compositing with a node-based procedural graphic… | 67 | 26939 | experimental |
| ashuoAI/SHUO-Canvas SHUO Canvas (formerly AI-CanvasPro) is an AI multimodal creation canvas application that lets users combine text, images, video, and audio … | 82 | 1261 | active |
| rlawjdghek/StableVITON StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent… | 47 | 1261 | stable |
| MemeMeow-Studio/MemeMeow MemeMeow is a self-hosted meme/sticker management and retrieval application that lets users find images by describing the desired scene in … | 65 | 1260 | active |
| ShiftHackZ/Stable-Diffusion-KMP SDAI is an open-source, cross-platform Stable Diffusion client app for Android and iOS built with Kotlin Multiplatform and Jetpack Compose.… | 91 | 1259 | active |
| Francis-Rings/StableAvatar StableAvatar is an end-to-end video diffusion transformer that generates infinite-length, high-quality talking avatar videos from a referen… | 46 | 1258 | active |
| VainF/pytorch-msssim A PyTorch library providing fast, differentiable SSIM and MS-SSIM image quality metrics using separable Gaussian filtering for speed. It ca… | 23 | 1253 | stable |
| meowtec/Imagine Imagine is a cross-platform desktop GUI app for optimizing PNG, JPEG, and WebP images, built on pngquant, mozjpeg, and WebP encoders with a… | 23 | 4405 | maintenance |
| antfu/qrcode-toolkit A web-based toolkit for generating base QR codes and refining AI-generated QR codes by comparing outputs to find misaligned pixels. It also… | 29 | 1250 | active |
| sentinel-hub/eo-learn eo-learn is a collection of open-source Python packages for accessing and processing spatio-temporal satellite imagery, built around modula… | 51 | 1247 | active |
| MikeKovarik/exifr exifr is a fast, dependency-free JavaScript library for reading EXIF and other image metadata (TIFF, XMP, ICC, IPTC, GPS, JFIF) from JPG, P… | 23 | 1247 | active |