domain: image-processing
1843 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| ggchivalrous/yiyin Yiyin (壹印) is a free, open-source desktop application for adding watermark frames to photos, generating styled image borders with customiza… | 83 | 1703 | active |
| zai-org/ImageReward ImageReward is the first general-purpose human preference reward model for text-to-image generation, trained on 137k expert comparison pair… | 52 | 1702 | stable |
| sihyun-yu/REPA REPA is the official PyTorch implementation of the ICLR 2025 paper 'Representation Alignment for Generation', a regularization technique th… | 27 | 1700 | active |
| GreycLab/CImg CImg is a small, open-source, header-only C++ template library for image processing. It provides a single image class supporting up to 4-di… | 76 | 1690 | stable |
| sedthh/pyxelate Pyxelate is a Python library and CLI tool that converts images into 8-bit pixel art by downsampling and learning a reduced color palette. I… | 51 | 1685 | active |
| williamyang1991/DualStyleGAN Official PyTorch implementation of DualStyleGAN, a CVPR 2022 model for exemplar-based high-resolution (1024px) portrait style transfer. It … | 32 | 1683 | stable |
| fire-keeper/BlindWatermark A Python library and CLI/GUI tool that embeds invisible blind watermarks into images using discrete wavelet transforms, protecting creators… | 23 | 1675 | active |
| franciszzj/Leffa Leffa is a diffusion-based framework for controllable person image generation, supporting virtual try-on and pose transfer via a regulariza… | 40 | 1672 | active |
| ZrrSkywalker/Personalize-SAM PerSAM is the official implementation of 'Personalize Segment Anything Model with One Shot', which customizes the Segment Anything Model (S… | 29 | 1671 | active |
| JiuhaiChen/BLIP3o Official implementation of the BLIP3o-Series, a unified autoregressive-plus-diffusion model for text-to-image generation and editing. It co… | 44 | 1666 | active |
| bamlab/react-native-image-resizer A React Native library for resizing and compressing local images on iOS and Android. It supports JPEG/PNG/WEBP formats, rotation, quality c… | 67 | 1661 | active |
| thunil/TecoGAN TecoGAN is the official source code for a temporally coherent GAN for video super-resolution, published at SIGGRAPH/ACM TOG. It includes in… | 32 | 6141 | maintenance |
| ermig1979/AntiDupl AntiDupl.NET is a free open-source Windows desktop application that finds duplicate and similar images on disk by comparing file contents, … | 70 | 1659 | active |
| chineseocr A Python OCR toolkit that combines YOLO3-based text detection with CRNN/Dense recognition for Chinese and English text in natural scene ima… | 32 | 6123 | maintenance |
| nolanx-ai/nolanx.ai NolanX is an open-source multi-modal agent platform for AI filmmaking that orchestrates text, image, audio, and video models into long-runn… | 52 | 1655 | active |
| davidbyttow/govips govips is a Go library that wraps the libvips image processing library, exposing fast image operations like resizing, format conversion, an… | 79 | 1654 | active |
| taki0112/UGATIT Official TensorFlow implementation of U-GAT-IT, an unsupervised image-to-image translation model using attention modules and adaptive layer… | 32 | 6116 | maintenance |
| cubiq/ComfyUI_IPAdapter_plus A ComfyUI custom node implementation of IPAdapter models for image-to-image conditioning in Stable Diffusion workflows. It transfers the su… | 35 | 6110 | maintenance |
| bytedance/DreamO DreamO is the official PyTorch implementation of a unified framework for image customization, built on FLUX diffusion models. It supports t… | 35 | 1649 | active |
| XueZeyue/DanceGRPO Official implementation of DanceGRPO, a framework applying Group Relative Policy Optimization (GRPO) to fine-tune visual generation models … | 40 | 1648 | active |
| InsightSoftwareConsortium/ITK The Insight Toolkit (ITK) is an open-source, cross-platform C++ library with Python bindings for image analysis, providing algorithms for p… | 94 | 1647 | stable |
| pnggroup/libpng libpng is the official reference C library for reading, creating, and manipulating PNG (Portable Network Graphics) raster image files. It h… | 72 | 1647 | stable |
| JIA-Lab-research/ControlNeXt ControlNeXt is the official implementation of a controllable generation method for images and videos, built on Stable Diffusion XL, Stable … | 24 | 1646 | active |
| nuno-faria/tiler Tiler is a Python tool that builds a large image out of many smaller tile images (circles, legos, minecraft blocks, cross stitches, etc.). … | 32 | 6058 | maintenance |
| CoinCheung/BiSeNet A PyTorch implementation of the BiSeNet V1 and V2 real-time semantic segmentation models, with pretrained weights for Cityscapes, COCO-Stuf… | 57 | 1638 | active |
| RQLuo/MixTeX-Latex-OCR MixTeX is a multimodal OCR application that recognizes LaTeX formulas, tables, and mixed Chinese/English text from images, running entirely… | 22 | 1637 | active |
| luanfujun/deep-painterly-harmonization Research code implementing the 'Deep Painterly Harmonization' algorithm, which seamlessly blends a pasted object into a painting's style us… | 32 | 6042 | maintenance |
| Picsart-AI-Research/StreamingT2V StreamingT2V is a research codebase (CVPR 2025) implementing an autoregressive technique that turns short text-to-video diffusion models li… | 31 | 1630 | active |
| Lucchetto/SuperImage SuperImage is an Android application that upscales low-resolution images using a Real-ESRGAN neural network running on-device via the MNN d… | 10 | 1630 | active |
| Akegarasu/stable-diffusion-inspector A web tool for reading PNG info (generation parameters) from Stable Diffusion generated images and inspecting Stable Diffusion model metada… | 33 | 1628 | active |
| InterDigitalInc/CompressAI CompressAI is a PyTorch library and evaluation platform for end-to-end learned data compression research. It provides custom layers, entrop… | 73 | 1627 | active |
| FuzzyIdeas/Clop Clop is a macOS menu bar application that automatically optimises images, videos, and PDFs as soon as they are copied to the clipboard, min… | 99 | 1623 | active |
| houseofsecrets/SdPaint A Python desktop application that provides a painting canvas where each stroke is sent to a Stable Diffusion (automatic1111) API and the ge… | 20 | 1618 | active |
| dtlnor/stable-diffusion-webui-localization-zh_CN A Simplified Chinese localization extension for AUTOMATIC1111's Stable Diffusion WebUI. It translates the UI and many popular extensions in… | 40 | 1615 | active |
| yossdotpro/removerized Removerized is an open-source, browser-based AI image toolkit for background removal and image upscaling, running entirely client-side via … | 81 | 1610 | active |
| bilibili/ailab Bilibili's AI lab repository, best known for Real-CUGAN, a deep learning model for anime image super-resolution (upscaling). It provides pr… | 23 | 5881 | maintenance |
| Tsuk1ko/cq-picsearcher-bot A Node.js QQ bot that performs reverse image searches via saucenao, ascii2d, soutubot.moe, and trace.moe, connecting to any OneBot 11-compa… | 74 | 1596 | active |
| trekhleb/js-image-carver A JavaScript/TypeScript implementation of the Seam Carving algorithm for content-aware image resizing and object removal. It ships as both … | 56 | 1594 | active |
| SonyResearch/micro_diffusion Official implementation of Sony Research's micro-budget approach to training large-scale text-to-image diffusion transformer models from sc… | 24 | 1594 | active |
| FoundationVision/Infinity Infinity is a bitwise autoregressive text-to-image generation model (CVPR 2025 Oral) with released training and inference code, checkpoints… | 56 | 1587 | active |
| XPixelGroup/HAT HAT (Hybrid Attention Transformer) is a PyTorch implementation of a state-of-the-art transformer model for image super-resolution and resto… | 32 | 1583 | stable |
| ai-forever/ghost GHOST (Generative High-fidelity One Shot Transfer) is a one-shot face swap pipeline for images and videos, published as an IEEE paper and i… | 26 | 1582 | active |
| pangxiaobin/image-matting A free open-source desktop AI image tool built with pywebview and Vue that performs local background removal (matting) using the RMBG-1.4 m… | 86 | 1580 | active |
| posva/catimg catimg is a small C program that prints images (JPEG, PNG, GIF) directly in the terminal using 256-color and unicode escape codes, with no … | 63 | 1577 | stable |
| photosynthesis-team/piq PyTorch Image Quality (PIQ) is a collection of measures and metrics for image quality assessment in image-to-image tasks, written in pure P… | 23 | 1574 | stable |
| IDEA-Research/Rex-Omni Rex-Omni is a 3B-parameter multimodal large language model that unifies object detection, OCR, pointing, keypoint detection, and visual pro… | 47 | 1561 | active |
| LibRaw/LibRaw LibRaw is a C++ library providing a unified interface for reading RAW files from digital cameras, extracting pixel data, processing metadat… | 88 | 1560 | active |
| kritiksoman/GIMP-ML GIMP-ML is a set of Python plugins that bring computer vision and deep learning models into the GNU Image Manipulation Program (GIMP). It p… | 23 | 1553 | active |
| toy/image_optim A Ruby gem and command line tool that optimizes (losslessly, or optionally lossily) JPEG, PNG, GIF, and SVG images by wrapping external uti… | 76 | 1551 | stable |
| tianrun-chen/SAM-Adapter-PyTorch A PyTorch library that adapts Meta AI's Segment Anything Model (SAM, SAM2, SAM3) to underperforming downstream segmentation tasks using lig… | 67 | 1551 | active |
| shrimbly/node-banana Node Banana is an open-source, node-based visual workflow editor for building AI media generation pipelines. Users connect nodes on an infi… | 81 | 1550 | active |
| ssitu/ComfyUI_UltimateSDUpscale A ComfyUI custom node pack implementing the Ultimate Stable Diffusion Upscale script, which runs image-to-image diffusion on large images i… | 69 | 1547 | active |
| baaivision/Emu3.5 Emu3.5 is BAAI's native multimodal foundation model that jointly predicts next states across vision and language, trained on 10T+ interleav… | 43 | 1547 | active |
| wasabeef/Blurry Blurry is an easy-to-use Android library for applying blur effects to views and bitmaps. It supports configurable radius, downsampling, col… | 32 | 5652 | maintenance |
| ShineChen1024/MagicClothing Official PyTorch implementation of Magic Clothing, a diffusion-based model for controllable garment-driven image synthesis (virtual try-on)… | 26 | 1543 | active |
| nuxt/image Nuxt Image is an official Nuxt module providing plug-and-play image optimization for Nuxt applications. It offers drop-in <nuxt-img> and <n… | 89 | 1542 | stable |
| hiroi-sora/PaddleOCR-json An offline OCR command-line executable compiled from PaddleOCR C++ that recognizes text in images and outputs results as JSON strings. It c… | 30 | 1542 | active |
| VNCCS VNCCS is a ComfyUI custom node suite providing an end-to-end pipeline for generating visual novel character sprites with consistent appeara… | 83 | 1541 | active |
| lucidrains/DALLE-pytorch A PyTorch implementation/replication of OpenAI's DALL-E, a text-to-image transformer, including a discrete VAE and optional CLIP for rankin… | 23 | 5627 | maintenance |
| fast-average-color/fast-average-color A small, fast TypeScript library that calculates the average or dominant color of images, videos, canvases, and raw pixel arrays in the bro… | 70 | 1537 | stable |
| Haidra-Org/AI-Horde AI Horde is a crowdsourced distributed cluster where volunteers share GPU/CPU compute to generate AI images and text for others for free. T… | 77 | 1536 | active |
| 16131zzzzzzzz/EveryoneNobel EveryoneNobel is a Python framework built on ComfyUI that generates personalized Nobel Prize portrait images, overlaying text via HTML temp… | 22 | 1534 | active |
| Shawn-Shan/fawkes Fawkes is a privacy protection tool from University of Chicago researchers that adds imperceptible adversarial perturbations to photos to p… | 23 | 5595 | maintenance |
| 4lex4/scantailor-advanced ScanTailor Advanced is an interactive post-processing tool for scanned pages, merging features from ScanTailor Featured and Enhanced while … | 23 | 1532 | stable |
| AIGODLIKE/ComfyUI-BlenderAI-node A Blender addon that integrates ComfyUI into Blender by converting ComfyUI nodes into Blender nodes, enabling AI image generation, material… | 47 | 1531 | active |
| JingyunLiang/SwinIR Official PyTorch implementation of SwinIR, a Swin Transformer-based model for image restoration tasks including super-resolution, denoising… | 23 | 5580 | maintenance |
| hustvl/LightningDiT LightningDiT is a research codebase for latent diffusion models implementing VA-VAE and LightningDiT, achieving FID 1.35 on ImageNet-256 wi… | 47 | 1529 | active |
| BennyKok/comfyui-deploy An open-source, Vercel-like deployment platform for ComfyUI workflows, letting teams share workflows, manage machines (on-premise or server… | 40 | 1529 | active |
| cchen156/Learning-to-See-in-the-Dark TensorFlow implementation of 'Learning to See in the Dark' (CVPR 2018), a deep learning model that brightens very dark, short-exposure RAW … | 51 | 5565 | maintenance |
| AlekPet/ComfyUI_Custom_Nodes_AlekPet A collection of custom nodes for ComfyUI that extend its capabilities with painting, pose control, prompt translation, and speech recogniti… | 72 | 1524 | active |
| TransparentLC/realesrgan-gui A cross-platform graphical interface for the Real-ESRGAN AI image upscaler (with Real-CUGAN support), built in Python with tkinter. It wrap… | 59 | 1524 | active |
| WenmuZhou/PytorchOCR A PyTorch-based OCR toolkit that ports PaddleOCR models to PyTorch, supporting common text detection and recognition algorithms like the PP… | 59 | 1523 | active |
| anishathalye/neural-style A Python command-line tool implementing the neural style transfer algorithm (Gatys et al.) in TensorFlow, applying the style of one image t… | 67 | 5541 | maintenance |
| tin2tin/Pallaidium Pallaidium is a free, open-source generative AI movie studio implemented as a Blender add-on integrated into the Video Sequence Editor (VSE… | 75 | 1520 | active |
| caiyuanhao1998/Retinexformer Retinexformer is a one-stage Retinex-based Transformer model and toolbox for low-light image enhancement, published at ICCV 2023. It suppor… | 66 | 1518 | active |
| BrokenSource/DepthFlow DepthFlow is a free, open-source Python application and library that converts still images into 3D parallax effect videos using monocular d… | 85 | 1512 | active |
| HiDream-ai/HiDream-O1-Image HiDream-O1-Image is an open-weights 8B image generation foundation model built on a Pixel-level Unified Transformer (UiT) that natively enc… | 53 | 1510 | active |
| hkchengrex/Tracking-Anything-with-DEVA DEVA is a decoupled video segmentation framework that combines task-specific image-level segmentation models with a universal bi-directiona… | 27 | 1508 | stable |
| CanHub/Android-Image-Cropper A Kotlin image cropping library for Android, optimized for images picked from the camera or gallery. It provides a customizable CropImageVi… | 65 | 1504 | active |
| meiqua/shape_based_matching A C++ library implementing Halcon-style shape-based matching (equivalent to LINE-MOD) using gradient orientation templates for robust 2D ob… | 32 | 1500 | active |
| piddnad/DDColor DDColor is the official PyTorch implementation of an ICCV 2023 paper on photo-realistic automatic image colorization using dual decoders an… | 59 | 1496 | active |
| NJU-PCALab/STAR STAR is a research implementation of an ICCV 2025 paper performing real-world video super-resolution using spatial-temporal augmentation wi… | 34 | 1495 | active |
| una/CSSgram CSSgram is a tiny CSS/Sass library that recreates Instagram-style photo filters using CSS filters and blend modes. Filters are applied by a… | 23 | 5395 | maintenance |
| emcconville/wand Wand is a ctypes-based Python binding for the ImageMagick MagickWand API, supporting Python 3.8+ and PyPy. It exposes the full MagickWand f… | 90 | 1478 | active |
| aws-solutions/dynamic-image-transformation-for-amazon-cloudfront An official AWS Solutions implementation that deploys a serverless architecture for real-time image transformation, optimization, and deliv… | 97 | 1467 | stable |
| huangserva/skill-prompt-generator A Claude Code Skills-based system that generates high-quality AI image prompts by intelligently combining elements from a 1,246-item Univer… | 52 | 1467 | active |
| hojonathanho/diffusion The official reference implementation of Denoising Diffusion Probabilistic Models (DDPM) from the 2020 paper by Jonathan Ho et al., written… | 32 | 5300 | maintenance |
| Hello-hao/Tbed Hellohao Image Hosting (Tbed) is an open-source, self-hosted image hosting application built with Java and SpringBoot using a front-end/bac… | 67 | 1457 | active |
| dcm4che/dcm4che dcm4che is a Java toolkit and library implementing the DICOM standard for medical imaging, including data set handling, network communicati… | 91 | 1450 | active |
| psd-tools/psd-tools psd-tools is a Python package for reading and manipulating Adobe Photoshop PSD/PSB files. It parses the low-level file structure, exports l… | 98 | 1449 | active |
| zsyOAOA/InvSR InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff… | 51 | 1443 | active |
| serratus/quaggaJS QuaggaJS is a barcode-scanner library written entirely in JavaScript that supports real-time localization and decoding of barcode types suc… | 23 | 5207 | maintenance |
| sdcb/PaddleSharp A .NET/C# wrapper around Baidu's PaddleInference C API, providing PaddleOCR, PaddleDetection, rotation detection, Chinese segmentation, and… | 72 | 1441 | active |
| tianweiy/DMD2 DMD2 is the official PyTorch implementation of Improved Distribution Matching Distillation, a NeurIPS 2024 method that distills diffusion m… | 28 | 1438 | active |
| neuralchen/SimSwap SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It… | 23 | 5188 | maintenance |
| jakowenko/double-take Double Take is a self-hosted Docker application providing a unified UI and API for facial recognition. It abstracts multiple face detection… | 40 | 1436 | active |
| Haneke Haneke is a lightweight generic cache library for iOS and tvOS written in Swift, with memory and LRU disk caching for UIImage, NSData, JSON… | 23 | 5155 | maintenance |
| qupath/qupath QuPath is an open-source desktop application for bioimage analysis, aimed especially at digital pathology and whole-slide imaging. It provi… | 79 | 1430 | active |
| Francis-Rings/StableAnimator StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po… | 41 | 1430 | active |
| zsyOAOA/ResShift ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi… | 62 | 1427 | active |