Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: image-processing

1843 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
ggchivalrous/yiyin
Yiyin (壹印) is a free, open-source desktop application for adding watermark frames to photos, generating styled image borders with customiza…
831703active
zai-org/ImageReward
ImageReward is the first general-purpose human preference reward model for text-to-image generation, trained on 137k expert comparison pair…
521702stable
sihyun-yu/REPA
REPA is the official PyTorch implementation of the ICLR 2025 paper 'Representation Alignment for Generation', a regularization technique th…
271700active
GreycLab/CImg
CImg is a small, open-source, header-only C++ template library for image processing. It provides a single image class supporting up to 4-di…
761690stable
sedthh/pyxelate
Pyxelate is a Python library and CLI tool that converts images into 8-bit pixel art by downsampling and learning a reduced color palette. I…
511685active
williamyang1991/DualStyleGAN
Official PyTorch implementation of DualStyleGAN, a CVPR 2022 model for exemplar-based high-resolution (1024px) portrait style transfer. It …
321683stable
fire-keeper/BlindWatermark
A Python library and CLI/GUI tool that embeds invisible blind watermarks into images using discrete wavelet transforms, protecting creators…
231675active
franciszzj/Leffa
Leffa is a diffusion-based framework for controllable person image generation, supporting virtual try-on and pose transfer via a regulariza…
401672active
ZrrSkywalker/Personalize-SAM
PerSAM is the official implementation of 'Personalize Segment Anything Model with One Shot', which customizes the Segment Anything Model (S…
291671active
JiuhaiChen/BLIP3o
Official implementation of the BLIP3o-Series, a unified autoregressive-plus-diffusion model for text-to-image generation and editing. It co…
441666active
bamlab/react-native-image-resizer
A React Native library for resizing and compressing local images on iOS and Android. It supports JPEG/PNG/WEBP formats, rotation, quality c…
671661active
thunil/TecoGAN
TecoGAN is the official source code for a temporally coherent GAN for video super-resolution, published at SIGGRAPH/ACM TOG. It includes in…
326141maintenance
ermig1979/AntiDupl
AntiDupl.NET is a free open-source Windows desktop application that finds duplicate and similar images on disk by comparing file contents, …
701659active
chineseocr
A Python OCR toolkit that combines YOLO3-based text detection with CRNN/Dense recognition for Chinese and English text in natural scene ima…
326123maintenance
nolanx-ai/nolanx.ai
NolanX is an open-source multi-modal agent platform for AI filmmaking that orchestrates text, image, audio, and video models into long-runn…
521655active
davidbyttow/govips
govips is a Go library that wraps the libvips image processing library, exposing fast image operations like resizing, format conversion, an…
791654active
taki0112/UGATIT
Official TensorFlow implementation of U-GAT-IT, an unsupervised image-to-image translation model using attention modules and adaptive layer…
326116maintenance
cubiq/ComfyUI_IPAdapter_plus
A ComfyUI custom node implementation of IPAdapter models for image-to-image conditioning in Stable Diffusion workflows. It transfers the su…
356110maintenance
bytedance/DreamO
DreamO is the official PyTorch implementation of a unified framework for image customization, built on FLUX diffusion models. It supports t…
351649active
XueZeyue/DanceGRPO
Official implementation of DanceGRPO, a framework applying Group Relative Policy Optimization (GRPO) to fine-tune visual generation models …
401648active
InsightSoftwareConsortium/ITK
The Insight Toolkit (ITK) is an open-source, cross-platform C++ library with Python bindings for image analysis, providing algorithms for p…
941647stable
pnggroup/libpng
libpng is the official reference C library for reading, creating, and manipulating PNG (Portable Network Graphics) raster image files. It h…
721647stable
JIA-Lab-research/ControlNeXt
ControlNeXt is the official implementation of a controllable generation method for images and videos, built on Stable Diffusion XL, Stable …
241646active
nuno-faria/tiler
Tiler is a Python tool that builds a large image out of many smaller tile images (circles, legos, minecraft blocks, cross stitches, etc.). …
326058maintenance
CoinCheung/BiSeNet
A PyTorch implementation of the BiSeNet V1 and V2 real-time semantic segmentation models, with pretrained weights for Cityscapes, COCO-Stuf…
571638active
RQLuo/MixTeX-Latex-OCR
MixTeX is a multimodal OCR application that recognizes LaTeX formulas, tables, and mixed Chinese/English text from images, running entirely…
221637active
luanfujun/deep-painterly-harmonization
Research code implementing the 'Deep Painterly Harmonization' algorithm, which seamlessly blends a pasted object into a painting's style us…
326042maintenance
Picsart-AI-Research/StreamingT2V
StreamingT2V is a research codebase (CVPR 2025) implementing an autoregressive technique that turns short text-to-video diffusion models li…
311630active
Lucchetto/SuperImage
SuperImage is an Android application that upscales low-resolution images using a Real-ESRGAN neural network running on-device via the MNN d…
101630active
Akegarasu/stable-diffusion-inspector
A web tool for reading PNG info (generation parameters) from Stable Diffusion generated images and inspecting Stable Diffusion model metada…
331628active
InterDigitalInc/CompressAI
CompressAI is a PyTorch library and evaluation platform for end-to-end learned data compression research. It provides custom layers, entrop…
731627active
FuzzyIdeas/Clop
Clop is a macOS menu bar application that automatically optimises images, videos, and PDFs as soon as they are copied to the clipboard, min…
991623active
houseofsecrets/SdPaint
A Python desktop application that provides a painting canvas where each stroke is sent to a Stable Diffusion (automatic1111) API and the ge…
201618active
dtlnor/stable-diffusion-webui-localization-zh_CN
A Simplified Chinese localization extension for AUTOMATIC1111's Stable Diffusion WebUI. It translates the UI and many popular extensions in…
401615active
yossdotpro/removerized
Removerized is an open-source, browser-based AI image toolkit for background removal and image upscaling, running entirely client-side via …
811610active
bilibili/ailab
Bilibili's AI lab repository, best known for Real-CUGAN, a deep learning model for anime image super-resolution (upscaling). It provides pr…
235881maintenance
Tsuk1ko/cq-picsearcher-bot
A Node.js QQ bot that performs reverse image searches via saucenao, ascii2d, soutubot.moe, and trace.moe, connecting to any OneBot 11-compa…
741596active
trekhleb/js-image-carver
A JavaScript/TypeScript implementation of the Seam Carving algorithm for content-aware image resizing and object removal. It ships as both …
561594active
SonyResearch/micro_diffusion
Official implementation of Sony Research's micro-budget approach to training large-scale text-to-image diffusion transformer models from sc…
241594active
FoundationVision/Infinity
Infinity is a bitwise autoregressive text-to-image generation model (CVPR 2025 Oral) with released training and inference code, checkpoints…
561587active
XPixelGroup/HAT
HAT (Hybrid Attention Transformer) is a PyTorch implementation of a state-of-the-art transformer model for image super-resolution and resto…
321583stable
ai-forever/ghost
GHOST (Generative High-fidelity One Shot Transfer) is a one-shot face swap pipeline for images and videos, published as an IEEE paper and i…
261582active
pangxiaobin/image-matting
A free open-source desktop AI image tool built with pywebview and Vue that performs local background removal (matting) using the RMBG-1.4 m…
861580active
posva/catimg
catimg is a small C program that prints images (JPEG, PNG, GIF) directly in the terminal using 256-color and unicode escape codes, with no …
631577stable
photosynthesis-team/piq
PyTorch Image Quality (PIQ) is a collection of measures and metrics for image quality assessment in image-to-image tasks, written in pure P…
231574stable
IDEA-Research/Rex-Omni
Rex-Omni is a 3B-parameter multimodal large language model that unifies object detection, OCR, pointing, keypoint detection, and visual pro…
471561active
LibRaw/LibRaw
LibRaw is a C++ library providing a unified interface for reading RAW files from digital cameras, extracting pixel data, processing metadat…
881560active
kritiksoman/GIMP-ML
GIMP-ML is a set of Python plugins that bring computer vision and deep learning models into the GNU Image Manipulation Program (GIMP). It p…
231553active
toy/image_optim
A Ruby gem and command line tool that optimizes (losslessly, or optionally lossily) JPEG, PNG, GIF, and SVG images by wrapping external uti…
761551stable
tianrun-chen/SAM-Adapter-PyTorch
A PyTorch library that adapts Meta AI's Segment Anything Model (SAM, SAM2, SAM3) to underperforming downstream segmentation tasks using lig…
671551active
shrimbly/node-banana
Node Banana is an open-source, node-based visual workflow editor for building AI media generation pipelines. Users connect nodes on an infi…
811550active
ssitu/ComfyUI_UltimateSDUpscale
A ComfyUI custom node pack implementing the Ultimate Stable Diffusion Upscale script, which runs image-to-image diffusion on large images i…
691547active
baaivision/Emu3.5
Emu3.5 is BAAI's native multimodal foundation model that jointly predicts next states across vision and language, trained on 10T+ interleav…
431547active
wasabeef/Blurry
Blurry is an easy-to-use Android library for applying blur effects to views and bitmaps. It supports configurable radius, downsampling, col…
325652maintenance
ShineChen1024/MagicClothing
Official PyTorch implementation of Magic Clothing, a diffusion-based model for controllable garment-driven image synthesis (virtual try-on)…
261543active
nuxt/image
Nuxt Image is an official Nuxt module providing plug-and-play image optimization for Nuxt applications. It offers drop-in <nuxt-img> and <n…
891542stable
hiroi-sora/PaddleOCR-json
An offline OCR command-line executable compiled from PaddleOCR C++ that recognizes text in images and outputs results as JSON strings. It c…
301542active
VNCCS
VNCCS is a ComfyUI custom node suite providing an end-to-end pipeline for generating visual novel character sprites with consistent appeara…
831541active
lucidrains/DALLE-pytorch
A PyTorch implementation/replication of OpenAI's DALL-E, a text-to-image transformer, including a discrete VAE and optional CLIP for rankin…
235627maintenance
fast-average-color/fast-average-color
A small, fast TypeScript library that calculates the average or dominant color of images, videos, canvases, and raw pixel arrays in the bro…
701537stable
Haidra-Org/AI-Horde
AI Horde is a crowdsourced distributed cluster where volunteers share GPU/CPU compute to generate AI images and text for others for free. T…
771536active
16131zzzzzzzz/EveryoneNobel
EveryoneNobel is a Python framework built on ComfyUI that generates personalized Nobel Prize portrait images, overlaying text via HTML temp…
221534active
Shawn-Shan/fawkes
Fawkes is a privacy protection tool from University of Chicago researchers that adds imperceptible adversarial perturbations to photos to p…
235595maintenance
4lex4/scantailor-advanced
ScanTailor Advanced is an interactive post-processing tool for scanned pages, merging features from ScanTailor Featured and Enhanced while …
231532stable
AIGODLIKE/ComfyUI-BlenderAI-node
A Blender addon that integrates ComfyUI into Blender by converting ComfyUI nodes into Blender nodes, enabling AI image generation, material…
471531active
JingyunLiang/SwinIR
Official PyTorch implementation of SwinIR, a Swin Transformer-based model for image restoration tasks including super-resolution, denoising…
235580maintenance
hustvl/LightningDiT
LightningDiT is a research codebase for latent diffusion models implementing VA-VAE and LightningDiT, achieving FID 1.35 on ImageNet-256 wi…
471529active
BennyKok/comfyui-deploy
An open-source, Vercel-like deployment platform for ComfyUI workflows, letting teams share workflows, manage machines (on-premise or server…
401529active
cchen156/Learning-to-See-in-the-Dark
TensorFlow implementation of 'Learning to See in the Dark' (CVPR 2018), a deep learning model that brightens very dark, short-exposure RAW …
515565maintenance
AlekPet/ComfyUI_Custom_Nodes_AlekPet
A collection of custom nodes for ComfyUI that extend its capabilities with painting, pose control, prompt translation, and speech recogniti…
721524active
TransparentLC/realesrgan-gui
A cross-platform graphical interface for the Real-ESRGAN AI image upscaler (with Real-CUGAN support), built in Python with tkinter. It wrap…
591524active
WenmuZhou/PytorchOCR
A PyTorch-based OCR toolkit that ports PaddleOCR models to PyTorch, supporting common text detection and recognition algorithms like the PP…
591523active
anishathalye/neural-style
A Python command-line tool implementing the neural style transfer algorithm (Gatys et al.) in TensorFlow, applying the style of one image t…
675541maintenance
tin2tin/Pallaidium
Pallaidium is a free, open-source generative AI movie studio implemented as a Blender add-on integrated into the Video Sequence Editor (VSE…
751520active
caiyuanhao1998/Retinexformer
Retinexformer is a one-stage Retinex-based Transformer model and toolbox for low-light image enhancement, published at ICCV 2023. It suppor…
661518active
BrokenSource/DepthFlow
DepthFlow is a free, open-source Python application and library that converts still images into 3D parallax effect videos using monocular d…
851512active
HiDream-ai/HiDream-O1-Image
HiDream-O1-Image is an open-weights 8B image generation foundation model built on a Pixel-level Unified Transformer (UiT) that natively enc…
531510active
hkchengrex/Tracking-Anything-with-DEVA
DEVA is a decoupled video segmentation framework that combines task-specific image-level segmentation models with a universal bi-directiona…
271508stable
CanHub/Android-Image-Cropper
A Kotlin image cropping library for Android, optimized for images picked from the camera or gallery. It provides a customizable CropImageVi…
651504active
meiqua/shape_based_matching
A C++ library implementing Halcon-style shape-based matching (equivalent to LINE-MOD) using gradient orientation templates for robust 2D ob…
321500active
piddnad/DDColor
DDColor is the official PyTorch implementation of an ICCV 2023 paper on photo-realistic automatic image colorization using dual decoders an…
591496active
NJU-PCALab/STAR
STAR is a research implementation of an ICCV 2025 paper performing real-world video super-resolution using spatial-temporal augmentation wi…
341495active
una/CSSgram
CSSgram is a tiny CSS/Sass library that recreates Instagram-style photo filters using CSS filters and blend modes. Filters are applied by a…
235395maintenance
emcconville/wand
Wand is a ctypes-based Python binding for the ImageMagick MagickWand API, supporting Python 3.8+ and PyPy. It exposes the full MagickWand f…
901478active
aws-solutions/dynamic-image-transformation-for-amazon-cloudfront
An official AWS Solutions implementation that deploys a serverless architecture for real-time image transformation, optimization, and deliv…
971467stable
huangserva/skill-prompt-generator
A Claude Code Skills-based system that generates high-quality AI image prompts by intelligently combining elements from a 1,246-item Univer…
521467active
hojonathanho/diffusion
The official reference implementation of Denoising Diffusion Probabilistic Models (DDPM) from the 2020 paper by Jonathan Ho et al., written…
325300maintenance
Hello-hao/Tbed
Hellohao Image Hosting (Tbed) is an open-source, self-hosted image hosting application built with Java and SpringBoot using a front-end/bac…
671457active
dcm4che/dcm4che
dcm4che is a Java toolkit and library implementing the DICOM standard for medical imaging, including data set handling, network communicati…
911450active
psd-tools/psd-tools
psd-tools is a Python package for reading and manipulating Adobe Photoshop PSD/PSB files. It parses the low-level file structure, exports l…
981449active
zsyOAOA/InvSR
InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff…
511443active
serratus/quaggaJS
QuaggaJS is a barcode-scanner library written entirely in JavaScript that supports real-time localization and decoding of barcode types suc…
235207maintenance
sdcb/PaddleSharp
A .NET/C# wrapper around Baidu's PaddleInference C API, providing PaddleOCR, PaddleDetection, rotation detection, Chinese segmentation, and…
721441active
tianweiy/DMD2
DMD2 is the official PyTorch implementation of Improved Distribution Matching Distillation, a NeurIPS 2024 method that distills diffusion m…
281438active
neuralchen/SimSwap
SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It…
235188maintenance
jakowenko/double-take
Double Take is a self-hosted Docker application providing a unified UI and API for facial recognition. It abstracts multiple face detection…
401436active
Haneke
Haneke is a lightweight generic cache library for iOS and tvOS written in Swift, with memory and LRU disk caching for UIImage, NSData, JSON…
235155maintenance
qupath/qupath
QuPath is an open-source desktop application for bioimage analysis, aimed especially at digital pathology and whole-slide imaging. It provi…
791430active
Francis-Rings/StableAnimator
StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po…
411430active
zsyOAOA/ResShift
ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi…
621427active

← prev page 7 / 19 next →