Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: image-processing

1843 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
zuruoke/watermark-removal
A machine learning tool that removes watermarks from images using deep learning image inpainting, based on Contextual Attention and Gated C…
835139maintenance
Tom94/tev
tev is a high dynamic range (HDR/EDR) image viewer written in C++ for people who care about accurate colors. It supports many formats (EXR,…
991426active
mdbloice/Augmentor
Augmentor is a standalone Python library for image augmentation in machine learning, providing a pipeline of stochastic operations like rot…
325133maintenance
apple/ml-aim
Apple's official repository for AIM (Autoregressive Image Models), providing code and pretrained checkpoints for AIMv1 and AIMv2 large visi…
421424active
mix1009/sdwebuiapi
A Python client library for the AUTOMATIC1111 stable-diffusion-webui REST API, wrapping endpoints like txt2img, img2img, and upscaling into…
321423active
affinelayer/pix2pix-tensorflow
A TensorFlow implementation of pix2pix, a conditional GAN that learns a mapping from input images to output images. It is a faithful port o…
325081maintenance
CSAILVision/semantic-segmentation-pytorch
A PyTorch implementation of semantic segmentation (scene parsing) models for the MIT ADE20K dataset, including pretrained model zoo and tra…
325078maintenance
khanamiryan/php-qrcode-detector-decoder
A pure PHP library for detecting and decoding QR codes from images, ported from the ZXing library. It works without any PHP extensions beyo…
371412active
SmileyChris/easy-thumbnails
A Django application for generating and managing image thumbnails, supporting predefined aliases, template tags, and Python API usage. It s…
601410active
IDEA-Research/DINO-X-API
DINO-X API is a Python client library and examples for accessing DINO-X, a hosted unified vision model for open-world object detection and …
361410active
gnobitab/InstaFlow
InstaFlow is a one-step text-to-image generation model based on Rectified Flow, enabling ultra-fast Stable Diffusion inference without iter…
281409active
willmiao/ComfyUI-Lora-Manager
A ComfyUI extension that provides a web-based manager for organizing, previewing, downloading, and applying LoRA models with metadata and r…
831404active
nachifur/MulimgViewer
MulimgViewer is a Python-based multi-image viewer that displays many images in a single interface for side-by-side comparison, parallel sel…
661400active
imagej/imagej2
ImageJ2 is an open-source Java framework and application for processing and analyzing N-dimensional scientific image data, built on the Img…
751399stable
yfeng95/PRNet
PRNet is a Python/TensorFlow implementation of the ECCV 2018 Position Map Regression Network for joint 3D face reconstruction and dense ali…
325013maintenance
liyue-aigc/female-portrait-director
A modular Codex Skill that directs and expands detailed AI female portrait prompts into coherent, camera-ready scenes, or generates images …
731396active
jexom/sd-webui-depth-lib
A depth map library and poser plugin for the Automatic1111 stable-diffusion-webui, designed to supply depth maps to the ControlNet extensio…
101395stable
logtd/ComfyUI-Fluxtapoz
A set of ComfyUI custom nodes for editing and stylizing images with Flux models, implementing techniques like RF-Inversion, RF-Edit, Firefl…
231392active
exif-js/exif-js
Exif.js is a JavaScript library for reading EXIF and IPTC metadata from JPEG and TIFF images in the browser. It works with image elements o…
234979maintenance
claviska/SimpleImage
A single-file PHP library that wraps the GD extension with a fluent, chainable API for loading, manipulating, and saving images. It support…
531388stable
cszn/BSRGAN
BSRGAN is a PyTorch implementation of a practical degradation model for deep blind image super-resolution, presented at ICCV 2021. It provi…
321386stable
14790897/handwriting-web
A self-hostable web application that converts typed text into images (or PDFs) simulating handwriting, using uploaded fonts, background ima…
921385active
OpenGVLab/DragGAN
An unofficial full-featured Python implementation of DragGAN, the interactive point-based image manipulation method built on StyleGAN gener…
294946maintenance
sdsykes/fastimage
FastImage is a Ruby library that finds the size or type of an image given its URI by fetching only the minimal bytes needed. It supports ma…
751380stable
zhixuhao/unet
A Keras implementation of the U-Net convolutional network architecture for image segmentation, based on the original biomedical segmentatio…
664941maintenance
keyu-tian/SparK
SparK is the official PyTorch implementation of an ICLR 2023 Spotlight paper that applies BERT/MAE-style masked image modeling to convoluti…
221376stable
qubvel/segmentation_models
A Python library providing neural network architectures for image segmentation (Unet, FPN, Linknet, PSPNet) built on Keras and TensorFlow K…
234923maintenance
ali-vilab/TeaCache
TeaCache is a training-free caching approach that accelerates inference for video diffusion models by estimating output differences across …
331369active
spatie/image
A PHP library for manipulating images with an expressive, fluent API, supporting drivers like Imagick and GD. It provides operations such a…
941364stable
ImprintLab/MedSegDiff
MedSegDiff is a diffusion probabilistic model framework for segmenting and reconstructing organs and tissues from medical images, with a tr…
501363active
bytedance/UNO
UNO is a research framework from ByteDance for subject-driven image generation with diffusion transformers, supporting both single- and mul…
381362active
scraed/LanPaint
LanPaint is a training-free inpainting sampler for stable diffusion models, implemented as a ComfyUI custom node. It uses iterative 'think …
851361active
mzucker/noteshrink
A Python command-line script that cleans up scans and photos of handwritten notes by separating background from ink, quantizing colors, and…
324841maintenance
lxtGH/OMG-Seg
Official research codebase for OMG-Seg (CVPR 2024) and OMG-LLaVA (NeurIPS 2024), unified models for image-level, object-level, and pixel-le…
471354active
zmrlft/GreenWall
A Wails-based desktop application that customizes your GitHub contribution graph by converting uploaded images into contribution heatmap pa…
681352active
wyhuai/DDNM
DDNM is a Python research codebase implementing the Denoising Diffusion Null-Space Model for zero-shot image restoration, published as an I…
321349stable
jhc13/taggui
TagGUI is a cross-platform desktop application for quickly adding and editing image tags and captions, aimed at creators of image datasets …
531346active
webtoon/psd
A fast, zero-dependency TypeScript parser for Adobe Photoshop PSD/PSB files that runs in web browsers and Node.js. It extracts layer inform…
231345active
jbilcke-hf/ai-comic-factory
AI Comic Factory is a web application that generates comic panels and pages from a single text prompt by combining an LLM for story/dialogu…
101345active
imgbot/Imgbot
Imgbot is a GitHub App backed by an Azure Functions solution that crawls a repository's image files and losslessly compresses them, then op…
241344active
muzishen/IMAGDressing
IMAGDressing-v1 is a diffusion-based framework for customizable virtual dressing that generates human images with fixed garments and contro…
441343active
zanllp/infinite-image-browsing
Infinite Image Browsing (IIB) is a full-featured image and video management application with fast thumbnail-based browsing, AI-generation m…
931341active
FireRedTeam/FireRed-Image-Edit
FireRed-Image-Edit is an open-source image editing foundation model built on diffusion models, released as PyTorch model weights with infer…
491341active
receyuki/stable-diffusion-prompt-reader
A standalone desktop application (with GUI and CLI) for reading generation prompts and parameters embedded in Stable Diffusion images, with…
211339stable
senguptaumd/Background-Matting
Official research code for 'Background Matting: The World is Your Green Screen' (CVPR 2020), a deep network that extracts per-pixel alpha m…
324769maintenance
christophschuhmann/improved-aesthetic-predictor
A CLIP+MLP neural network that predicts how much people on average like an image, trained on AVA dataset ratings. It is widely used for fil…
321338stable
ByteDance-Seed/SeedVR
SeedVR/SeedVR2 are diffusion-transformer based models for generic real-world and AIGC video and image restoration, with SeedVR2 using adver…
471334active
jtydhr88/ComfyUI-qwenmultiangle
A ComfyUI custom node providing an interactive Three.js 3D viewport for controlling camera azimuth, elevation, and zoom. It outputs formatt…
561332active
ldqk/ImageSearch
A .NET 10 desktop demo application that performs reverse image search (search by image) over local hard drives with tens of millions of ima…
931329active
Gourieff/ComfyUI-ReActor
ComfyUI-ReActor is a fast and simple face swap extension node for ComfyUI, based on the ReActor face-swapping engine. It includes a nudity …
651327active
numandev1/react-native-compressor
A React Native library that compresses images, videos, and audio with WhatsApp-like quality, plus background upload, file download, and vid…
961325active
cvzone/cvzone
CVZone is a Python computer vision helper library that wraps OpenCV and MediaPipe to simplify image processing and AI functions like hand t…
321325active
nateraw/stable-diffusion-videos
A Python library for generating videos with Stable Diffusion by walking the latent space and morphing between text prompts. It supports bea…
564705maintenance
thephpleague/color-extractor
A PHP library that extracts the most representative colors from an image, building a color palette sorted by pixel count. It handles transp…
881323stable
meta-pytorch/segment-anything-fast
A fast, batched offline inference-oriented fork of Meta's Segment Anything (SAM) image segmentation model. It applies optimizations like bf…
451321active
Decimation/SmartImage
SmartImage is a reverse image search tool that queries multiple engines (SauceNao, IQDB, Ascii2D, trace.moe, TinEye, Yandex, and more) and …
971320active
nroduit/Weasis
Weasis is an open-source DICOM viewer for medical imaging that runs standalone or embedded in web applications. It integrates with PACS, VN…
951317active
Capsize-Games/airunner
AI Runner is a privacy-focused desktop application for running local AI models offline, combining an AI chat companion with voice conversat…
791314active
ali-vilab/MimicBrush
MimicBrush is the official implementation of a zero-shot image editing method that lets users mask a region in a source image and provide a…
241311active
flozz/StackBlur
StackBlur.js is a JavaScript library implementing the fast, almost-Gaussian StackBlur algorithm for images and canvas elements. It works in…
731310stable
spatie/laravel-image-optimizer
A Laravel package that optimizes PNG, JPG, SVG, and GIF images by running them through a chain of installed optimization binaries. It is th…
771309stable
derrian-distro/LoRA_Easy_Training_Scripts
A PySide6 desktop GUI that wraps Kohya's sd-scripts to simplify training LoRA, LoCon, and other LoRA-type models for Stable Diffusion. It s…
331307active
seetaface/SeetaFaceEngine
SeetaFace Engine is an open-source C++ face recognition engine comprising face detection, face alignment, and face identification modules. …
324636maintenance
Vincentqyw/image-matching-webui
A Gradio-based web UI that matches keypoints between two images using many state-of-the-art image matching algorithms (LoFTR, SuperGlue, Li…
911302active
luosiallen/latent-consistency-model
Official implementation of Latent Consistency Models (LCM), a diffusion-based approach for synthesizing high-resolution images with few-ste…
274615maintenance
rsmbl/Resemble.js
Resemble.js is a JavaScript library for analyzing and comparing images using HTML5 canvas, producing pixel-level diff images and analysis d…
324612maintenance
sighook/pixload
pixload is a set of Perl CLI tools for creating and injecting payloads into image files (BMP, GIF, JPG, PNG, WebP). It is used in offensive…
231300active
Uminosachi/sd-webui-inpaint-anything
A Stable Diffusion Web UI extension that performs inpainting and outpainting using masks generated by Segment Anything models (SAM 2, SAM-H…
301298active
PyImageSearch/imutils
A Python library of convenience functions that simplify common OpenCV image processing tasks such as translation, rotation, resizing, skele…
324590maintenance
ARM-software/astc-encoder
The Arm ASTC Encoder (astcenc) is a command-line tool and codec library for compressing and decompressing images in the Adaptive Scalable T…
921295stable
W2GenAI-Lab/LucidFlux
LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a…
551293active
Parskatt/RoMa
RoMa (romatch) is a Python library for robust dense feature matching between image pairs, estimating pixel-dense warps and reliable certain…
521293active
ClownsharkBatwing/RES4LYF
RES4LYF is a ComfyUI custom node collection providing advanced diffusion samplers (RES samplers) that achieve high-quality image generation…
671289active
t3dotgg/quickpic
QuickPic is a free, open-source web app by Theo that converts SVGs to high-resolution PNGs and offers other quick image utilities like squa…
221289active
azagaya/laigter
Laigter is an open-source desktop tool that automatically generates normal, specular, parallax, and occlusion maps from 2D textures, design…
901288active
RickdeJager/stegseek
Stegseek is a lightning-fast command-line cracker for steghide steganography, built as a fork of the original steghide project that can tes…
231288stable
bytedance/Bernini
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer perf…
571287active
lzhgus/Capso
Capso is a free, open-source native macOS app for screenshots and screen recording, built with Swift 6.0 and SwiftUI as an alternative to C…
811286active
huanngzh/MV-Adapter
MV-Adapter is a plug-and-play adapter that turns pre-trained text-to-image diffusion models (e.g., SDXL, SD2.1) into multi-view consistent …
341285active
zju3dv/MatchAnything
MatchAnything is a deep learning model for universal cross-modality image matching, released as research code accompanying a TPAMI 2026 pap…
641279active
plemeri/transparent-background
A Python tool and CLI that removes backgrounds from images and videos using the InSPyReNet deep learning model (ACCV 2022). It supports ima…
631278active
Tencent-Hunyuan/SRPO
SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion…
541278active
NVlabs/stylegan2-ada-pytorch
Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi…
324487maintenance
manycore-maas/Painter
Painter is a JSON-driven canvas drawing library for WeChat mini programs (also usable in Node and HTML5) that renders images from a declara…
234475maintenance
vipshop/cache-dit
Cache-DiT is a PyTorch-native inference engine that accelerates Diffusion Transformer (DiT) models with hybrid caching, parallelism, quanti…
821267active
DimitarPetrov/stegify
stegify is a Go command line tool and library for LSB (Least Significant Bit) steganography that hides any file inside images such as PNG a…
231267stable
blueimp/JavaScript-Load-Image
A JavaScript library that loads images from File/Blob objects or URLs and returns optionally scaled, cropped, or rotated HTML img or canvas…
324456maintenance
brendan-duncan/image
A pure-Dart library for decoding, encoding, and manipulating images in many formats (PNG, JPEG, GIF, WebP, TIFF, BMP, and more). It works w…
761263active
bryandlee/animegan2-pytorch
A PyTorch implementation of AnimeGANv2, a GAN-based image-to-image style transfer model that converts photos into anime-style images. It pr…
324452maintenance
GraphiteEditor/Graphite
Graphite is a free, open source 2D graphics editor built in Rust that combines layer-based compositing with a node-based procedural graphic…
6726939experimental
ashuoAI/SHUO-Canvas
SHUO Canvas (formerly AI-CanvasPro) is an AI multimodal creation canvas application that lets users combine text, images, video, and audio …
821261active
rlawjdghek/StableVITON
StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent…
471261stable
MemeMeow-Studio/MemeMeow
MemeMeow is a self-hosted meme/sticker management and retrieval application that lets users find images by describing the desired scene in …
651260active
ShiftHackZ/Stable-Diffusion-KMP
SDAI is an open-source, cross-platform Stable Diffusion client app for Android and iOS built with Kotlin Multiplatform and Jetpack Compose.…
911259active
Francis-Rings/StableAvatar
StableAvatar is an end-to-end video diffusion transformer that generates infinite-length, high-quality talking avatar videos from a referen…
461258active
VainF/pytorch-msssim
A PyTorch library providing fast, differentiable SSIM and MS-SSIM image quality metrics using separable Gaussian filtering for speed. It ca…
231253stable
meowtec/Imagine
Imagine is a cross-platform desktop GUI app for optimizing PNG, JPEG, and WebP images, built on pngquant, mozjpeg, and WebP encoders with a…
234405maintenance
antfu/qrcode-toolkit
A web-based toolkit for generating base QR codes and refining AI-generated QR codes by comparing outputs to find misaligned pixels. It also…
291250active
sentinel-hub/eo-learn
eo-learn is a collection of open-source Python packages for accessing and processing spatio-temporal satellite imagery, built around modula…
511247active
MikeKovarik/exifr
exifr is a fast, dependency-free JavaScript library for reading EXIF and other image metadata (TIFF, XMP, ICC, IPTC, GPS, JFIF) from JPG, P…
231247active

← prev page 8 / 19 next →