Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: image-processing

1843 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
cbh123/stickerbaker
StickerBaker is an open-source web application that generates AI stickers from text prompts or face uploads, powered by Replicate models (f…
251100active
eszdman/PhotonCamera
PhotonCamera is an open-source Android camera app that applies enhanced computational photography image processing to captured photos. It u…
681099active
HswAI2026/JuZhou-V1
JuZhou 1.0 is an ultra-lightweight 0.387B-parameter text-to-image foundation model designed for fully offline, on-device execution on mobil…
541097active
boona13/image-extender
An open-source Next.js web app for AI image outpainting that extends images in any direction using Gemini models via OpenRouter, with Poiss…
521096active
WangLibo1995/GeoSeg
GeoSeg is an open-source PyTorch-based semantic segmentation toolbox focused on Vision Transformers for remote sensing imagery, featuring t…
321096active
LujiaJin/One-Pot_Multi-Frame_Denoising
Official PyTorch implementation of the One-Pot Multi-frame Denoising (OPD) method published at BMVC 2022 and extended in IJCV. It provides …
601094stable
benhowdle89/grade
Grade is a small JavaScript library that generates complementary gradient backgrounds from the top two dominant colors of supplied images, …
323759maintenance
Eyeline-Labs/Go-with-the-Flow
Official implementation of the CVPR 2025 Oral paper 'Go-with-the-Flow', which controls motion in video diffusion models by replacing i.i.d.…
411093active
welltop-cn/ComfyUI-TeaCache
A ComfyUI plugin integrating TeaCache, a training-free caching method that accelerates diffusion model inference by exploiting output diffe…
351092active
jarun/imgp
imgp is a fast command-line batch image resizer and rotator for JPEG and PNG images, powered by multiprocessing and Pillow-SIMD. It support…
871091active
image-js/image-js
ImageJS is a JavaScript/TypeScript library for image processing and manipulation, offering features like resizing, cropping, filtering, col…
911090stable
mlc-ai/web-stable-diffusion
A project that compiles and runs Stable Diffusion text-to-image models entirely inside web browsers using WebGPU and WebAssembly, with no s…
303721maintenance
SimpleITK/SimpleITK
SimpleITK is a simplified C++ interface to the Insight Toolkit (ITK) for multi-dimensional image analysis, including filtering, segmentatio…
981084stable
Alpha-VLLM/Lumina-mGPT-2.0
Lumina-mGPT 2.0 is a stand-alone decoder-only autoregressive model trained from scratch that unifies a broad range of image generation task…
421084active
openai/glide-text2im
Official codebase for GLIDE, a diffusion-based text-conditional image synthesis model from OpenAI. It provides pretrained models and notebo…
103684maintenance
lolishinshi/imsearch
A Rust-based large-scale similar image search tool that uses feature point matching (ORB features with a FAISS-style index) to find full im…
941074active
yeates/PromptFix
PromptFix is a PyTorch implementation of a diffusion-model-based image restoration model that follows natural language instructions to fix …
241070active
ototadana/sd-face-editor
A Stable Diffusion Web UI extension that detects and regenerates faces in generated images to fix broken faces, change facial expressions, …
231070active
jeffbass/imagezmq
imageZMQ is a set of Python classes that transport OpenCV images between computers using PyZMQ messaging. It enables distributed computer v…
621069stable
RL-VIG/LibFewShot
LibFewShot is a comprehensive PyTorch library for few-shot learning, implementing many fine-tuning, meta-learning, and metric-learning meth…
541069active
tomayac/SVGcode
SVGcode is a Progressive Web App that converts raster images (JPG, PNG, GIF, WebP, AVIF, etc.) into SVG vector graphics. It runs in the bro…
701066active
hujie-frank/SENet
Official Caffe/CUDA implementation of Squeeze-and-Excitation Networks (SENet), channel-attention building blocks for convolutional neural n…
323646maintenance
fossasia/magic-epaper-app
Magic ePaper is an open-source Flutter mobile app for designing content and transferring it to battery-free NFC ePaper badges. It offers dr…
771063active
rust-cv/cv
Rust CV is a mono-repo of pure-Rust computer vision crates aiming to encapsulate capabilities of OpenCV, OpenMVG, and vSLAM frameworks in c…
471063active
AIrjen/OneButtonPrompt
OneButtonPrompt is a Python tool that generates complete Stable Diffusion prompts from scratch with a single click, for beginners and users…
441062active
vijishmadhavan/ArtLine
ArtLine is a deep learning project that converts portrait photos into line art portraits, with a ControlNet-based variant that adjusts styl…
323630maintenance
clovaai/stargan-v2
The official PyTorch implementation of StarGAN v2, a CVPR 2020 paper on diverse image-to-image translation across multiple domains using a …
323617maintenance
sb-ai-lab/EmotiEffLib
EmotiEffLib (formerly HSEmotion) is a lightweight library for facial emotion and engagement recognition in photos and videos, available in …
651057active
yoyo-nb/Thin-Plate-Spline-Motion-Model
The official PyTorch implementation of the CVPR 2022 paper 'Thin-Plate Spline Motion Model for Image Animation'. It animates a source image…
323604maintenance
hnvn/flutter_image_cropper
A Flutter plugin that provides image cropping and rotation on Android, iOS, and Web by wrapping native libraries (uCrop, TOCropViewControll…
751055active
williamyang1991/VToonify
Official PyTorch implementation of VToonify, a SIGGRAPH Asia 2022 framework for controllable high-resolution portrait video style transfer …
323584maintenance
psoho/fast-poster
fastposter is a self-hostable poster/image generation service with a drag-and-drop web editor for composing text, images, QR codes, and ava…
321050active
showlab/MotionDirector
MotionDirector is a research library for customizing text-to-video diffusion models to generate videos with desired motions from a small se…
271050active
qqlu/Entity
EntitySeg is an open-source PyTorch toolbox for open-world, high-quality image segmentation, built on Detectron2. It aggregates multiple re…
321048active
anuragxel/salt
SALT is a Python-based image labeling tool built on Meta AI's Segment Anything Model, providing a barebones GUI for annotating images with …
301048active
addyosmani/squish
Squish is a browser-based batch image compression tool that uses WebAssembly codecs to compress and convert images entirely client-side. It…
221048active
zhbhun/idify
Idify is a browser-based application for creating ID, passport, and visa photos with all processing done locally in the browser. It require…
561047active
kikoso/android-stackblur
An Android library that applies a StackBlur (gaussian-like) blur effect to Bitmaps with configurable radius or gradient. It offers Java, ND…
323566maintenance
nv-tlabs/PiD
PiD is a plug-and-play pixel diffusion decoder from NVIDIA that replaces VAE/RAE decoders, decoding latent representations directly into hi…
561045active
liuyuan-pal/SyncDreamer
SyncDreamer is a synchronized multiview diffusion model that generates multiview-consistent images from a single-view image, released with …
501045active
Tencent-Hunyuan/InstantCharacter
InstantCharacter is a tuning-free framework built on diffusion transformers that generates character-consistent images from a single refere…
291045active
muapi CLI
muapi-cli and its Generative Media Skills provide a schema-driven CLI, skill library, and MCP server that let AI agents (Claude Code, Curso…
931043active
iptag/jimeng-api
A self-hosted API service that reverse-engineers Jimeng AI (China) and Dreamina (international) to expose free AI image and video generatio…
101041active
foolwood/SiamMask
Official PyTorch implementation of SiamMask, a deep learning framework for fast online visual object tracking and video object segmentation…
353547maintenance
Nutlope/blinkshot
BlinkShot is an open-source web application that generates AI images in real time as you type, powered by the Flux Schnell model via Togeth…
671040active
ibireme/YYWebImage
YYWebImage is an asynchronous image loading framework for iOS, part of YYKit, offering remote/local image loading, animated WebP/APNG/GIF d…
233526maintenance
NVlabs/DiffusionNFT
DiffusionNFT is a research library implementing an online reinforcement learning paradigm for diffusion models that optimizes policy direct…
471034active
antimatter15/ocrad.js
Ocrad.js is a pure-JavaScript port of the Ocrad OCR engine, compiled to JavaScript via Emscripten, that converts scanned images of text bac…
323517maintenance
TTPlanetPig/Comfyui_TTP_Toolset
A collection of ComfyUI custom nodes for tiled image processing, including object-aware Smart Tile 2.0 workflows for detail img2img upscali…
661033active
HarborYuan/ovsam
Official PyTorch implementation of Open-Vocabulary SAM (ECCV 2024), a model that unifies SAM's interactive segmentation with CLIP's open-vo…
421033active
nodeca/probe-image-size
A small JavaScript library that reads image dimensions (width, height, type, mime, orientation) from URLs, streams, or buffers without down…
761032active
podgorskiy/ALAE
Official PyTorch implementation of Adversarial Latent Autoencoders (ALAE/StyleALAE), a CVPR 2020 paper combining autoencoders with GAN trai…
323511maintenance
t3mujinpack/t3mujinpack
A collection of film emulation presets for the open-source RAW photo developer Darktable, emulating classic films like Fuji Velvia, Kodak P…
291031active
MemeCrafters/meme-generator
A Python meme generator library that produces various humorous meme images from user-provided avatars and text using built-in templates. It…
731030active
awentzonline/image-analogies
A Python library implementing neural image analogies using VGG16 feature maps with PatchMatch-based matching and blending, based on the 'Im…
233502maintenance
continue-revolution/sd-webui-segment-anything
A Stable Diffusion WebUI extension that integrates Segment Anything and GroundingDINO to generate segmentation masks from clicks or text pr…
303499maintenance
kuprel/min-dalle
min(DALL·E) is a fast, minimal PyTorch port of DALL·E Mini/Mega stripped down for text-to-image inference, with only numpy, requests, pillo…
313494maintenance
lovasoa/dezoomify-rs
dezoomify-rs is a desktop application and CLI tool written in Rust that downloads high-resolution zoomable (tiled) images from websites and…
1001026active
zhyever/PatchFusion
PatchFusion is a CVPR 2024 end-to-end tile-based framework for high-resolution monocular metric depth estimation from single images. It fus…
571026active
JackAILab/ConsistentID
ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using …
521026active
chrisgoringe/cg-use-everywhere
A ComfyUI custom node plugin that provides 'Anything Everywhere' nodes which broadcast data (like MODEL, CLIP, VAE) to all nodes that need …
691024active
DingXiaoH/RepVGG
RepVGG is a PyTorch implementation of the VGG-style ConvNet architecture from the CVPR 2021 paper, achieving over 84% top-1 ImageNet accura…
323478maintenance
steffest/DPaint-js
DPaint.js is a browser-based image editor modeled after Deluxe Paint, with strong support for retro Amiga file formats like IFF ILBM images…
771021active
TencentARC/SEED-Voken
SEED-Voken is a collection of visual tokenizers (Open-MAGVIT2 and IBQ) that convert images and videos into discrete tokens for autoregressi…
481021active
yangxy/PASD
PASD (Pixel-Aware Stable Diffusion) is a Python research codebase implementing an ECCV 2024 method for realistic image super-resolution and…
281021active
JiahuiYu/generative_inpainting
An open-source implementation of DeepFill v1/v2 generative image inpainting models, featuring Contextual Attention (CVPR 2018) and Gated Co…
323466maintenance
tensorlayer/SRGAN
Reference implementation of SRGAN, a generative adversarial network for photo-realistic single image super-resolution, built on TensorLayer…
233466maintenance
ocropus-archive/DUP-ocropy
OCRopy is a collection of Python-based tools for document analysis and OCR, covering binarization, page layout analysis, and text line reco…
103465maintenance
richzhang/colorization
A Python library implementing automatic colorization of grayscale photos using deep neural networks from the ECCV 2016 'Colorful Image Colo…
323461maintenance
zai-org/GLM-Image
GLM-Image is an open-source image generation model combining a 9B autoregressive generator with a 7B diffusion decoder, excelling at text r…
481019active
google/prompt-to-prompt
Google's official implementation of the Prompt-to-Prompt paper, which enables text-driven image editing in Latent Diffusion and Stable Diff…
103456maintenance
evanw/glfx.js
glfx.js is a JavaScript library for applying real-time image effects and photo adjustments in the browser using WebGL. It leverages the GPU…
323451maintenance
Sxela/WarpFusion
WarpFusion is a Stable Diffusion-based video-to-video style transfer tool distributed as a Jupyter/Colab notebook. It applies AI animation …
311016active
eliukblau/pixterm
PIXterm is a Go CLI tool that renders images directly in ANSI terminals using true color escape codes and unicode half-block characters. It…
851015active
Alpha-VLLM/Lumina-DiMOO
Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outpu…
551015active
alessandrofrancesconi/gimp-plugin-bimp
BIMP is a GIMP plugin that applies a set of image manipulations (resize, crop, rotate, watermark, color correction, format conversion, etc.…
231015active
fallenshock/FlowEdit
Official PyTorch implementation of FlowEdit, an ICCV 2025 method for inversion-free, text-based editing of real images using pre-trained fl…
661014active
Gregwar/Image
A PHP library providing a simple object-oriented API for image handling, resizing, cropping, and applying filters, with built-in caching of…
321014stable
addyosmani/bg-remove
A React + Vite web application that removes image backgrounds entirely in the browser using Transformers.js with the RMBG-1.4 model (and op…
221014active
hezarai/hezar
Hezar is an all-in-one Python AI library for the Persian language, covering NLP, speech recognition, OCR, and image captioning through a ta…
781013active
fiji/fiji
Fiji is a batteries-included distribution of ImageJ, bundling thousands of plugins for scientific image processing and analysis into a port…
741013active
Alpha-VLLM/Lumina-Image-2.0
Lumina-Image 2.0 is an open-source 2.6B-parameter text-to-image generation framework built on a unified Next-DiT architecture with a unifie…
581013active
eragonruan/text-detection-ctpn
A TensorFlow implementation of the Connectionist Text Proposal Network (CTPN) for detecting horizontal scene text in images. It includes pr…
233429maintenance
chrissimpkins/Crunch
Crunch is a lossy PNG image optimization tool that combines bit depth, color type, and palette reduction with zopfli DEFLATE compression vi…
233425maintenance
sail-sg/EditAnything
Edit Anything is a Python application for text-guided image editing and generation, combining Segment Anything, ControlNet, BLIP2, and Stab…
333422maintenance
RupertAvery/DiffusionToolkit
Diffusion Toolkit is a Windows desktop application that indexes and views metadata (prompts, models, settings) embedded in AI-generated ima…
651009active
waifu2x (nunif)
waifu2x is an image super-resolution and noise-reduction tool for anime-style art (and photos) using deep convolutional neural networks, or…
783418maintenance
shaoanlu/faceswap-GAN
A Jupyter Notebook-based implementation of face swapping using a denoising autoencoder architecture enhanced with adversarial losses, VGGFa…
323416maintenance
facebookresearch/Mask2Former
Mask2Former is the official PyTorch implementation of the CVPR 2022 paper 'Masked-attention Mask Transformer for Universal Image Segmentati…
103416maintenance
makegirlsmoe/makegirlsmoe_web
The React front-end for MakeGirlsMoe, a web app that generates anime character portraits using a GAN model. It provides the interactive UI …
323414maintenance
MeiGen-AI/PosterCraft
PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight…
481007active
libtv-labs/libtv-skills
A collection of AI agent skill packages that expose LibLib.tv's AIGC capabilities (AI image and video generation) via its OpenAPI. It follo…
471007active
jhfmat/ISP-pipeline-hdrplus
A C/C++ image processing library (Matlib) implementing a fast ISP pipeline with HDR+ multi-frame denoising, super-low-light processing, and…
321005active
Kosinkadink/ComfyUI-Advanced-ControlNet
A set of ComfyUI custom nodes providing advanced ControlNet scheduling, weighting, and masking for Stable Diffusion workflows. It supports …
711004active
dlbeer/quirc
Quirc is a small, dependency-free C library for extracting and decoding QR codes from images, fast enough for realtime video. It handles ro…
421004active
clovaai/CRAFT-pytorch
Official PyTorch implementation of CRAFT (Character Region Awareness for Text Detection), a scene text detector that localizes text by pred…
323398maintenance
AkiraBit/PicSharp
PicSharp is a modern, cross-platform desktop application for high-performance image compression, built with Tauri and TypeScript. It suppor…
591003active
mit-han-lab/efficientvit
A collection of efficient vision foundation models from MIT Han Lab, including EfficientViT backbones for perception, EfficientViT-SAM for …
483354maintenance
eladrich/pixel2style2pixel
Official PyTorch implementation of pixel2style2pixel (pSp), a StyleGAN encoder from CVPR 2021 that maps real images directly into the W+ la…
323350maintenance
tamarott/SinGAN
Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat…
323344maintenance

← prev page 10 / 19 next →