Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: image-processing

1843 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
kean/Nuke
Nuke is a Swift image loading and caching framework for Apple platforms, providing an ImagePipeline for fetching, processing, and displayin…
998656stable
lllyasviel/IC-Light
IC-Light is a Python tool for manipulating the illumination of images using diffusion models, offering text-conditioned and background-cond…
278508active
Nutlope/logocreator
An open-source AI logo generator web app that creates brand-ready logos using FLUX models on Together AI, with logo editing, style presets,…
648435active
photopea/photopea
Photopea is a free online image editor for raster and vector graphics that runs entirely in the browser, supporting PSD, AI, Sketch, and do…
458405active
CASIA-LMC-Lab/FastSAM
FastSAM is a CNN-based Segment Anything Model trained on only 2% of the SA-1B dataset, achieving comparable segmentation performance to SAM…
198401active
XPixelGroup/BasicSR
BasicSR is an open-source PyTorch toolbox for image and video restoration tasks such as super-resolution, denoising, deblurring, and JPEG a…
238367stable
bytedeco/javacv
JavaCV is a Java library that wraps OpenCV, FFmpeg, and other computer vision and multimedia libraries via JavaCPP Presets, with utility cl…
868335active
vietnh1009/ASCII-generator
A Python tool that converts images and videos into ASCII art, outputting text files or image/video files in grayscale or color. It supports…
328318stable
exadel-inc/CompreFace
Exadel CompreFace is a free, open-source face recognition system that provides REST APIs for face recognition, verification, detection, lan…
238273stable
QwenLM/Qwen-Image
Qwen-Image is a 20B MMDiT image generation foundation model from the Qwen team, with strong complex text rendering (especially Chinese) and…
488265active
lltcggie/waifu2x-caffe
A Windows GUI/CLI application that reimplements the waifu2x image upscaling and noise-reduction tool using the Caffe deep learning framewor…
568221stable
carson-katri/dream-textures
A Blender add-on that integrates Stable Diffusion for generating textures, concept art, and background assets directly inside Blender. It s…
238195active
brycedrennan/imaginAIry
A Python library and CLI tool (imaginairy/aimg) for generating images and videos with Stable Diffusion and Stable Video Diffusion models. I…
638179active
zumerlab/snapdom
SnapDOM is a high-performance, dependency-free browser library that captures DOM elements as self-contained SVG representations and exports…
868040active
SixLabors/ImageSharp
ImageSharp is a fully managed, high-performance 2D graphics and image processing library for .NET 8+. It provides loading, resizing, format…
958034stable
nadermx/backgroundremover
A free, open-source command line tool that removes backgrounds from images and videos using the U2Net AI model built on PyTorch. It is inst…
938020active
bingoogolapple/BGAQRCode-Android
An Android library for scanning and generating QR codes and barcodes, offering both ZXing and ZBar engines behind customizable scan views. …
648008stable
PixiEditor/PixiEditor
PixiEditor is a free, open-source, cross-platform desktop 2D graphics editor built in C# with Avalonia. It combines pixel art, painting, an…
997999active
MochiDiffusion/MochiDiffusion
A native macOS app built with SwiftUI for running Stable Diffusion and FLUX.2 Klein image generation locally on Apple Silicon Macs. It uses…
787926active
TheLastBen/fast-stable-diffusion
A collection of Google Colab notebooks for quickly running Stable Diffusion UIs (AUTOMATIC1111, ComfyUI) and training DreamBooth models for…
577910active
TencentARC/GFPGAN
GFPGAN is a Python library built on PyTorch that restores and enhances real-world degraded face photos using GAN-based priors. It provides …
2337657maintenance
geekyutao/Inpaint-Anything
Inpaint Anything combines Segment Anything (SAM) with inpainting models like LaMa and Stable Diffusion to remove, fill, or replace objects …
657703active
lllyasviel/Omost
Omost is a Python application that converts LLM coding capability into image composition by having pretrained LLMs (based on Llama3 and Phi…
247610active
RapidAI/RapidOCR
RapidOCR is an open-source, multi-language OCR toolkit that performs text detection and recognition using models converted to run on ONNX R…
977599active
1adrianb/face-alignment
A Python library built on PyTorch that detects 2D and 3D facial landmarks in images using the FAN deep learning face alignment network. It …
707538active
phoboslab/qoi
QOI is the 'Quite OK Image Format', a fast, lossless image compression format with a single-file MIT-licensed C/C++ reference implementatio…
707522stable
hybridgroup/gocv
GoCV is a Go language binding for the OpenCV 4 computer vision library, supporting Linux, macOS, Windows, and Docker. It includes support f…
767491active
EutropicAI/Final2x
Final2x is a cross-platform desktop application for image super-resolution (upscaling) built with Electron, Vue3, and a PyTorch-based Pytho…
747323active
vladmandic/sdnext
SD.Next is an open-source, self-hosted WebUI server application for AI generative image and video creation built on Stable Diffusion and Di…
767320active
AbdullahAlfaraj/Auto-Photoshop-StableDiffusion-Plugin
A Photoshop plugin (built on Adobe UXP) that lets users generate Stable Diffusion images directly inside Photoshop, using Automatic1111 Web…
227288active
imgly/background-removal-js
An npm package (browser and Node.js variants) that removes image backgrounds using ONNX-based image segmentation/matting models running ent…
437287active
liuliu/ccv
ccv is a modern, minimalist computer vision library written in C/C++ with an application-driven set of state-of-the-art algorithms includin…
777243active
civitai/civitai
Civitai is a web platform for sharing, discovering, and discussing Stable Diffusion models, textual inversions, LoRAs, VAEs, and other gene…
777238active
zetbaitsu/Compressor
Compressor is a lightweight Android image compression library written in Kotlin. It lets developers shrink large photos into smaller files …
547227stable
bubkoo/html-to-image
A TypeScript library that generates images (PNG, JPEG, SVG, blob, canvas, or pixel data) from DOM nodes using HTML5 canvas and SVG foreignO…
727222stable
kohya-ss/sd-scripts
A collection of Python training, generation, and utility scripts for Stable Diffusion and other image generation models, most widely used f…
897210active
PeterL1n/BackgroundMattingV2
Official PyTorch implementation of the CVPR 2021 paper 'Real-Time High-Resolution Background Matting'. It produces state-of-the-art alpha m…
237189stable
ControlNet
ControlNet is a neural network architecture that adds conditional control (edges, poses, depth, etc.) to pretrained text-to-image diffusion…
3134091maintenance
zxing/zxing
ZXing ('Zebra Crossing') is an open-source, multi-format 1D/2D barcode image processing library implemented in Java, with ports to other la…
7334077maintenance
xushengfeng/eSearch
eSearch is a cross-platform desktop application (Electron) combining screenshot capture, offline OCR based on PaddleOCR, screen search, tra…
987036active
latentcat/qrbtf
QRBTF is an AI and parametric QR code generator that creates stylized, scannable QR codes, available as a hosted web app at qrbtf.com with …
306996active
mapbox/pixelmatch
A tiny, dependency-free JavaScript library for pixel-level image comparison, originally built for comparing screenshots in tests. It works …
896931stable
sczhou/ProPainter
ProPainter is a PyTorch-based video inpainting model from ICCV 2023 that combines dual-domain propagation with a mask-guided sparse video T…
226916stable
leejet/stable-diffusion.cpp
A pure C/C++ inference engine for diffusion models (Stable Diffusion, FLUX, Wan, Qwen Image, Z-Image, and more) built on ggml in the style …
916846active
Dooy/chatgpt-web-midjourney-proxy
A unified web/desktop UI for ChatGPT plus AI image, music, and video generation services like Midjourney, Suno, Luma, Runway, and Flux. It …
766785active
visioncortex/vtracer
VTracer is an open-source tool and library that converts raster images (JPG, PNG) into vector graphics (SVG), handling both colored images …
986745active
tencent-ailab/IP-Adapter
IP-Adapter is a lightweight (22M parameter) adapter that adds image prompt capability to pretrained text-to-image diffusion models like Sta…
286677stable
halide/Halide
Halide is an embedded DSL (in C++ and Python) for writing high-performance, data-parallel image and array processing pipelines. It separate…
706590stable
11cafe/jaaz
Jaaz is an open-source multimodal creative assistant application that serves as a privacy-focused, locally usable alternative to Canva and …
536589active
scikit-image/scikit-image
scikit-image is a Python library providing a collection of peer-reviewed image processing algorithms built on NumPy and SciPy. It offers ro…
826577stable
iperov/DeepFaceLive
DeepFaceLive is a real-time face-swap application for PC streaming and video calls, using trained face models (DFM) applied to webcam or vi…
1031011maintenance
HVision-NKU/StoryDiffusion
StoryDiffusion is the official implementation of a NeurIPS 2024 Spotlight paper introducing Consistent Self-Attention for character-consist…
246452active
sz3/libcimbar
libcimbar is an optimized C++ implementation of the cimbar (color icon matrix) high-density 2D barcode format for air-gapped data transfer …
956426active
OpenDroneMap/ODM
OpenDroneMap (ODM) is an open source command line toolkit that processes aerial drone, balloon, or kite imagery into classified point cloud…
906417active
madmaze/pytesseract
Python-tesseract is a Python wrapper for Google's Tesseract-OCR engine that recognizes and extracts text embedded in images. It supports al…
646383stable
GNOME/gimp
GIMP (GNU Image Manipulation Program) is a free, open-source raster graphics editor for photo retouching, image composition, and authoring.…
776378active
mindee/doctr
docTR is a Python OCR library that extracts text from documents and images using a two-stage deep learning approach: text detection followe…
906315active
Lymphatus/caesium-image-compressor
Caesium Image Compressor is a free, open-source desktop application for compressing JPG, PNG, WebP, and TIFF images while preserving visual…
686266active
szad670401/HyperLPR
HyperLPR3 is a high-performance open-source framework for recognizing Chinese license plates, built with deep learning and available as a P…
276255active
RangiLyu/nanodet
NanoDet-Plus is a super fast, lightweight anchor-free object detection model implemented in PyTorch, with model sizes as small as 980KB (IN…
236252stable
Akegarasu/lora-scripts
SD-Trainer is a GUI application and set of scripts for training LoRA and Dreambooth fine-tunes of Stable Diffusion diffusion models, wrappi…
666110active
h2non/imaginary
Imaginary is a fast HTTP microservice written in Go for high-level image processing, backed by bimg and libvips. It exposes image operation…
466079stable
shimat/opencvsharp
OpenCvSharp is a cross-platform .NET wrapper for the OpenCV computer vision library, published as NuGet packages with bundled native binari…
986072active
basketikun/chatgpt2api
A self-hosted reverse-engineered implementation of ChatGPT's official web interfaces, exposing OpenAI-compatible API endpoints for text gen…
776012active
Doubiiu/ToonCrafter
ToonCrafter is a generative model that interpolates two cartoon images into a short animation by leveraging pre-trained image-to-video diff…
296003stable
chaiNNer-org/chaiNNer
chaiNNer is a free, open-source, node-based desktop application for building image processing pipelines by connecting nodes on a canvas. Or…
815994active
lxfater/inpaint-web
A free, open-source browser-based tool for image inpainting (object removal) and image upscaling (super-resolution), built with WebGPU and …
535912active
image-rs/image
A Rust library providing encoding and decoding for many common image formats (PNG, JPEG, GIF, WebP, AVIF, TIFF, and more) plus basic image …
775862stable
ChaoningZhang/MobileSAM
MobileSAM is the official implementation of a lightweight version of Meta's Segment Anything Model (SAM), replacing the heavyweight image e…
655858stable
PaddlePaddle/PaddleClas
PaddleClas is a Python library and toolkit for image classification, recognition, and retrieval built on the PaddlePaddle deep learning fra…
665838active
fengyuanchen/compressorjs
Compressor.js is a JavaScript library that compresses images in the browser using the native HTMLCanvasElement.toBlob() method. It is typic…
815764stable
disintegration/imaging
Imaging is a Go library providing basic image processing functions such as resize, rotate, crop, and brightness/contrast adjustments. It wo…
235756stable
kornelski/pngquant
pngquant is a command-line lossy PNG compressor that converts images to efficient 8-bit palette PNGs with alpha support, often reducing fil…
725742stable
imagemin/imagemin
A Node.js library that minifies images (JPEG, PNG, GIF, SVG) through a pluggable architecture, accepting file globs or buffers and returnin…
275722active
mozilla/mozjpeg
MozJPEG is an improved JPEG encoder built as a patch on libjpeg-turbo, producing smaller files at higher visual quality while remaining ful…
355716stable
zhongerxin/Cowart
Cowart is a Codex plugin that provides a native infinite canvas widget built on tldraw for brainstorming, annotating images, and AI-driven …
585683active
OpenSenseNova/SenseNova-U1
SenseNova-U is a series of open-weight unified multimodal models (e.g., SenseNova-U1.5-8B-MoT) built on the NEO-unify architecture that com…
595668active
idealo/imagededup
imagededup is a Python library for finding exact and near-duplicate images in a collection using perceptual hashing algorithms (PHash, DHas…
485666stable
Fanghua-Yu/SUPIR
SUPIR is a Python-based photo-realistic image restoration system built on SDXL diffusion priors and LLaVA captioning, presented at CVPR 202…
365649active
basketikun/infinite-canvas
An open-source infinite canvas workbench for AI-driven visual creation, combining canvas orchestration, AI image generation, reference-imag…
805629active
ATH-MaaS/ComfyUI-Copilot
ComfyUI-Copilot is an AI-powered custom node for ComfyUI that acts as an intelligent assistant for building, debugging, and optimizing imag…
535489active
coobird/thumbnailator
Thumbnailator is a Java library for generating high-quality image thumbnails with a simple fluent API. It wraps the Java Image I/O and Java…
635425active
mayocream/koharu
Koharu is a local-first desktop application that automates manga translation using machine learning, combining text/bubble detection, OCR, …
825410active
GargantuaX/gemini-watermark-remover
A 100% client-side tool that removes Gemini AI watermarks from images and videos using a mathematically precise Reverse Alpha Blending algo…
825409active
Hillobar/Rope
Rope is a GUI-focused desktop application for face swapping that implements the insightface inswapper_128 model. It offers batch swapping, …
795367active
novicezk/midjourney-proxy
A Java-based proxy service that wraps MidJourney's Discord channel into a REST API, enabling programmatic AI image generation. It supports …
455347active
wiltodelta/remove-ai-watermarks
A Python library and CLI for removing AI watermarks and provenance metadata from images and video the user generated themselves. It handles…
775278active
timesler/facenet-pytorch
A PyTorch library providing pretrained face detection (MTCNN) and facial recognition (Inception ResNet V1) models, ported from the TensorFl…
425162stable
yisol/IDM-VTON
Official implementation of IDM-VTON, an ECCV 2024 paper that improves diffusion models for high-fidelity virtual try-on, swapping garments …
305156active
aigc-apps/sd-webui-EasyPhoto
EasyPhoto is a Stable Diffusion WebUI plugin for generating AI portraits by training a personal 'digital doppelganger' from 5-20 user photo…
285155active
philz1337x/clarity-upscaler
Clarity AI is a free and open-source AI image upscaler and enhancer built on Stable Diffusion, positioned as an alternative to Magnific. It…
305115active
ai-dawang/PlugNPlay-Modules
A curated collection of plug-and-play deep learning modules (convolutions, attention mechanisms, downsampling, and feature fusion blocks) i…
385105active
dmMaze/BallonsTranslator
A desktop GUI application that uses deep learning to automatically translate comics and manga, combining text detection, OCR, inpainting, a…
995065active
RandyGaul/cute_headers
A collection of cross-platform, single-file C/C++ header-only libraries with no dependencies, primarily aimed at game development. It inclu…
755052active
soruly/trace.moe
trace.moe is an anime scene search engine that identifies which anime, episode, and exact timestamp a screenshot comes from. This repositor…
765022active
pollinations/pollinations
Pollinations is an open-source generative AI platform offering free APIs for text, image, and vision model inference without requiring API …
774996active
pkuliyi2015/multidiffusion-upscaler-for-automatic1111
An Automatic1111 Stable Diffusion WebUI extension that enables generating and upscaling ultra-large images (2K+) on GPUs with limited VRAM …
304994active
TheJoeFin/Text-Grab
Text Grab is a Windows OCR utility that extracts text from anywhere on screen — screenshots, images, videos, PDFs, or app windows — entirel…
944978active
TimOliver/TOCropViewController
TOCropViewController is an open-source iOS view controller for cropping UIImage objects and performing basic rotations, designed to feel li…
924942stable
wuyoscar/GPT-Image2-Skill
A curated prompt gallery and library for OpenAI's GPT Image 2 model, packaged as an agentic skill and Python CLI for image generation and e…
744923active

← prev page 2 / 19 next →