Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: image-processing

1843 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
javierbyte/img2css
A web tool that converts any image into pure CSS, recreating it as a matrix of box-shadows on a single pixel div or as a base64-embedded im…
672498active
yifan123/flow_grpo
Flow-GRPO is the official PyTorch implementation of a NeurIPS 2025 paper that trains flow matching models (e.g., SD3.5, FLUX.1, Qwen-Image,…
562498active
mosch/react-avatar-editor
A React component for cropping, resizing, and rotating uploaded avatar/profile pictures via an intuitive canvas-based UI. It supports round…
582497active
yformer/EfficientSAM
EfficientSAM is an efficient image segmentation model that leverages masked image pretraining to provide a lightweight alternative to Meta'…
272491active
wasabeef/glide-transformations
An Android library that provides a collection of bitmap transformations (crop, blur, grayscale, rounded corners, masks, and GPU-based filte…
329883maintenance
ipazc/mtcnn
A Python library implementing the MTCNN (Multitask Cascaded Convolutional Networks) algorithm for face detection and facial landmark alignm…
232485stable
twistedfall/opencv-rust
Rust bindings for the OpenCV computer vision library, generated automatically via Clang. It exposes OpenCV 3.4 (deprecated), 4.x, and 5.x A…
752483active
xuebinqin/U-2-Net
Official PyTorch implementation of U^2-Net, a nested U-structure deep network for salient object detection, published in Pattern Recognitio…
329853maintenance
alexjc/neural-doodle
A Python implementation of Semantic Style Transfer (Champandard, 2016) based on the Neural Patches algorithm. It turns rough doodles into r…
109852maintenance
DMarby/picsum-photos
Lorem Picsum is a self-hostable service that serves stylish placeholder photos, like Lorem Ipsum but for images. Written in Go, it provides…
622478active
GaParmar/img2img-turbo
A research library implementing one-step image-to-image translation models (CycleGAN-Turbo and pix2pix-turbo) built on SD-Turbo diffusion m…
412476active
kevmo314/magic-copy
Magic Copy is a browser extension (Chrome, Firefox, and Figma) that uses Meta's Segment Anything Model to segment a foreground object from …
202458active
ruslanskorb/RSKImageCropper
RSKImageCropper is an Objective-C library providing an image/photo crop view controller for iOS, styled like the Contacts app crop UI, with…
692449stable
unjs/ipx
IPX is a high-performance, secure image optimizer powered by sharp and svgo that serves images in any size, format, and quality via URL mod…
942447active
antirez/h3.c
A native C inference engine for the MiniMax H3 model on Apple Silicon, using Metal for GPU acceleration. It generates video (and audio) fro…
562443active
Vibrant-Colors/node-vibrant
node-vibrant is a TypeScript library that extracts prominent color palettes (vibrant, muted, light/dark variants) from images. It provides …
732442active
wolny/pytorch-3dunet
A PyTorch implementation of 3D U-Net and its variants (residual, squeeze-and-excitation) for volumetric semantic segmentation, with 2D U-Ne…
632416active
glidea/banana-prompt-quicker
A Chrome extension that lets users quickly insert curated and custom prompts into Google AI Studio, Gemini, and any website input box via r…
612405active
PyWavelets/pywt
PyWavelets is an open-source Python library for wavelet transforms, offering discrete, stationary, wavelet packet, and continuous wavelet t…
672396stable
alexvasilkov/GestureViews
An Android library providing ImageView and FrameLayout widgets with built-in gesture control (pan, zoom, fling, rotation, double tap) and s…
802386stable
ossappscollective/OSS-DocumentScanner
OSS Document Scanner is a free, open-source, privacy-focused mobile app for scanning documents with automatic edge detection, editing, OCR,…
912385active
chillerlan/php-qrcode
A PHP library for generating Model 2 QR Codes (versions 1-40, all ECC levels, mixed encoding modes) with extensible output modules for rast…
812384active
webmproject/libwebp
libwebp is the reference C library for encoding and decoding images in the WebP format, maintained by the WebM project. It also ships comma…
772380stable
pydn/ComfyUI-to-Python-Extension
A ComfyUI extension and CLI tool that translates ComfyUI node-graph workflows into executable Python scripts. It lets users export workflow…
792375active
pixpark/gpupixel
GPUPixel is a high-performance, cross-platform real-time image and video filter library written in C++11 and built on OpenGL/ES. It provide…
882371active
sumelabs/clawra
Clawra is an OpenClaw skill/plugin that gives an AI agent the ability to generate consistent selfies via xAI Grok Imagine on fal.ai and sen…
452352active
AcademySoftwareFoundation/OpenImageIO
OpenImageIO is a C++ library and toolset for reading, writing, and processing images in nearly any file format through a format-agnostic pl…
982349stable
matthewwithanm/django-imagekit
django-imagekit is a Django app for automated image processing, generating derived images like thumbnails or cropped versions from source i…
782349active
ComflowySpace
ComflowySpace is an open-source desktop application that wraps ComfyUI/Stable Diffusion into a friendlier, app-like interface for generatin…
172345active
heshengtao/comfyui_LLM_party
A ComfyUI plugin providing a comprehensive set of nodes for building LLM agent workflows, including MCP server support, RAG/GraphRAG, TTS, …
652343active
thoas/picfit
Picfit is a reusable Go HTTP server that resizes, crops, and generates thumbnails of images on the fly, acting as a proxy over storage back…
722341active
lvandeve/lodepng
LodePNG is a standalone PNG encoder and decoder written in C and C++ with no external dependencies, distributed as just two source files. I…
702340stable
MouseLand/cellpose
Cellpose is a generalist deep learning algorithm for cellular and nucleus segmentation in microscopy images, with human-in-the-loop capabil…
862331active
NextLevel/NextLevel
NextLevel is a Swift camera capture library for iOS built on AVFoundation, providing photo and video capture, multi-clip recording, ARKit i…
712331active
wasabeef/android-gpuimage
An Android library for applying GPU-accelerated image and video filters using OpenGL ES 2.0, ported from the iOS GPUImage framework. It pro…
329155maintenance
Cadene/pretrained-models.pytorch
A Python library providing pretrained ConvNet models (ResNet, ResNeXt, InceptionV4, Xception, NASNet, SENet, DPN, etc.) for PyTorch behind …
329099maintenance
Stability-AI/StableStudio
StableStudio is Stability AI's open-source, web-based variant of DreamStudio for creating and editing AI-generated images. It features a pl…
209048maintenance
Brooooooklyn/canvas
A high-performance Node.js canvas implementation backed by Google's Skia graphics library, built with Rust and Node-API (napi-rs). It provi…
952306active
tannerhelland/PhotoDemon
PhotoDemon is a free, open-source, portable photo editor for Windows built in Visual Basic 6. It offers pro-grade tools like layers, RAW su…
742304active
strukturag/libheif
libheif is a C++ library that decodes and encodes HEIF/HEIC and AVIF image files, plus HEIF containers using VVC, AVC, JPEG, and JPEG-2000 …
992302active
Achno/gowall
Gowall is a Go-based CLI tool that converts images (especially wallpapers) to custom color schemes and offers a broad suite of image proces…
732301active
emgucv/emgucv
Emgu CV is a cross-platform .NET wrapper for the OpenCV image processing library, allowing OpenCV functions to be called from .NET-compatib…
742294active
mflux-community/mflux
MFLUX is a native MLX implementation of state-of-the-art generative image and video models (Flux, Qwen-Image, Z-Image, and others), ported …
902291active
adieyal/sd-dynamic-prompts
An extension for AUTOMATIC1111's stable-diffusion-webui that adds a template language for random and combinatorial prompt generation using …
322287active
andrewssobral/bgslibrary
BGSLibrary is a C++ framework for background subtraction in video, offering 43 algorithms for foreground-background separation built on Ope…
612277active
onevcat/APNGKit
APNGKit is a high-performance Swift framework for loading, decoding, and displaying Animated PNG (APNG) images on iOS and macOS. It offers …
802276active
siwangqishiq/ImageEditor-Android
An open-source Android image editing control (Java) supporting stickers, filters, rotation, cropping, text overlays, doodles, skin smoothin…
492271active
OlafenwaMoses/ImageAI
ImageAI is a Python library that lets developers add computer vision capabilities like image classification, object detection, and video ob…
238877maintenance
aigc-apps/EasyAnimate
EasyAnimate is an end-to-end Python pipeline for high-resolution, long video and image generation based on transformer diffusion (DiT) mode…
202270active
yawiii/ComfyUI-Prompt-Assistant
A ComfyUI plugin that provides an all-in-one prompt assistant, connecting to cloud LLM/VLM APIs (Zhipu, SiliconFlow, Gemini, Baidu) and loc…
702269active
ermig1979/Simd
Simd Library is a free open-source C++ image processing and machine learning library with a C API and Python wrapper. Its algorithms are ha…
982265active
stepfun-ai/Step1X-Edit
Step1X-Edit is an open-source state-of-the-art instruction-based image editing model from StepFun, designed to rival closed-source editors …
552256active
THU-MIG/yoloe
YOLOE is the official PyTorch implementation of an open-vocabulary object detection and segmentation model presented at ICCV 2025. It unifi…
322256active
Alpha-VLLM/Lumina-T2X
Lumina-T2X is a unified framework for text-to-any-modality generation built on flow-based large diffusion transformers. It supports generat…
282250active
mrousavy/react-native-blurhash
A React Native library that renders BlurHash strings as colorful blurred image placeholders while content loads. It provides a native compo…
652237stable
facebookresearch/DiT
Official PyTorch implementation of Diffusion Transformers (DiT) from the paper 'Scalable Diffusion Models with Transformers', including mod…
108689maintenance
NVlabs/MambaVision
MambaVision is NVIDIA's official PyTorch implementation of a hybrid Mamba-Transformer vision backbone, published at CVPR 2025. It provides …
492224active
photopea/UPNG.js
UPNG.js is a small, fast JavaScript library for encoding and decoding PNG and APNG images, serving as the main PNG engine for the Photopea …
232215stable
aigc-apps/VideoX-Fun
VideoX-Fun is a Python-based video generation pipeline built on Diffusion Transformer models (CogVideoX-Fun, Wan-Fun) that generates videos…
672210active
MVIG-SJTU/AlphaPose
AlphaPose is an open-source real-time multi-person full-body pose estimation and tracking system built on PyTorch. It detects human keypoin…
328596maintenance
kijai/ComfyUI-LivePortraitKJ
ComfyUI custom nodes that integrate the LivePortrait face animation and retargeting model, supporting image-to-video, video-to-video, and n…
232200active
pydicom/pydicom
pydicom is a pure Python library for reading, modifying, and writing DICOM medical imaging files and File-sets in a pythonic way. It option…
892197stable
LigphiDonk/academic-figure-generator
A self-hosted AI-powered platform that generates high-quality academic paper figures: users upload a paper (PDF/DOCX/TXT), Claude analyzes …
622190active
TimmyOVO/deepseek-ocr.rs
A Rust implementation of the DeepSeek-OCR inference stack with multiple OCR/VLM backends (DeepSeek-OCR, PaddleOCR-VL, DotsOCR), DSQ quantiz…
602182active
MetalPetal/MetalPetal
MetalPetal is a GPU-accelerated image and video processing framework built on Apple's Metal API. It provides an image/filter/render pipelin…
232179active
sirfz/tesserocr
A Python wrapper around the tesseract-ocr C++ API built with Cython for optical character recognition. It is Pillow-friendly, works with im…
932171active
yatengLG/ISAT_with_segment_anything
ISAT_with_segment_anything is an interactive semi-automatic image annotation tool built on the Segment Anything Model family (SAM, SAM2, SA…
842166active
Scholar01/sd-webui-mov2mov
A Mov2mov plugin for the Automatic1111 stable-diffusion-webui that applies Stable Diffusion to videos by processing frames and repackaging …
312166active
AOMediaCodec/libavif
libavif is a portable C library for encoding and decoding AVIF (AV1 Image File Format) images, supporting all AV1 YUV formats and bit depth…
902163active
BishopFox/unredacter
Unredacter is an Electron-based desktop tool that demonstrates how pixelated redactions in images can be reversed by brute-force guessing t…
328382maintenance
zombieyang/sd-ppp
SD-PPP is an open-source Photoshop plugin (built on Adobe UXP) that integrates AI image generation platforms like ComfyUI, Replicate, and R…
662151active
jd-opensource/JoyAI-Image
JoyAI-Image is a unified multimodal foundation model for image understanding, text-to-image generation, and instruction-guided image editin…
582148active
rupeshs/fastsdcpu
FastSD CPU is a Python application that runs Stable Diffusion image generation quickly on CPUs and Intel AI PCs using Latent Consistency Mo…
782143active
haraldk/TwelveMonkeys
TwelveMonkeys ImageIO is a collection of plugins and extensions for Java's javax.imageio framework, adding read/write support for many imag…
892142stable
xinsir6/ControlNetPlus
ControlNet++ is an all-in-one ControlNet model and architecture supporting 10+ control types for text-to-image generation and image editing…
232138active
espressif/esp-who
ESP-WHO is an image processing development platform from Espressif providing face detection, face recognition, pedestrian detection, and QR…
672133active
lukemelas/EfficientNet-PyTorch
A PyTorch implementation of the EfficientNet convolutional neural network family with pretrained ImageNet weights. It provides a simple pip…
238222maintenance
autonomousvision/sdfstudio
SDFStudio is a unified and modular framework for neural implicit surface reconstruction built on top of nerfstudio. It provides unified imp…
312120active
Spu7Nix/obamify
obamify is a Rust desktop and web application that transforms any image into a morphing animation of Barack Obama using pixel assignment al…
522111active
River-Zhang/ICEdit
ICEdit (In-Context Edit) is a research framework for instruction-based image editing built on large-scale Diffusion Transformers, using a L…
452102active
path/FastImageCache
FastImageCache is an Objective-C iOS library for persistently storing and retrieving images at high speed, designed to keep scrolling smoot…
238059maintenance
PaddlePaddle/PaddleGAN
PaddleGAN is a Python library providing high-performance implementations of classic and state-of-the-art Generative Adversarial Networks bu…
238048maintenance
tumuyan/RealSR-NCNN-Android
An Android application for image super-resolution and upscaling built on NCNN and MNN inference engines, bundling models like RealSR, Real-…
842087active
1038lab/ComfyUI-RMBG
A ComfyUI custom node package for advanced image background removal and segmentation of objects, faces, clothing, and fashion elements. It …
662086active
ali-vilab/In-Context-LoRA
Official repository for In-Context LoRA (IC-LoRA), a framework for adapting Diffusion Transformers to diverse visual generation tasks via L…
222083active
alex-damian/pulse
PULSE is a Python research implementation of a CVPR 2020 paper that upscales low-resolution face photos by searching the latent space of a …
328023maintenance
ascorbic/unpic-img
unpic-img is a cross-framework responsive image component library for React, Vue, Svelte, Astro, Angular, SolidJS, and more. It generates c…
842075active
DanBloomberg/leptonica
Leptonica is an open-source C library providing a broad set of image processing and image analysis operations, with a focus on document ima…
742074stable
esimov/triangle
A Go CLI tool and library that converts images into abstract computer-generated art using Delaunay triangulation. It blurs, grayscales, and…
232069active
oxylabs/how-to-scrape-google-images
A Python-based command-line tool that scrapes Google Images search results, including reverse image search based on a provided image URL. I…
642055active
shanglianlm0525/PyTorch-Networks
A collection of PyTorch implementations of classic and modern CNN architectures, covering classification, detection, segmentation, face, an…
532055active
visomaster/VisoMaster
VisoMaster is a Python-based desktop application for AI-powered face swapping and face editing in images and videos. It supports multiple s…
272052active
discord/lilliput
A Go library for resizing and transcoding images, backed by mature C libraries (JPEG, PNG, WebP, AVIF, animated GIF) via cgo. It minimizes …
742051active
Sygil-Dev/sygil-webui
A browser-based web UI for generating images with Stable Diffusion, built in Python with Gradio and Streamlit frontends. It supports text-t…
747870maintenance
storytold/artcraft
ArtCraft is an open-source desktop application for interactive AI image and video creation, described as 'the IDE for artists'. It provides…
942044active
PRIS-CV/DemoFusion
DemoFusion is a CVPR 2024 framework that extends open-source latent diffusion models like SDXL to generate high-resolution images without a…
482041stable
deep-floyd/IF
DeepFloyd IF is an open-source text-to-image model library implementing a cascaded pixel diffusion architecture with a frozen T5 text encod…
227804maintenance
serengil/retinaface
RetinaFace is a Python library for deep learning based face detection, built on TensorFlow and derived from the insightface project's Retin…
612027active
renzhezhilu/webp2jpg-online
A browser-based, pure front-end image format converter that converts between jpeg, png, gif, webp, svg, ico, bmp, psd, heic and more withou…
232024stable
TianZerL/Anime4KCPP
Anime4KCPP is a high-performance anime image and video upscaler built on CNN-based algorithms, written in C++. It ships as a library plus V…
782022active

← prev page 5 / 19 next →