Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: image-processing

4273 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
SatDump/SatDump
SatDump is a general-purpose satellite data processing application that receives, records, demodulates, and decodes signals from weather sa…
742088active
tumuyan/RealSR-NCNN-Android
An Android application for image super-resolution and upscaling built on NCNN and MNN inference engines, bundling models like RealSR, Real-…
842087active
WebAV-Tech/WebAV
WebAV is a TypeScript SDK for creating and editing audio/video files entirely in the browser, built on the WebCodecs API. It provides compo…
582087active
1038lab/ComfyUI-RMBG
A ComfyUI custom node package for advanced image background removal and segmentation of objects, faces, clothing, and fashion elements. It …
662086active
ali-vilab/In-Context-LoRA
Official repository for In-Context LoRA (IC-LoRA), a framework for adapting Diffusion Transformers to diverse visual generation tasks via L…
222083active
alex-damian/pulse
PULSE is a Python research implementation of a CVPR 2020 paper that upscales low-resolution face photos by searching the latent space of a …
328023maintenance
vapoursynth/vapoursynth
VapourSynth is a video processing framework with a C++ core library and a Python module for writing video processing scripts. It supports m…
942079active
hilongjw/vue-lazyload
A lightweight Vue.js plugin that lazy-loads images and components via a v-lazy directive, with loading/error placeholders and CSS state cla…
237993maintenance
ascorbic/unpic-img
unpic-img is a cross-framework responsive image component library for React, Vue, Svelte, Astro, Angular, SolidJS, and more. It generates c…
842075active
DanBloomberg/leptonica
Leptonica is an open-source C library providing a broad set of image processing and image analysis operations, with a focus on document ima…
742074stable
MiniMax-AI/cli
The official CLI for the MiniMax AI Platform, written in TypeScript, that generates text, images, video, speech, and music from the termina…
812073active
GPUOpen-Effects/FidelityFX-FSR2
AMD FidelityFX Super Resolution 2 (FSR 2) is an open-source, high-quality temporal upscaling solution that reconstructs high-resolution fra…
232073stable
esimov/triangle
A Go CLI tool and library that converts images into abstract computer-generated art using Delaunay triangulation. It blurs, grayscales, and…
232069active
facebookresearch/ConvNeXt-V2
Official PyTorch implementation of ConvNeXt V2, a family of pure convolutional neural network models co-designed with a fully convolutional…
102069stable
Flipboard/FLAnimatedImage
FLAnimatedImage is a performant animated GIF engine for iOS that plays GIFs with desktop-browser-like speed while handling variable frame d…
237949maintenance
dai-shi/excalidraw-animate
A web-based tool and npm package that converts Excalidraw drawings into animations, with configurable per-element animation order and durat…
722066active
3rd/image.nvim
A Neovim plugin written in Lua that adds image display support inside the editor using Kitty's Graphics Protocol, ueberzugpp, or Sixel back…
812064active
rlxone/Equinox
Equinox is a free, open-source native macOS app for creating dynamic wallpapers such as Dynamic Desktop and Light & Dark Desktop. Users cho…
752054active
marcoslucianops/DeepStream-Yolo
A collection of configuration files, parsers, and conversion utilities for running YOLO-family object detection models on NVIDIA DeepStream…
612054active
visomaster/VisoMaster
VisoMaster is a Python-based desktop application for AI-powered face swapping and face editing in images and videos. It supports multiple s…
272052active
extesy/hoverzoom
Hover Zoom+ is an open-source browser extension that enlarges images and videos to full size when you hover your mouse over them on support…
932051active
discord/lilliput
A Go library for resizing and transcoding images, backed by mature C libraries (JPEG, PNG, WebP, AVIF, animated GIF) via cgo. It minimizes …
742051active
Sygil-Dev/sygil-webui
A browser-based web UI for generating images with Stable Diffusion, built in Python with Gradio and Streamlit frontends. It supports text-t…
747870maintenance
alganzory/HaramBlur
HaramBlur is a browser extension that automatically detects and blurs inappropriate images and videos on web pages using on-device machine …
182048active
storytold/artcraft
ArtCraft is an open-source desktop application for interactive AI image and video creation, described as 'the IDE for artists'. It provides…
942044active
PRIS-CV/DemoFusion
DemoFusion is a CVPR 2024 framework that extends open-source latent diffusion models like SDXL to generate high-resolution images without a…
482041stable
emilianavt/OpenSeeFace
OpenSeeFace is a robust realtime face and facial landmark tracking library that runs on CPU at 30-60 fps using ONNX-converted MobileNetV3 m…
492038active
deep-floyd/IF
DeepFloyd IF is an open-source text-to-image model library implementing a cascaded pixel diffusion architecture with a frozen T5 text encod…
227804maintenance
YUZU-Hub/appscreen
A free, open-source browser-based tool for creating App Store screenshots with customizable backgrounds, text overlays, and 2D/3D device mo…
522032active
jaywcjlove/DevHub
DevHub is an offline, local-first developer toolbox application for macOS built with SwiftUI, bundling 100+ everyday utilities such as JSON…
712027active
serengil/retinaface
RetinaFace is a Python library for deep learning based face detection, built on TensorFlow and derived from the insightface project's Retin…
612027active
renzhezhilu/webp2jpg-online
A browser-based, pure front-end image format converter that converts between jpeg, png, gif, webp, svg, ico, bmp, psd, heic and more withou…
232024stable
TianZerL/Anime4KCPP
Anime4KCPP is a high-performance anime image and video upscaler built on CNN-based algorithms, written in C++. It ships as a library plus V…
782022active
RexanWONG/text-behind-image
An open-source web application for creating text-behind-image designs, where text appears layered behind the subject of a photo. It is avai…
622020active
XavierXiao/Dreambooth-Stable-Diffusion
An implementation of Google's Dreambooth fine-tuning method applied to Stable Diffusion, enabling personalization of a text-to-image diffus…
327738maintenance
instantX-research/InstantStyle
InstantStyle is a framework for style-preserving text-to-image generation that disentangles style and content from reference images using f…
262018active
NVlabs/SPADE
Official PyTorch implementation of SPADE (GauGAN), a CVPR 2019 method for synthesizing photorealistic images from semantic segmentation map…
327717maintenance
sindresorhus/capture-website
A Node.js library for capturing screenshots of websites using Puppeteer (headless Chrome) under the hood. It supports saving screenshots to…
612013active
julyx10/lap
Lap is an open-source, local-first desktop photo manager for macOS, Windows, and Linux built for large personal photo libraries. It offers …
892012active
webp-sh/webp_server_go
A Go-based HTTP server that serves JPEG, PNG, BMP, GIF, SVG and other images as WebP/AVIF/JXL on the fly, without changing the original URL…
952006active
nhn/tui.image-editor
TOAST UI Image Editor is a full-featured photo image editor built on HTML5 Canvas, providing crop, flip, rotate, draw, shape, text, mask, a…
237669maintenance
AntixK/PyTorch-VAE
A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f…
387665maintenance
flytkgl/PDFQFZ
PDFQFZ is a small desktop tool for adding cross-page (riding) seals to PDF documents. It takes a full seal image, randomly splits it across…
902003active
Badge Magic
Badge Magic is a cross-platform mobile and desktop app for creating text, drawings, and animations on LED name badges and transferring them…
802001active
WhatDreamsCost/WhatDreamsCost-ComfyUI
A collection of free custom ComfyUI nodes and workflows, centered on LTX Director, a timeline-based tool for directing LTX video generation…
572001active
bytetriper/RAE
Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u…
482001active
fogleman/sdf
A Python library for generating 3D meshes from signed distance functions (SDFs) with a simple, operator-based API supporting constructive s…
322000stable
DEIM
DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp…
621999active
whomwah/rqrcode
RQRCode is a Ruby library for generating QR codes with a simple interface exposing standard QR code options like error correction level, si…
771998stable
bluefireteam/photo_view
A Flutter library providing a customizable zoomable image widget with gesture support for pinch, rotate, and drag. It can also display arbi…
231997stable
fex-team/webuploader
WebUploader is a JavaScript file upload component that uses HTML5 as its primary runtime with a Flash fallback for legacy browsers like IE6…
107634maintenance
facebookresearch/dino
PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th…
107611maintenance
NVIDIA/Stable-Diffusion-WebUI-TensorRT
An NVIDIA extension for the Stable Diffusion Web UI (Automatic1111) that accelerates image generation using TensorRT-optimized engines on R…
181989active
diana7127/mpv.net-DW
A personally customized Windows build of mpv.net (based on mpv.net_CM and mpv_lazy) that bundles a ModernX playback UI, thumbfast seekbar t…
231988active
zxing-cpp/zxing-cpp
ZXing-C++ is an open-source, multi-format 1D/2D barcode image processing library written in pure C++20, ported from the Java ZXing library …
971987active
svg-sprite/svg-sprite
A low-level Node.js module that takes a batch of SVG files, optimizes them with SVGO, and generates SVG sprites of several types (CSS sprit…
541986stable
LinwoodDev/Butterfly
Butterfly is a powerful, minimalistic, open-source note-taking app built with Flutter that supports drawing, handwriting, and rich text on …
981984active
manisandro/gImageReader
gImageReader is a graphical GTK/Qt front-end to the tesseract-ocr engine for recognizing text in images, PDFs, scans, and screenshots. It s…
511984active
op7418/logo-generator-skill
A Claude Code skill that generates professional SVG logos in 6+ design variants and produces high-end showcase images using Gemini 3.1 Flas…
491984active
xingyizhou/CenterNet
CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi…
327573maintenance
antirez/iris.c
Iris is a pure C inference pipeline that generates images from text prompts using open-weights diffusion transformer models like FLUX.2 Kle…
461983active
hkchengrex/XMem
XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in…
231983stable
Belval/pdf2image
A Python module that wraps the pdftoppm and pdftocairo command-line utilities (from Poppler) to convert PDF pages into PIL Image objects. I…
231982stable
laravolt/avatar
A PHP/Laravel package that generates avatar images from names, emails, or arbitrary strings, producing initials-based avatars as base64, PN…
921981active
blend2d/blend2d
Blend2D is a high-performance 2D vector graphics engine written in C++ that uses a built-in JIT compiler (via AsmJit) to generate optimized…
571981active
thx/resvg-js
A high-performance SVG renderer and toolkit for Node.js, Deno, and browsers, built on the Rust resvg library via napi-rs with a WebAssembly…
821980active
patrikhuber/eos
A lightweight, header-only 3D Morphable Face Model (3DMM) fitting library written in modern C++11/14, with Python bindings. It provides mod…
311980active
gcui-art/markdown-to-image
A React component library that renders Markdown into visually appealing poster images optimized for social media sharing, with support for …
281980active
JIA-Lab-research/DreamOmni2
DreamOmni2 is the official PyTorch implementation of a CVPR 2026 Highlight model for multimodal instruction-based image editing and generat…
511978active
Linzaer/Ultra-Light-Fast-Generic-Face-Detector-1MB
An ultra-lightweight face detection model (~1MB FP32, ~300KB quantized) designed for edge computing devices, with slim and RFB variants tra…
327542maintenance
tandpfun/wardrobe
A self-hosted web application that detects garments in photos, extracts clean product cutouts, and generates modeled editorial previews usi…
541975active
showlab/Show-o
Show-o is a research repository implementing a unified transformer model that combines autoregressive and discrete diffusion modeling for m…
501973active
mborgerding/kissfft
KISS FFT is a mixed-radix Fast Fourier Transform library written in C that supports fixed-point and floating-point data types. It is design…
711972stable
android/androidify
An open-source Android sample app from Google that lets users create custom Android bot avatars using AI image generation via the Gemini AP…
681972active
crystian/ComfyUI-Crystools
A collection of utility custom nodes and UI extensions for ComfyUI, including real-time CPU/GPU/RAM resource monitors, progress bars, and m…
391972active
theamusing/perfectPixel
A Python library that automatically detects the optimal grid size in AI-generated pixel art images and refines them into clean, perfectly a…
451971active
astrofox-io/astrofox
Astrofox is a free, open-source motion graphics application that turns audio into audio-reactive visuals, available as both a web app and a…
671970active
Netflix/void-model
VOID (Video Object and Interaction Deletion) is a research model from Netflix that removes objects from videos along with the physical inte…
541965active
boycy815/PinchImageView
PinchImageView is a lightweight Android image gesture control that extends ImageView with pinch-to-zoom, swipe inertia, double-tap zoom, an…
321961stable
open-mmlab/mmagic
MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri…
237457maintenance
SizheAn/PanoHead
PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis…
291956active
tpaviot/pythonocc-core
pythonocc-core is a Python package providing 3D modeling and data exchange features based on the OpenCascade Technology (OCCT) CAD kernel. …
741955active
alibaba/EasyCV
EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo…
321954active
eriklindernoren/PyTorch-YOLOv3
A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights…
327440maintenance
xdan/jodit
Jodit is an open-source WYSIWYG rich text editor written in pure TypeScript with zero dependencies, offering a built-in file browser and im…
941953stable
ollm/OpenComic
OpenComic is a cross-platform comic and manga reader desktop application built with Node.js and Electron. It supports a wide range of image…
831953active
Fafa-DL/Awesome-Backbones
A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin …
331953active
66HEX/frame
Frame is a native desktop GUI for FFmpeg built in Rust with GPUI-CE, providing media conversion for video, audio, and image files with gran…
771952active
chn-lee-yumi/MaterialSearch
MaterialSearch is a self-hosted semantic search tool that indexes local photos and videos using a CLIP multimodal model, letting users find…
731949active
LTH14/mar
Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive …
541949stable
openai/guided-diffusion
OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio…
327419maintenance
multiavatar/Multiavatar
Multiavatar is an open-source multicultural avatar generator library that converts any input string into a unique SVG avatar, capable of pr…
321946stable
ruyo/VRM4U
VRM4U is an Unreal Engine (UE4/UE5) plugin that imports VRM 3D avatar files at runtime and in-editor. It generates skeletal rigs, morph tar…
971945active
lgarron/folderify
A Rust CLI tool that generates pixel-perfect macOS folder icons in the native style from a PNG mask file, producing .icns and .iconset file…
801941active
clawsoftware/clawPDF
clawPDF is an open-source virtual printer for Windows that converts printed output into PDF, PDF/A, OCR text, SVG, and various image format…
231940active
PixArt-alpha/PixArt-sigma
PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with…
251939active
starik222/BooruDatasetTagManager
A desktop tag editor for managing booru-style tagged image and video datasets used to train Stable Diffusion models such as LoRAs, embeddin…
761938active
NVlabs/RADIO
Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro…
641933active
wysaid/android-gpuimage-plus
A C++ and Java library for Android that applies GPU-accelerated image, camera, and video filters using OpenGL shaders. It supports rule-str…
881927active
riddleling/iOS-OCR-Server
An iOS app that turns an iPhone into a local OCR server using Apple's Vision Framework, exposing an HTTP API and web interface for image te…
741927active

← prev page 11 / 43 next →