Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: image-processing

4273 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
mg-chao/snow-apps
Snow Apps is a C++ open-source suite containing Snow Shot, a screenshot capture and annotation tool, and Snow Image Viewer. Snow Shot offer…
674926active
wuyoscar/GPT-Image2-Skill
A curated prompt gallery and library for OpenAI's GPT Image 2 model, packaged as an agentic skill and Python CLI for image generation and e…
744923active
Augani/openreel-video
OpenReel Video is a fully browser-based, open source video editor built with React, TypeScript, WebCodecs, and WebGPU. All editing and expo…
774914active
eKoopmans/html2pdf.js
html2pdf.js is a JavaScript library that converts webpages or DOM elements into printable PDF files entirely in the browser, built on html2…
874913active
Chlumsky/msdfgen
A C++ library and console tool that generates multi-channel signed distance fields (MSDFs) from vector shapes and font glyphs. The resultin…
664912stable
googlefonts/noto-emoji
Google's open-source Noto Emoji project providing Unicode-compliant emoji fonts, including a color emoji font (CBDT/CBLC format), a monochr…
474904active
arrayfire/arrayfire
ArrayFire is a general-purpose tensor/numerical computing library for C, C++, and Python that accelerates array operations on GPUs (CUDA, O…
574902stable
KaiyangZhou/deep-person-reid
Torchreid is a PyTorch library for deep-learning person re-identification, supporting both image and video reid with end-to-end training an…
504900stable
mpetroff/pannellum
Pannellum is a lightweight, free, open source panorama viewer for the web built with HTML5, CSS3, JavaScript, and WebGL. It ships as a sing…
754881stable
tyxsspa/AnyText
AnyText is the official implementation of a diffusion-based model for multilingual visual text generation and editing in images, accepted a…
324874active
ZzzLc0405/photo-abstract-editorial
A Codex Skill / prompt template that transforms a photo into a vertical editorial composition combining the original photo with an abstract…
574868active
neilsonnn/image-blaster
A Claude skillset that converts a single input image into a full 3D environment, including meshed 3D models (.glb/.obj), Gaussian splats (.…
514822active
google/wuffs
Wuffs is a memory-safe programming language plus a standard library for safely parsing, decoding and encoding untrusted file formats such a…
754820active
lolcommits/lolcommits
lolcommits is a Ruby CLI tool that automatically captures a webcam snapshot every time you make a git commit, archiving a lolcat-style self…
784815active
aloshdenny/reverse-SynthID
A research tool that reverse-engineers Google's SynthID watermark embedded in Gemini-generated images using spectral analysis and signal pr…
664815active
endroid/qr-code
A PHP library for generating QR codes with support for multiple writers (PNG, WebP, SVG, EPS, binary), logos, labels, and error correction.…
624806stable
palxiao/poster-design
XunPai Design (poster-design) is an open-source online image editor and poster designer built with Vue3, Vite, and Express, inspired by too…
434802active
charmbracelet/freeze
Freeze is a Go CLI tool that generates PNG, SVG, and WebP images of code snippets and terminal output. It supports syntax highlighting, ANS…
714799active
UX-Decoder/Segment-Everything-Everywhere-All-At-Once
SEEM is the official PyTorch implementation of the NeurIPS 2023 paper 'Segment Everything Everywhere All at Once', a model for universal im…
204794stable
zju3dv/EasyMocap
EasyMocap is an open-source Python toolbox for markerless human motion capture and novel view synthesis from RGB videos. It fits parametric…
544783active
Bing-su/adetailer
ADetailer is an extension for the Stable Diffusion WebUI (A1111) that automatically detects objects such as faces and hands in generated im…
744781active
EFPrefix/EFQRCode
EFQRCode is a lightweight, pure-Swift library for generating stylized QR code images (with watermarks, icons, or GIFs) and recognizing QR c…
664755stable
cvg/LightGlue
LightGlue is a deep neural network library that matches sparse local features across image pairs with high accuracy and fast inference. It …
504728stable
esimov/pigo
Pigo is a pure Go library for fast face detection, pupil/eye localization, and facial landmark detection based on the Pixel Intensity Compa…
314728stable
CloudCompare/CloudCompare
CloudCompare is a 3D point cloud and triangular mesh processing application, originally built to compare point clouds from laser scanners a…
674693active
yangjian102621/geekai
GeekAI is an open-source, self-hosted AI content creation platform integrating chat, image generation (MidJourney, DALL-E, Stable Diffusion…
954686active
cf-pages/Telegraph-Image
A free, self-deployable image hosting application that serves as a Flickr/Imgur alternative, built on Cloudflare Pages with uploads stored …
754655active
f3d-app/f3d
F3D is a fast, minimalist open-source 3D viewer desktop application supporting many formats (glTF, USD, STL, STEP, OBJ, FBX, Alembic) with …
884650active
FooIbar/EhViewer
A modern Android client for E-Hentai, overhauled with Material Design 3 and dynamic color support, forked from Ehviewer-Overhauled. It is a…
864647active
Kwai-Kolors/Kolors
Kolors is a large-scale latent diffusion model for photorealistic text-to-image synthesis, trained with bilingual (Chinese and English) tex…
234615active
sensity-ai/dot
dot (Deepfake Offensive Toolkit) is a Python tool that generates real-time, controllable deepfakes from a webcam feed and injects them into…
234586active
joanrod/star-vector
StarVector is a foundation model that generates scalable vector graphics (SVG) code from images and text by treating vectorization as a cod…
494560active
xyxiao001/vue-cropper
A Vue.js component plugin for cropping images in the browser, supporting Vue 2 and Vue 3. It offers rotation, zooming, fixed aspect ratios,…
554559active
spipm/Depixelization_poc
Depix is a proof-of-concept tool that recovers plaintext from pixelized screenshots by matching pixelated blocks against a rendered font se…
104551active
crabbly/Print.js
Print.js is a tiny JavaScript library that helps printing from the web, supporting printing of PDF, HTML, image, and JSON content directly …
654545stable
TencentARC/InstantMesh
InstantMesh is a feed-forward framework for generating 3D meshes from a single image using sparse-view large reconstruction models (LRM/Ins…
254509active
mcmonkeyprojects/SwarmUI
SwarmUI (formerly StableSwarmUI) is a modular, self-hosted web user interface for AI image generation, supporting models like Stable Diffus…
764499active
burhanrashid52/PhotoEditor
An Android photo editing library that lets apps add paint drawing, text, filters, emoji, and stickers to images, similar to Instagram/Faceb…
734499stable
Aidoku/Aidoku
Aidoku is a free, open-source manga reading application for iOS, iPadOS, and macOS with no ads. It supports local CBZ files, self-hosted me…
924495active
royshil/obs-backgroundremoval
An OBS Studio plugin that removes and replaces the background in portrait video using ONNX-based machine learning segmentation, acting as a…
984492active
KnpLabs/snappy
Snappy is a PHP library that wraps the wkhtmltopdf and wkhtmltoimage binaries to generate PDFs, snapshots, or thumbnails from URLs or HTML …
924476stable
php-imagine/Imagine
Imagine is an object-oriented image manipulation library for PHP, inspired by Python's PIL. It provides a unified API over GD2, Imagick, an…
894471stable
cyanfish/naps2
NAPS2 is a free, open-source document scanning application for Windows, Mac, and Linux that supports WIA, TWAIN, SANE, and ESCL scanners an…
974464active
astriaai/headshots-starter
An open-source Next.js starter kit that generates professional AI headshots from user-uploaded selfies using Astria.ai's fine-tuning and in…
404463active
metatube-community/jellyfin-plugin-metatube
A metadata provider plugin for Jellyfin and Emby media servers that fetches movie and actor metadata from various internet providers via th…
614454active
layumi/Person_reID_baseline_pytorch
A small, friendly PyTorch baseline implementation for person and vehicle re-identification (ReID). It reproduces strong top-conference resu…
654446stable
rom1504/img2dataset
A Python tool that downloads large sets of image URLs and packages them into machine learning datasets, with resizing and caption support. …
564443active
Zeejay0/gathered-scenes-zine-skill
A collection of image-generation skills (prompt packs) written for Codex that transform ordinary photos into zine-style paper artworks. It …
574440active
xlite-dev/lite.ai.toolkit
A lightweight C++ toolkit providing unified APIs for 100+ pre-trained AI models across inference backends like ONNX Runtime, MNN, TensorRT,…
744427active
Nutlope/restorePhotos
A Next.js web application that restores old and blurry face photos using the GFPGAN ML model via the Replicate API. It provides a hosted se…
314427active
huawei-noah/Efficient-AI-Backbones
A collection of efficient neural network backbone architectures (GhostNet, TNT, ViG, WaveMLP, TinyNet, etc.) from Huawei Noah's Ark Lab, wi…
284418active
codeforreal1/compressO
CompressO is a free, open-source desktop app for compressing videos and images to tiny sizes entirely offline, built with Tauri, Rust, and …
864415active
imazen/imageflow
Imageflow is a high-performance, memory-safe image manipulation suite written in Rust, offering a C ABI library (libimageflow), a CLI tool …
924412active
pyqtgraph/pyqtgraph
PyQtGraph is a pure-Python graphics and GUI library built on PyQt/PySide and NumPy for fast 2D/3D scientific plotting and interactive data …
734405stable
libjpeg-turbo/libjpeg-turbo
libjpeg-turbo is a SIMD-accelerated JPEG image codec library that is API/ABI-compatible with libjpeg and generally 2-6x faster. It provides…
884401stable
iperov/DeepFaceLab
DeepFaceLab is the leading open-source Windows application for creating deepfakes, allowing users to swap, de-age, or replace faces and hea…
1019292maintenance
bowang-lab/MedSAM
MedSAM is a fine-tuned Segment Anything Model (SAM) foundation model for universal medical image segmentation, trained on over 1.5 million …
294379active
s1dashu/ip-as-logo-skill
A compact Agent Skill that guides AI agents to generate highly simplified, rounded, subtly neo-skeuomorphic IP mascot logos. It follows the…
574374active
dreamgaussian/dreamgaussian
DreamGaussian is the official PyTorch implementation of an ICLR 2024 Oral paper for efficient 3D content creation using generative Gaussian…
184352active
VectorSpaceLab/OmniGen
OmniGen is a unified diffusion-based image generation model that produces and edits images from multi-modal prompts without auxiliary modul…
474340active
OHIF/Viewers
OHIF Viewer is an open-source, zero-footprint web-based medical imaging viewer for DICOM images, built as a configurable and extensible pro…
984309active
kohler/gifsicle
Gifsicle is a command-line tool for creating, editing, and optimizing GIF images and animations, with companion programs gifview (a viewer)…
624307stable
Baseflow/PhotoView
PhotoView is an Android library providing a drop-in replacement for ImageView that supports zooming via multi-touch and double-tap gestures…
2318813maintenance
yuzutech/kroki
Kroki is a unified HTTP API service that converts textual diagram descriptions (PlantUML, Mermaid, GraphViz, Ditaa, Excalidraw, and many mo…
944297active
Tencent-Hunyuan/HunyuanDiT
Hunyuan-DiT is Tencent's open-source diffusion transformer model for text-to-image generation with fine-grained Chinese language understand…
494291active
richzhang/PerceptualSimilarity
A PyTorch library implementing the LPIPS (Learned Perceptual Image Patch Similarity) metric, which measures perceptual distance between ima…
234269stable
LycheeOrg/Lychee
Lychee is a free, open-source, self-hosted photo management system with a web interface for uploading, organizing, and sharing photos. It r…
954265active
hoothin/UserScripts
A collection of Greasemonkey/Tampermonkey userscripts by hoothin, including Pagetual (auto-pager infinite scrolling), Picviewer CE+ (online…
764264active
SysCV/sam-hq
HQ-SAM (Segment Anything in High Quality) upgrades Meta's Segment Anything Model with a learnable High-Quality Output Token for accurate ze…
484255active
ali-vilab/AnyDoor
AnyDoor is the official implementation of a diffusion-based model that teleports target objects into new scenes at user-specified locations…
284238active
TheAlphamerc/flutter_twitter_clone
Fwitter is a fully functional Twitter clone mobile app built with the Flutter framework, using Firebase authentication, realtime database, …
234237active
mausimus/ShaderGlass
ShaderGlass is a Windows (and Wine) overlay application that applies GPU shader effects on top of the desktop, a window, or a captured sour…
864214active
ArcReel/ArcReel
ArcReel is an open-source, self-hosted AI video production workspace that turns novels, scripts, or product material into characters, scene…
824211active
anthonynsimon/bild
bild is a collection of parallel image processing algorithms written in pure Go, usable both as a Go library and as a CLI tool. It supports…
964204active
TokisanGames/Terrain3D
Terrain3D is a high-performance, editable terrain system for Godot 4, written in C++ as a GDExtension addon. It supports sculpting, texture…
844199active
evanw/thumbhash
ThumbHash is a compact encoding of an image placeholder that can be stored inline with data and rendered while the real image loads. It is …
304197stable
cvg/Hierarchical-Localization
hloc is a modular Python toolbox for state-of-the-art 6-DoF visual localization, combining image retrieval and feature matching (SuperPoint…
484194active
oxipng/oxipng
Oxipng is a multithreaded, lossless PNG/APNG compression optimizer written in Rust. It can be used as a command-line utility or as a Rust l…
944188stable
fatihak/InkyPi
InkyPi is an open-source E-Ink display application that runs on a Raspberry Pi and is controlled through a web interface. It supports plugi…
664184active
lllyasviel/style2paints
Style2Paints is an AI-driven tool that colorizes lineart sketches, optionally guided by human hints, style reference images, and lighting. …
3218179maintenance
hackerb9/lsix
lsix is a shell script that displays image thumbnails directly in the terminal using sixel graphics, powered by ImageMagick. It works like …
234174stable
Fannovel16/comfyui_controlnet_aux
A collection of ComfyUI custom nodes providing ControlNet auxiliary preprocessors that generate hint images (Canny edges, lineart, depth ma…
634163active
torchgeo/torchgeo
TorchGeo is a PyTorch domain library, similar to torchvision, providing datasets, samplers, transforms, and pre-trained models specific to …
954159active
bingoogolapple/BGABanner-Android
An Android UI library providing banner/carousel components with swipe navigation for guide screens, infinite auto-looping for one or more p…
754156active
nodeca/pica
A browser-side JavaScript library for high-quality, high-speed image resizing using canvas, web workers, WebAssembly, and createImageBitmap…
764144stable
RawTherapee/RawTherapee
RawTherapee is a free, cross-platform raw photo processing program for developing images from digital cameras, written in C++ with a GTK fr…
874134active
justadudewhohacks/face-api.js
A JavaScript face detection and face recognition library built on top of tensorflow.js, usable in the browser and Node.js. It provides mode…
2317945maintenance
opengeos/segment-geospatial
SamGeo (segment-geospatial) is a Python package that applies Meta AI's Segment Anything Model (SAM, SAM2, SAM3, HQ-SAM) to geospatial data …
974122active
WebODM/WebODM
WebODM is a user-friendly, commercial-grade application for drone image processing that generates georeferenced maps, point clouds, elevati…
984116active
lllyasviel/sd-forge-layerdiffuse
A Stable Diffusion WebUI (Forge) extension implementing Layer Diffusion to generate transparent images and separate foreground/background l…
254116active
VectorSpaceLab/OmniGen2
OmniGen2 is an open-source unified multimodal generation model supporting text-to-image generation, instruction-guided image editing, and i…
524112active
dominictobias/react-image-crop
A dependency-free React component for responsive image cropping with pixel or percentage coordinates. It supports touch, keyboard accessibi…
804103active
cdcseacave/openMVS
OpenMVS is an open-source C++ library for Multi-View Stereo 3D reconstruction, taking camera poses and a sparse point-cloud as input and pr…
764100active
ZhengPeng7/BiRefNet
BiRefNet is a PyTorch implementation of the CAAI AIR 2024 paper 'Bilateral Reference for High-Resolution Dichotomous Image Segmentation'. I…
654098active
twitter/twemoji
Twemoji is a library from Twitter that provides standard Unicode emoji support across all platforms via a set of SVG/PNG assets and a JavaS…
6417763maintenance
GuyTevet/motion-diffusion-model
Official PyTorch implementation of the Human Motion Diffusion Model (MDM) paper, generating 3D human motion sequences from text prompts usi…
524092active
aFarkas/lazysizes
lazysizes is a high-performance, SEO-friendly lazy loader for images (including responsive picture/srcset), iframes, scripts, and widgets. …
2317718maintenance
lllyasviel/Paints-UNDO
Paints-UNDO is a family of deep learning models that take an image as input and generate the step-by-step drawing sequence (sketching, inki…
404067active
Dimezis/BlurView
An Android library providing a BlurView, a FrameLayout-like view that dynamically blurs its underlying content in an iOS-like fashion and r…
774045active
armory3d/armorpaint
ArmorPaint is a stand-alone 3D PBR texture painting application that runs entirely on the GPU, supporting node-based procedural materials, …
674039active

← prev page 5 / 43 next →