function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| muzishen/IMAGDressing IMAGDressing-v1 is a diffusion-based framework for customizable virtual dressing that generates human images with fixed garments and contro… | 44 | 1343 | active |
| zanllp/infinite-image-browsing Infinite Image Browsing (IIB) is a full-featured image and video management application with fast thumbnail-based browsing, AI-generation m… | 93 | 1341 | active |
| FireRedTeam/FireRed-Image-Edit FireRed-Image-Edit is an open-source image editing foundation model built on diffusion models, released as PyTorch model weights with infer… | 49 | 1341 | active |
| insoxin/API A free, self-hostable REST API platform (姬长信 API) packaged as a Docker image, offering dozens of endpoints for daily-life services, news, w… | 66 | 1339 | active |
| receyuki/stable-diffusion-prompt-reader A standalone desktop application (with GUI and CLI) for reading generation prompts and parameters embedded in Stable Diffusion images, with… | 21 | 1339 | stable |
| senguptaumd/Background-Matting Official research code for 'Background Matting: The World is Your Green Screen' (CVPR 2020), a deep network that extracts per-pixel alpha m… | 32 | 4769 | maintenance |
| christophschuhmann/improved-aesthetic-predictor A CLIP+MLP neural network that predicts how much people on average like an image, trained on AVA dataset ratings. It is widely used for fil… | 32 | 1338 | stable |
| wormtql/yas Yas is a fast screen-scanning tool that uses a custom-trained SVTR OCR model to read Genshin Impact and Honkai: Star Rail artifact stats di… | 40 | 1336 | active |
| ByteDance-Seed/SeedVR SeedVR/SeedVR2 are diffusion-transformer based models for generic real-world and AIGC video and image restoration, with SeedVR2 using adver… | 47 | 1334 | active |
| liwenxi/SWIFT-AI SWIFT-AI is a deep learning system for extremely fast gigapixel-level visual understanding in scientific applications, such as detecting st… | 29 | 1334 | active |
| fayazara/Screendrop Screendrop is a native macOS app for screenshots and screen recording, serving as a free, self-hostable Loom alternative. It offers annotat… | 81 | 1333 | active |
| wenqsun/DimensionX DimensionX is a research framework that generates photorealistic 3D and 4D scenes from a single image using controllable video diffusion mo… | 43 | 1333 | active |
| microshow/RxFFmpeg RxFFmpeg is an Android framework built on FFmpeg 4.0 (with X264, mp3lame, fdk-aac, opencore-amr, and OpenSSL) for fast audio and video edit… | 23 | 4745 | maintenance |
| chocoford/ExcalidrawZ ExcalidrawZ is a native Excalidraw client for macOS, iPadOS, and iOS built with SwiftUI. It wraps the familiar Excalidraw canvas with nativ… | 98 | 1332 | active |
| embedded-graphics/embedded-graphics A no_std 2D graphics library for Rust targeting memory-constrained embedded devices. It draws primitives, text, and images using an iterato… | 76 | 1332 | stable |
| jtydhr88/ComfyUI-qwenmultiangle A ComfyUI custom node providing an interactive Three.js 3D viewport for controlling camera azimuth, elevation, and zoom. It outputs formatt… | 56 | 1332 | active |
| dapi-labs/react-nice-avatar A React component library for generating illustrated SVG avatars from seeds or custom configuration objects. It offers customizable attribu… | 23 | 1332 | stable |
| moshstudio/TAICHI-flet TAICHI-flet is a Windows desktop entertainment application built with the Flet framework that lets users browse images, music, novels, comi… | 75 | 4735 | maintenance |
| ldqk/ImageSearch A .NET 10 desktop demo application that performs reverse image search (search by image) over local hard drives with tens of millions of ima… | 93 | 1329 | active |
| leviarista/github-profile-header-generator A web-based generator that creates customizable header/banner images for GitHub profile READMEs and repository banners. Users can tweak tex… | 81 | 1329 | active |
| bytedance/Lance Lance is a 3B-parameter native unified multimodal model from ByteDance for image and video understanding, generation, and editing, trained … | 55 | 1329 | active |
| Swati4star/Images-to-PDF An open-source Android app that converts images (JPG and others) from the camera or gallery into PDF files. It also offers PDF management f… | 70 | 1328 | active |
| jackmoore/colorbox Colorbox is a lightweight, customizable lightbox plugin for jQuery that displays images, galleries, slideshows, ajax, inline, and iframed c… | 10 | 4724 | maintenance |
| Gourieff/ComfyUI-ReActor ComfyUI-ReActor is a fast and simple face swap extension node for ComfyUI, based on the ReActor face-swapping engine. It includes a nudity … | 65 | 1327 | active |
| numandev1/react-native-compressor A React Native library that compresses images, videos, and audio with WhatsApp-like quality, plus background upload, file download, and vid… | 96 | 1325 | active |
| cvzone/cvzone CVZone is a Python computer vision helper library that wraps OpenCV and MediaPipe to simplify image processing and AI functions like hand t… | 32 | 1325 | active |
| nateraw/stable-diffusion-videos A Python library for generating videos with Stable Diffusion by walking the latent space and morphing between text prompts. It supports bea… | 56 | 4705 | maintenance |
| thephpleague/color-extractor A PHP library that extracts the most representative colors from an image, building a color palette sorted by pixel count. It handles transp… | 88 | 1323 | stable |
| galilai-group/lejepa LeJEPA is a Python framework for scalable, theoretically grounded self-supervised representation learning based on Joint-Embedding Predicti… | 45 | 1322 | active |
| ImprintLab/Medical-SAM-Adapter Medical SAM Adapter (MSA) is a Python framework that fine-tunes Meta's Segment Anything Model for medical image segmentation using lightwei… | 39 | 1322 | active |
| neozhaoliang/surround-view-system-introduction A Python implementation of a vehicle surround-view (bird's-eye view) camera system, covering fisheye camera calibration, projection, image … | 66 | 1321 | active |
| meta-pytorch/segment-anything-fast A fast, batched offline inference-oriented fork of Meta's Segment Anything (SAM) image segmentation model. It applies optimizations like bf… | 45 | 1321 | active |
| gavrielc/Nano-PDF A Python CLI tool that edits PDF slides using natural language prompts, powered by Google's Gemini 3 Pro Image model. It renders pages to i… | 41 | 1321 | active |
| lsky-org/lsky-pro Lsky Pro is a self-hosted PHP/Laravel image hosting and photo album application for storing, managing, and sharing images on the cloud. It … | 53 | 4692 | maintenance |
| Decimation/SmartImage SmartImage is a reverse image search tool that queries multiple engines (SauceNao, IQDB, Ascii2D, trace.moe, TinEye, Yandex, and more) and … | 97 | 1320 | active |
| yzane/vscode-markdown-pdf A Visual Studio Code extension that converts Markdown files to PDF, HTML, PNG, or JPEG. It supports PlantUML and Mermaid diagrams, KaTeX ma… | 93 | 1319 | active |
| transmute-app/transmute Transmute is a self-hosted web application for converting and compressing files, supporting images, video, audio, JSON, Excel, PDF and more… | 79 | 1319 | active |
| nroduit/Weasis Weasis is an open-source DICOM viewer for medical imaging that runs standalone or embedded in web applications. It integrates with PACS, VN… | 95 | 1317 | active |
| netdcy/FlowVision FlowVision is a waterfall-style image viewer for macOS with smooth scrolling and immersive browsing. It supports video playback, HDR displa… | 86 | 1317 | active |
| Chlumsky/msdf-atlas-gen A C++ utility and library that generates multi-channel signed distance field (MSDF) font atlases from TTF/OTF fonts, packing glyphs into co… | 71 | 1317 | active |
| yoksel/url-encoder A browser-based tool that URL-encodes SVG markup so it can be embedded in CSS as data URIs for background-image, border-image, or mask prop… | 34 | 1317 | active |
| python-escpos/python-escpos A Python library for controlling ESC/POS receipt printers as defined by Epson, supporting text, images, barcodes, and QR codes over USB, se… | 67 | 1316 | active |
| Arnklit/Waterways A Godot Engine add-on that generates river meshes with flow and foam maps from bezier curves, entirely within the editor. It includes path … | 48 | 1316 | active |
| jstkdng/ueberzugpp Überzug++ is a C++ command line utility that draws images directly in terminals using X11/Wayland child windows, sixels, or the kitty and i… | 77 | 1315 | active |
| BigBadaboom/androidsvg AndroidSVG is a parser and renderer for SVG files on Android, supporting nearly all static visual elements of SVG 1.1 and SVG 1.2 Tiny (exc… | 34 | 1315 | stable |
| ali-vilab/MimicBrush MimicBrush is the official implementation of a zero-shot image editing method that lets users mask a region in a source image and provide a… | 24 | 1311 | active |
| flozz/StackBlur StackBlur.js is a JavaScript library implementing the fast, almost-Gaussian StackBlur algorithm for images and canvas elements. It works in… | 73 | 1310 | stable |
| ndl-lab/ndlocr-lite NDLOCR-Lite is a lightweight Japanese OCR application developed by the National Diet Library that converts digitized images of books and ma… | 77 | 1309 | active |
| spatie/laravel-image-optimizer A Laravel package that optimizes PNG, JPG, SVG, and GIF images by running them through a chain of installed optimization binaries. It is th… | 77 | 1309 | stable |
| svg-net/SVG SVG.NET is a C# library for reading, writing, and rendering SVG 1.1 images in .NET applications. It targets .NET Standard 2.0 and is distri… | 74 | 1308 | active |
| draftbit/avatar-generator Personas is an open-source avatar generator web app by Draftbit that lets users mix and match skin, hair, facial hair, body, eyes, mouth, n… | 76 | 1307 | active |
| huawei-noah/Efficient-Computing A collection of efficient deep learning methods from Huawei Noah's Ark Lab, covering model compression, knowledge distillation, pruning, qu… | 32 | 1307 | active |
| seetaface/SeetaFaceEngine SeetaFace Engine is an open-source C++ face recognition engine comprising face detection, face alignment, and face identification modules. … | 32 | 4636 | maintenance |
| igordanchenko/yet-another-react-lightbox Yet Another React Lightbox is a modern, performant lightbox component for React for displaying image (and optionally video) galleries. It i… | 98 | 1305 | active |
| Vincentqyw/image-matching-webui A Gradio-based web UI that matches keypoints between two images using many state-of-the-art image matching algorithms (LoFTR, SuperGlue, Li… | 91 | 1302 | active |
| vye16/shape-of-motion Shape of Motion is a Python research codebase for 4D reconstruction of dynamic scenes from a single monocular video, based on the ICCV 2025… | 30 | 1302 | active |
| luosiallen/latent-consistency-model Official implementation of Latent Consistency Models (LCM), a diffusion-based approach for synthesizing high-resolution images with few-ste… | 27 | 4615 | maintenance |
| rsmbl/Resemble.js Resemble.js is a JavaScript library for analyzing and comparing images using HTML5 canvas, producing pixel-level diff images and analysis d… | 32 | 4612 | maintenance |
| playcanvas/splat-transform SplatTransform is an open-source CLI tool and library for converting, editing, and optimizing 3D Gaussian splat files. It supports reading … | 84 | 1300 | active |
| besscroft/PicImpact PicImpact is a self-hostable photography portfolio website built with Next.js and Hono.js for photographers to showcase their work. It feat… | 80 | 1300 | active |
| sighook/pixload pixload is a set of Perl CLI tools for creating and injecting payloads into image files (BMP, GIF, JPG, PNG, WebP). It is used in offensive… | 23 | 1300 | active |
| danvergara/morphos Morphos is a self-hosted file converter server written in Go that lets users convert files between formats privately, without third-party s… | 10 | 1300 | active |
| HG-ha/MTools MTools is a cross-platform desktop application built with Python and Flet that bundles image processing, audio/video editing, text operatio… | 83 | 1298 | active |
| Uminosachi/sd-webui-inpaint-anything A Stable Diffusion Web UI extension that performs inpainting and outpainting using masks generated by Segment Anything models (SAM 2, SAM-H… | 30 | 1298 | active |
| jipika/WaifuX WaifuX is an open-source macOS application that aggregates static and dynamic anime wallpapers from sources like Wallhaven and MotionBGs, w… | 77 | 1297 | active |
| donydchen/mvsplat MVSplat is a PyTorch implementation of an ECCV 2024 Oral model that predicts 3D Gaussians from sparse multi-view images in a single feed-fo… | 61 | 1296 | active |
| PyImageSearch/imutils A Python library of convenience functions that simplify common OpenCV image processing tasks such as translation, rotation, resizing, skele… | 32 | 4590 | maintenance |
| ARM-software/astc-encoder The Arm ASTC Encoder (astcenc) is a command-line tool and codec library for compressing and decompressing images in the Adaptive Scalable T… | 92 | 1295 | stable |
| HSLix/LixAssistantLimbusCompany LixAssistantLimbusCompany (LALC) is a free, open-source Windows desktop assistant for the game Limbus Company that automates daily gameplay… | 90 | 1295 | active |
| cnr-isti-vclab/vcglib VCGlib is a templated, header-only C++ library with no external dependencies for manipulating, processing, cleaning, and simplifying triang… | 67 | 1295 | stable |
| fundamentalvision/BEVFormer BEVFormer is the official PyTorch implementation of an ECCV 2022 paper that learns bird's-eye-view (BEV) representations from multi-camera … | 23 | 4579 | maintenance |
| W2GenAI-Lab/LucidFlux LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a… | 55 | 1293 | active |
| Parskatt/RoMa RoMa (romatch) is a Python library for robust dense feature matching between image pairs, estimating pixel-dense warps and reliable certain… | 52 | 1293 | active |
| artur-graniszewski/DLSS-Enabler DLSS-Enabler is a Windows installer/tool that lets users simulate NVIDIA DLSS upscaling and DLSS-G frame generation on any DirectX 12 compa… | 86 | 1290 | active |
| streamlit/demo-self-driving A Streamlit demo app that provides an interactive image browser for the Udacity self-driving-car dataset with realtime YOLO object detectio… | 60 | 1290 | stable |
| BishopFox/eyeballer Eyeballer is a convolutional neural network tool that classifies screenshots of web hosts taken during large-scope penetration tests. It la… | 55 | 1290 | active |
| ClownsharkBatwing/RES4LYF RES4LYF is a ComfyUI custom node collection providing advanced diffusion samplers (RES samplers) that achieve high-quality image generation… | 67 | 1289 | active |
| hku-mars/livox_camera_calib A C++/ROS tool from HKU MARS for automatic extrinsic calibration between high-resolution LiDAR (e.g., Livox) and cameras in targetless envi… | 32 | 1289 | stable |
| t3dotgg/quickpic QuickPic is a free, open-source web app by Theo that converts SVGs to high-resolution PNGs and offers other quick image utilities like squa… | 22 | 1289 | active |
| azagaya/laigter Laigter is an open-source desktop tool that automatically generates normal, specular, parallax, and occlusion maps from 2D textures, design… | 90 | 1288 | active |
| bytedance/Bernini Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer perf… | 57 | 1287 | active |
| lzhgus/Capso Capso is a free, open-source native macOS app for screenshots and screen recording, built with Swift 6.0 and SwiftUI as an alternative to C… | 81 | 1286 | active |
| SebLague/Fluid-Sim A particle-based fluid simulation built in Unity with C#, implementing SPH-style methods from academic papers. It includes GPU-accelerated … | 48 | 1286 | active |
| huanngzh/MV-Adapter MV-Adapter is a plug-and-play adapter that turns pre-trained text-to-image diffusion models (e.g., SDXL, SD2.1) into multi-view consistent … | 34 | 1285 | active |
| BoboTiG/python-mss Python MSS is an ultra-fast, cross-platform screenshot library in pure Python using ctypes, capable of capturing one or all monitors with n… | 81 | 1280 | stable |
| Tianxiaomo/pytorch-YOLOv4 A minimal PyTorch implementation of YOLOv4 (and YOLOv4-tiny) supporting inference and training, with tools to convert Darknet weights to Py… | 32 | 4521 | maintenance |
| zju3dv/MatchAnything MatchAnything is a deep learning model for universal cross-modality image matching, released as research code accompanying a TPAMI 2026 pap… | 64 | 1279 | active |
| plemeri/transparent-background A Python tool and CLI that removes backgrounds from images and videos using the InSPyReNet deep learning model (ACCV 2022). It supports ima… | 63 | 1278 | active |
| Tencent-Hunyuan/SRPO SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion… | 54 | 1278 | active |
| city-super/Scaffold-GS Scaffold-GS is a research implementation of a structured 3D Gaussian splatting method that uses anchor points on a sparse voxel grid to dis… | 27 | 1278 | active |
| PrunaAI/pruna Pruna is an open-source Python model optimization framework that makes AI models faster, smaller, cheaper, and greener via caching, quantiz… | 84 | 1275 | active |
| nyanmisaka/ffmpeg-rockchip A fork of FFmpeg adding full hardware transcoding pipelines for Rockchip SoCs via MPP decoders/encoders and RGA 2D accelerator filters, wit… | 70 | 1275 | active |
| google-research/simclr Google Research's official implementation of SimCLR and SimCLRv2, a framework for contrastive learning of visual representations, with 65 p… | 10 | 4502 | maintenance |
| Stable-X/Stable3DGen Stable3DGen is a modular Python framework for generating 3D assets from images, adapted from Microsoft's TRELLIS with NVIDIA library depend… | 33 | 1274 | active |
| DreamTechAI/Direct3D-S2 Direct3D-S2 is a research framework for high-resolution 3D shape generation from images, built on sparse volumetric representations and a n… | 29 | 1274 | active |
| dcharatan/pixelsplat pixelSplat is a PyTorch implementation of a feed-forward model that reconstructs 3D radiance fields parameterized by 3D Gaussian primitives… | 27 | 1274 | stable |
| ziguishian/xhs-visual-director-skill An Agent Skill (Codex Skill) that acts as a 'visual director' for planning Xiaohongshu (Little Red Book) image-and-text posts. It interview… | 54 | 1272 | active |
| NVlabs/stylegan2-ada-pytorch Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi… | 32 | 4487 | maintenance |
| Yummypets/YPImagePicker YPImagePicker is an Instagram-like photo and video picker library for iOS written in pure Swift. It provides a feature-rich, highly customi… | 87 | 4483 | maintenance |
| GarrettGunnell/AcerolaFX A modular HDR post-processing shader suite written in HLSL for GShade, targeting Final Fantasy XIV gameplay and gpose screenshots. It provi… | 23 | 1269 | active |