function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| Yuanshi9815/OminiControl OminiControl is a universal control framework for Diffusion Transformer models like FLUX, supporting subject-driven and spatial control (ed… | 62 | 1927 | active |
| AbdBarho/stable-diffusion-webui-docker A Docker Compose setup that runs Stable Diffusion locally with popular web UIs like AUTOMATIC1111, ComfyUI, and InvokeAI. It packages model… | 23 | 7309 | maintenance |
| Code-with-Beto/snapai SnapAI is a Node.js CLI that generates 1024x1024 mobile app icons and 1024x500 Google Play feature graphics using OpenAI or Google Gemini i… | 76 | 1923 | active |
| kfrlib/kfr KFR is a fast, modern C++ digital signal processing framework providing FFT/DFT, FIR/IIR/biquad filter design and processing, sample rate c… | 94 | 1917 | stable |
| pymatting/pymatting PyMatting is a Python library for alpha matting that estimates an alpha matte from an input image and a hand-drawn trimap to extract foregr… | 67 | 1914 | active |
| strukturag/libde265 libde265 is an open source implementation of the H.265 (HEVC) video codec written from scratch in C++ with a plain C API for easy integrati… | 95 | 1910 | stable |
| HoshinoSuzumi/chronoframe ChronoFrame is a self-hosted personal photo gallery application for photographers, built with Nuxt 4 and TypeScript. It supports online pho… | 69 | 1910 | active |
| diffgram/diffgram Diffgram is a self-hosted AI datastore for managing schemas, BLOBs, and predictions, with built-in human supervision (data labeling), data … | 62 | 1909 | active |
| Folder11 Folder11 is a Windows application for customizing folder icons, accompanied by an icon gallery and an automated icon repository (Folder-Ico… | 77 | 1908 | active |
| derf/feh feh is a fast, lightweight, and highly configurable image viewer for X11, aimed at command-line users but also launchable from file manager… | 77 | 1908 | stable |
| scaleflex/filerobot-image-editor Filerobot Image Editor is an easy-to-integrate JavaScript image editing library for web applications, supporting resize, crop, flip, finetu… | 69 | 1908 | active |
| arcsin1/oh-my-ppt Oh My PPT is a local-first, open-source desktop application (Electron + React + TypeScript) that uses AI to generate editable HTML-based pr… | 81 | 1907 | active |
| sicxu/Deep3DFaceRecon_pytorch A PyTorch implementation of Deep3DFaceReconstruction, a weakly-supervised CNN method for reconstructing 3D face geometry from a single imag… | 32 | 1907 | stable |
| NVlabs/nvdiffrast Nvdiffrast is a PyTorch library from NVIDIA providing high-performance, GPU-accelerated primitive operations for rasterization-based differ… | 55 | 1905 | stable |
| visual-layer/fastdup fastdup is a free Python tool for rapidly analyzing image and video datasets to surface duplicates, outliers, broken, dark, bright, blurry,… | 67 | 1904 | active |
| sindresorhus/gulp-imagemin A Gulp plugin that minifies PNG, JPEG, GIF, and SVG images using imagemin, with bundled optimizers (gifsicle, mozjpeg, optipng, svgo). It i… | 51 | 1904 | stable |
| Faceplugin-ltd/Open-Source-Face-Recognition-SDK An open-source face recognition SDK by Faceplugin providing face detection, landmark extraction, feature embedding generation, and face tem… | 64 | 1903 | active |
| HashLips Art Engine HashLips Art Engine is a Node.js tool that generates multiple unique instances of generative artwork by compositing user-provided image lay… | 23 | 7207 | maintenance |
| tilltue/TLPhotoPicker TLPhotoPicker is a Swift library for iOS that provides a Facebook-style photo and video picker supporting multiple PHAsset selection across… | 79 | 1898 | active |
| amebalabs/TRex TRex is a macOS menu bar application that extracts text from any visible screen area using OCR, copying it directly to the clipboard. It wo… | 84 | 1896 | active |
| garris/BackstopJS BackstopJS is a Node.js tool that automates visual regression testing of web apps by capturing screenshots with Chrome Headless and compari… | 23 | 7176 | maintenance |
| CristianOlivera1/openvid Openvid is an open-source, browser-based video editor for recording screens and creating polished product demos with cinematic zooms, 3D de… | 60 | 1894 | active |
| MinJieLiu/react-photo-view A lightweight (7KB gzipped) React photo preview/lightbox component with polished touch gestures, animations, and keyboard navigation. It su… | 52 | 1894 | stable |
| rdumasia303/deepseek_ocr_app A self-hosted OCR web application combining a React frontend with a FastAPI backend, powered by the DeepSeek-OCR model. It processes images… | 50 | 1892 | active |
| d8ahazard/sd_dreambooth_extension A Stable Diffusion WebUI extension that adds DreamBooth fine-tuning capabilities, ported from Shivam Shrirao's optimized Diffusers implemen… | 42 | 1887 | active |
| laugh12321/TensorRT-YOLO A C++/Python deployment toolkit for running YOLO-family models (YOLOv3 through YOLO26) on NVIDIA GPUs using TensorRT, with custom plugins, … | 63 | 1880 | active |
| nitrain/nitrain Nitrain is a framework-agnostic Python library for sampling, augmenting, and training AI models on medical imaging datasets, with support f… | 23 | 1880 | active |
| react-designer/react-designer A lightweight React library for embedding editable vector graphics canvases in web apps. It supports drawing and manipulating shapes, paths… | 23 | 1877 | active |
| vt-vl-lab/3d-photo-inpainting A Python research codebase from a CVPR 2020 paper that converts a single RGB-D image into a 3D photo using layered depth inpainting. It hal… | 32 | 7093 | maintenance |
| sixthsurge/photon Photon is a gameplay-focused shader pack for Minecraft written in GLSL, adding realistic lighting, shadows, water, clouds, and post-process… | 75 | 1875 | active |
| robertknight/ocrs Ocrs is a Rust library and CLI tool for optical character recognition that extracts text from images such as scanned documents, photos, and… | 69 | 1875 | active |
| zcpua/midjourney-api An unofficial Node.js/TypeScript client library for interacting with the MidJourney image generation service via Discord. It wraps MidJourn… | 65 | 1872 | active |
| we0091234/Chinese_license_plate_detection_recognition A PyTorch-based Chinese license plate detection and recognition system built on YOLOv5 for detection and CRNN for recognition. It supports … | 71 | 1868 | active |
| jingsongliujing/OnnxOCR A lightweight multilingual OCR library rebuilt from PaddleOCR models to run on ONNXRuntime, removing the PaddlePaddle dependency for fast i… | 74 | 1860 | active |
| D-Ogi/WatermarkRemover-AI An AI-powered desktop application that detects and removes watermarks from images and videos using Microsoft's Florence-2 for detection and… | 56 | 1859 | active |
| django-cms/django-filer django-filer is a file and image management application for Django, providing an admin-integrated interface for organizing, uploading, and … | 95 | 1856 | stable |
| facebookresearch/MetaCLIP Meta's research code and models for Meta CLIP, a reimplementation and scaling recipe for CLIP-style contrastive vision-language models, inc… | 82 | 1854 | active |
| gaomingqi/Track-Anything Track-Anything is an interactive tool for video object tracking and segmentation built on Segment Anything, XMem, and E2FGVI. Users specify… | 56 | 6994 | maintenance |
| rosuH/EasyWatermark EasyWatermark is an open-source Android app for adding text or image watermarks to photos, built to protect sensitive images from leaking o… | 72 | 1853 | active |
| thygate/stable-diffusion-webui-depthmap-script An extension for AUTOMATIC1111's Stable Diffusion WebUI that generates high-resolution depth maps from images using models like Marigold, M… | 32 | 1853 | active |
| yihong0618/GitHubPoster A Python CLI tool that turns activity data from many sources (Strava, GPX, LeetCode, Twitter, Bilibili, WakaTime, Kindle, and more) into Gi… | 77 | 1852 | active |
| LuChengTHU/dpm-solver Official PyTorch implementation of DPM-Solver and DPM-Solver++, fast high-order ODE solvers for diffusion probabilistic model sampling that… | 32 | 1852 | stable |
| bootchk/resynthesizer A suite of third-party plugins for GIMP implementing the Resynthesizer algorithm for texture synthesis and transfer among images. It enable… | 70 | 1850 | active |
| CesiumGS/obj2gltf A Node.js tool and library that converts Wavefront OBJ 3D model assets to glTF 2.0, supporting both .gltf and binary .glb output. It handle… | 56 | 1849 | active |
| ConsistentlyInconsistentYT/Pixeltovoxelprojector A Python tool that projects the motion of pixels onto a voxel representation, converting 2D pixel movement into 3D voxel space. It is a pop… | 42 | 1849 | active |
| AutoFigure AutoFigure-Edit is a Python application that converts scientific paper method sections into fully editable SVG figures using large language… | 56 | 1845 | active |
| ivmartel/dwv DWV (DICOM Web Viewer) is an open source, zero-footprint JavaScript/HTML5 library for viewing and manipulating DICOM medical images in any … | 94 | 1844 | active |
| NVlabs/stylegan3 Official PyTorch implementation of StyleGAN3 (Alias-Free GANs), a state-of-the-art generative adversarial network for high-fidelity image s… | 32 | 6943 | maintenance |
| paintdotnet/release The official download repository for Paint.NET, providing offline installer EXEs, portable ZIPs, and deployment MSIs for the Windows image … | 78 | 1842 | active |
| YangLing0818/RPG-DiffusionMaster Official implementation of RPG (Recaption, Plan, Generate), a training-free framework that uses multimodal LLMs as prompt recaptioners and … | 28 | 1842 | active |
| 188080501/JQTools JQTools is an open-source developer toolbox built with Qt/QML/C++ that bundles common small utilities: text processing, hash and encryption… | 90 | 1840 | active |
| Idered/chalk.ist Chalk.ist is a web application for creating beautiful images of source code, similar to tools like Carbon. Users paste code, customize the … | 60 | 1840 | active |
| RanFeng/clipsketch-ai ClipSketch AI is a web-based AI content creation workbench that imports videos from Bilibili and Xiaohongshu links, lets users frame-accura… | 43 | 1840 | active |
| NVIDIA/pix2pixHD PyTorch implementation of pix2pixHD, a conditional GAN method for synthesizing and manipulating high-resolution (2048x1024) photorealistic … | 32 | 6923 | maintenance |
| MaximeRivest/riddle A diary-style application for the reMarkable Paper Pro e-ink tablet that turns handwritten pages into conversations with an LLM. You write … | 74 | 1837 | active |
| AcademySoftwareFoundation/openexr OpenEXR is the specification and reference C/C++ implementation of the EXR file format, the professional high-dynamic-range image storage f… | 99 | 1836 | stable |
| leslievan/semi-utils A Python CLI tool that batch-adds EXIF watermarks (camera model, lens, focal length, aperture, shutter, ISO, shooting time, brand logo) to … | 49 | 1835 | active |
| tdewolff/canvas A Go library providing a common vector drawing target that can output SVG, PDF, EPS, raster images, HTML Canvas via WASM, OpenGL, and Gio. … | 77 | 1833 | active |
| openai/point-e Point-E is OpenAI's official release of models and code for generating 3D point clouds from text prompts or images using diffusion models. … | 32 | 6895 | maintenance |
| timothybrooks/instruct-pix2pix PyTorch implementation of InstructPix2Pix, a diffusion-based model that edits images according to natural language instructions (e.g., 'tur… | 31 | 6885 | maintenance |
| norlab-ulaval/libpointmatcher libpointmatcher is a modular C++ library implementing the Iterative Closest Point (ICP) algorithm for aligning 2D and 3D point clouds, with… | 45 | 1828 | active |
| ZFTurbo/Weighted-Boxes-Fusion A Python library implementing several methods for ensembling bounding boxes from multiple object detection models, including Non-maximum Su… | 65 | 1827 | stable |
| lds133/weather_landscape A Python application that renders weather forecasts as a stylized landscape image instead of numeric dashboards, encoding time, temperature… | 56 | 1827 | active |
| Xanashi/Icaros Icaros is a collection of lightweight Windows Shell Extensions that provide Windows Explorer thumbnails for virtually any FFmpeg-supported … | 87 | 1826 | active |
| Zheng-Chong/CatVTON CatVTON is a lightweight diffusion model for virtual try-on that swaps clothing onto a person image using a concatenation-based architectur… | 40 | 1824 | active |
| octref/polacode Polacode is a VS Code extension that turns code snippets into beautiful, shareable screenshot images using your existing editor theme, gram… | 23 | 6840 | maintenance |
| gkjohnson/three-gpu-pathtracer A GPU-accelerated path tracing renderer for three.js built on three-mesh-bvh and WebGL 2. It provides physically based rendering with GGX m… | 78 | 1817 | active |
| potamides/DeTikZify DeTikZify is a Python library and research tool that uses multimodal large language models to synthesize TikZ/LaTeX graphics programs from … | 49 | 1817 | active |
| andrewsbarbaro/for-the-badge For the Badge is a web application for creating and sharing custom SVG badges with custom text, colors, and icons, famous for its 'badges f… | 75 | 1815 | active |
| Gregwar/Captcha A PHP library for generating CAPTCHA images to protect web forms from bots and spam. It builds distorted text images, exposes the phrase fo… | 83 | 1814 | active |
| protonemedia/laravel-ffmpeg A Laravel package that wraps PHP-FFMpeg to provide an elegant, testable API for video and audio processing, with deep integration into Lara… | 72 | 1814 | active |
| LingyiChen-AI/AIComicBuilder AI Comic Builder is a self-hosted Next.js web application that turns scripts into fully animated comic videos through an automated AI pipel… | 66 | 1814 | active |
| tjko/jpegoptim jpegoptim is a command-line utility for optimizing and compressing JPEG files, offering lossless optimization via Huffman table optimizatio… | 56 | 1809 | active |
| kyechan99/capsule-render capsule-render is a dynamic image generation service that renders colorful decorative header/footer images (waving, cylinder, venom, etc.) … | 75 | 1808 | active |
| apple/ml-4m 4M is a framework from Apple and EPFL for training any-to-any multimodal foundation models using masked modeling over discrete tokens acros… | 35 | 1808 | active |
| all-in-aigc/aicover A full-stack Next.js web application that generates AI-designed cover images using DALL-E 3. It ships with user auth (Clerk), payments (Str… | 27 | 1807 | active |
| rust-skia/rust-skia Safe Rust bindings for Google's Skia 2D graphics library, providing idiomatic Rust access to Skia's C++ API. It supports GPU rendering back… | 86 | 1806 | active |
| picqer/php-barcode-generator A lightweight, framework-independent PHP library that generates 1D barcodes (Code 128, EAN-13, UPC, Code 39, etc.) as SVG, PNG, JPG, or HTM… | 95 | 1805 | stable |
| GordenSun/GordenSuperPPTSkills A set of Codex agent skills that use GPT image generation and vision to create image-format PPT slides and convert them into fully editable… | 52 | 1805 | active |
| triple-mu/YOLOv8-TensorRT A library for running YOLOv8 inference accelerated with NVIDIA TensorRT, supporting detection, segmentation, pose estimation, oriented boun… | 75 | 1804 | active |
| hako-mikan/sd-webui-regional-prompter A custom script extension for AUTOMATIC1111's stable-diffusion-webui that lets users assign different prompts to different regions of a gen… | 55 | 1804 | active |
| ethereal-developers/OpenScan OpenScan is an open-source Android document scanner app built with Flutter that converts photos of documents, notes, and business cards int… | 67 | 1801 | active |
| OpenImagingLab/FlashVSR FlashVSR is a one-step diffusion-based streaming video super-resolution framework that runs at ~17 FPS for 768x1408 video on a single A100 … | 61 | 1799 | active |
| jazzband/sorl-thumbnail sorl-thumbnail is a Django app that generates and manages image thumbnails with pluggable engines (Pillow, ImageMagick, Wand, etc.) and key… | 88 | 1794 | stable |
| Stability-AI/stable-fast-3d Stable Fast 3D (SF3D) is Stability AI's open-source model that reconstructs a textured, UV-unwrapped 3D mesh from a single input image in a… | 24 | 1794 | active |
| TotallyNotChase/glitch-this A Python library and command-line tool that applies customizable glitch effects to images and converts images into glitched GIFs. It offers… | 23 | 1793 | stable |
| Kosinkadink/ComfyUI-VideoHelperSuite A ComfyUI custom node suite providing video workflow nodes such as Load Video, Load Image Sequence, and Video Combine. It converts videos t… | 72 | 1792 | active |
| pencil2d/pencil Pencil2D is a free, open-source desktop application for creating traditional 2D hand-drawn animations using both bitmap and vector graphics… | 89 | 1791 | active |
| thorvg/thorvg ThorVG is a lightweight, production-ready C++ vector graphics engine for rendering SVG, Lottie animations, and vector scenes across embedde… | 99 | 1784 | stable |
| Webreaper/Damselfly Damselfly is a server-based photograph management application designed to index and search very large image collections using metadata such… | 88 | 1783 | active |
| wseemann/FFmpegMediaMetadataRetriever A reimplementation of Android's MediaMetadataRetriever class backed by FFmpeg, providing a unified interface for retrieving frames and meta… | 87 | 1783 | active |
| xyTom/snippai Snippai is an AI-powered snipping tool that captures screenshots and uses AI to extract structured content such as LaTeX formulas, text, ta… | 89 | 1782 | active |
| mdc-ng/mdc-ng A self-hosted media metadata scraper and organizer for adult video libraries, written in Rust with a Next.js web UI. It scrapes metadata fr… | 83 | 1782 | active |
| SchneeHertz/exhentai-manga-manager A Windows desktop application for managing and reading locally downloaded ExHentai manga with tag-based organization. It extracts covers fr… | 89 | 1779 | active |
| huchenlei/ComfyUI-layerdiffuse A ComfyUI custom node plugin implementing LayerDiffuse, enabling generation of transparent images with RGBA output and foreground/backgroun… | 29 | 1779 | active |
| Emu Series Emu3 is a suite of state-of-the-art multimodal models from BAAI trained solely with next-token prediction, tokenizing images, text, and vid… | 57 | 1778 | active |
| nsfw-filter/nsfw-filter A free, open-source, privacy-focused browser extension that blocks NSFW images using on-device AI classification with TensorFlow.js. It hid… | 85 | 1775 | active |
| ianzhao/textshot TextShot is a Python command-line tool that lets you draw a rectangle over any screen region and copies the recognized text to your clipboa… | 32 | 1774 | active |
| puffinsoft/jscanify jscanify is an open-source pure JavaScript document scanning library powered by OpenCV.js. It detects and highlights documents in images an… | 73 | 1769 | active |
| DTolm/VkFFT VkFFT is an open-source, GPU-accelerated multidimensional Fast Fourier Transform library supporting Vulkan, CUDA, HIP, OpenCL, Level Zero, … | 57 | 1769 | active |