function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| mazzzystar/Queryable Queryable is an open-source iOS app that runs Apple's MobileCLIP (formerly OpenAI's CLIP) entirely on-device to search your photo album wit… | 62 | 2977 | active |
| mypaint/mypaint MyPaint is a fast, simple open-source painting and drawing program designed for use with Wacom-style graphics tablets. It features an infin… | 62 | 2963 | active |
| wasserth/TotalSegmentator TotalSegmentator is a Python command-line tool that robustly segments over 100 anatomical structures in CT and MR images using deep learnin… | 66 | 2952 | active |
| gre/react-native-view-shot A React Native library that captures a view and saves it as an image, supporting both old and new architectures (Fabric + TurboModules). It… | 94 | 2947 | active |
| sylikc/jpegview JPEGView is a fast, lean, and highly configurable image viewer and editor for Windows with a minimal GUI, supporting JPEG, PNG, WEBP, TIFF,… | 23 | 2942 | active |
| KichangKim/DeepDanbooru DeepDanbooru is a Python/TensorFlow system that estimates Danbooru-style tags for anime-style girl images using multi-label classification.… | 63 | 2937 | active |
| hero8152/Infinite-Canvas An infinite canvas desktop application for orchestrating AI image, video, and LLM generation workflows. It integrates with ComfyUI, OpenAI-… | 57 | 2934 | active |
| signintech/gopdf gopdf is a Go library for programmatically generating PDF documents. It supports Unicode font embedding, drawing shapes and images, text al… | 75 | 2932 | active |
| ArtifexSoftware/mupdf MuPDF is a lightweight open-source C library and toolkit for viewing, rendering, parsing, and converting PDF, XPS, and e-book documents. It… | 77 | 2930 | stable |
| xelatihy/yocto-gl Yocto/GL is a collection of small C++17 libraries for building physically-based graphics algorithms, written in a data-oriented style and r… | 23 | 2925 | active |
| kozakdenys/qr-code-styling A JavaScript/TypeScript library for generating styled QR codes with custom colors, gradients, dot shapes, corner styles, and embedded logos… | 64 | 2924 | active |
| Tencent-Hunyuan/HunyuanWorld-1.0 Tencent HunyuanWorld-1.0 is an open-source 3D world generation model that creates immersive, explorable, and interactive 3D worlds from tex… | 53 | 2918 | active |
| gildor2/UEViewer UE Viewer (umodel) is a desktop application for viewing and exporting visual resources (meshes, animations, textures, sounds) from games bu… | 23 | 2917 | active |
| ogkalu2/comic-translate An AI-powered application and browser extension that automatically translates comics, manga, manhwa, webtoons, and BDs across many language… | 90 | 2911 | active |
| SimpleSoftwareIO/simple-qrcode Simple QrCode is an easy-to-use PHP library for generating QR codes, built as a wrapper around Bacon/BaconQrCode. It offers first-party int… | 23 | 2903 | active |
| idootop/MagicMirror MagicMirror is a desktop application for instant AI face swapping in photos, built with Tauri. It runs entirely offline on standard hardwar… | 37 | 2891 | active |
| maptiler/tileserver-gl TileServer GL is an open-source map tile server that serves vector and raster map tiles from MBTiles files with GL styles. It supports serv… | 91 | 2887 | active |
| kane50613/takumi Takumi is a Rust-based rendering engine that converts JSX, HTML, and CSS into images (PNG, JPEG, WebP, SVG, animated GIF/WebP) and paged PD… | 81 | 2886 | active |
| nimiq/qr-scanner A lightweight JavaScript/TypeScript QR code scanner library based on Cosmo Wolfe's port of Google's ZXing library. It supports webcam video… | 32 | 2886 | stable |
| spatie/image-optimizer A PHP library that optimizes PNG, JPG, WEBP, AVIF, SVG and GIF images by running them through a chain of installed optimization binaries li… | 81 | 2876 | stable |
| BatchDrake/SigDigger SigDigger is a free Qt-based digital signal analyzer for analyzing unknown radio signals with software-defined radio devices via SoapySDR. … | 53 | 2876 | active |
| insidegui/AssetCatalogTinkerer A macOS application for opening Apple asset catalog (.car) files and browsing, copying, or exporting the images inside them. It also ships … | 58 | 2874 | active |
| NVlabs/FoundationStereo FoundationStereo is NVIDIA's official PyTorch implementation of a foundation model for zero-shot stereo depth estimation, published as a CV… | 47 | 2874 | active |
| duongductrong/Snapzy Snapzy is a free, open-source native macOS app for screenshots, screen recording, annotation, and OCR, built with SwiftUI, AppKit, and Scre… | 78 | 2870 | active |
| minimagick/minimagick MiniMagick is a lightweight Ruby wrapper around the ImageMagick command-line tool, serving as a memory-efficient alternative to RMagick. It… | 97 | 2863 | stable |
| pissang/claygl ClayGL is a WebGL graphics library for building scalable Web3D applications in the browser. It offers a modular, tree-shakeable API (down t… | 48 | 2863 | stable |
| deforum/sd-webui-deforum Deforum is the official extension for AUTOMATIC1111's Stable Diffusion webui that generates AI animations from text prompts using keyframed… | 23 | 2859 | active |
| UX-Decoder/Semantic-SAM Official PyTorch implementation of Semantic-SAM, a universal image segmentation model that segments and recognizes anything at any desired … | 33 | 2854 | active |
| joye61/pic-smaller Pic Smaller is a free, open-source batch image compressor that runs entirely in the browser, supporting JPEG, PNG, WebP, GIF, SVG, AVIF, an… | 69 | 2852 | active |
| microsoft/DirectXTK The DirectX Tool Kit (DirectXTK) is a collection of helper classes for writing Direct3D 11 C++ code in Win32 desktop, UWP, and Xbox One app… | 85 | 2851 | stable |
| openmv/openmv OpenMV is an open-source machine vision platform consisting of camera hardware firmware programmable in Python 3 (MicroPython). The firmwar… | 91 | 2850 | active |
| bytedeco/javacpp-presets JavaCPP Presets provides Java bindings for commonly used native C++ libraries such as OpenCV, FFmpeg, and many others, built on the JavaCPP… | 86 | 2850 | active |
| TMElyralab/MuseV MuseV is a diffusion-based framework for generating high-fidelity virtual human videos of infinite length using a Visual Conditioned Parall… | 25 | 2846 | active |
| pmndrs/postprocessing A post processing library for three.js that adds fullscreen image effects like bloom to WebGL scenes via passes and effects managed by an E… | 98 | 2835 | active |
| metadata-extractor A Java library (with a .NET port) for reading metadata such as Exif, IPTC, XMP, and ICC profiles from image, video, and audio files. It sup… | 87 | 2825 | stable |
| openalpr/openalpr OpenALPR is an open-source Automatic License Plate Recognition (ALPR) library written in C++ that analyzes images and video streams to dete… | 23 | 11452 | maintenance |
| eliemichel/MapsModelsImporter A Blender add-on that imports 3D models of buildings and terrain captured from Google Maps (and Google Earth, Mapy CZ) by recording WebGL G… | 23 | 2809 | active |
| google-research/kubric Kubric is a data generation pipeline from Google Research for creating semi-realistic synthetic multi-object videos with rich annotations l… | 60 | 2808 | active |
| microsoft/MoGe MoGe is a deep learning model from Microsoft Research that recovers 3D geometry from a single open-domain image, predicting metric point ma… | 66 | 2807 | active |
| pikepdf/pikepdf pikepdf is a Python library for reading, writing, repairing, and transforming PDF files, built on the mature qpdf C++ library. It supports … | 95 | 2798 | active |
| danbooru/danbooru Danbooru is a taggable image board web application written in Ruby on Rails, the software behind the well-known anime image board. It suppo… | 67 | 2791 | active |
| dimsemenov/Magnific-Popup Magnific Popup is a fast, light, and responsive lightbox and modal dialog plugin for jQuery and Zepto.js. It can display images, image gall… | 23 | 11314 | maintenance |
| numz/ComfyUI-SeedVR2_VideoUpscaler The official ComfyUI integration of ByteDance's SeedVR2 model for high-quality video and image upscaling, provided as custom nodes. It can … | 54 | 2786 | active |
| gcacace/android-signaturepad An Android library providing a custom SignaturePad View for capturing smooth, handwritten signatures. It uses variable-width Bézier curve i… | 84 | 2785 | stable |
| imanoop7/Ollama-OCR A Python package and Streamlit web app that performs OCR on images and PDFs using vision language models served through Ollama. It supports… | 26 | 2780 | active |
| ValentinH/react-easy-crop react-easy-crop is a React component for cropping images and videos with drag, zoom, and rotate interactions. It returns crop dimensions in… | 97 | 2772 | active |
| espressif/esp32-camera Espressif's official camera driver library for ESP32-series SoCs (ESP32, ESP32-S2, ESP32-S3), supporting a wide range of image sensors like… | 89 | 2771 | active |
| ideogram-oss/ideogram4 Ideogram 4 is an open-weight text-to-image foundation model trained from scratch, with inference code and weights released in Python. It fe… | 54 | 2766 | active |
| autodistill/autodistill Autodistill is a Python library that uses large foundation vision models (like Grounding DINO, Grounded SAM, and CLIP) to automatically lab… | 29 | 2763 | active |
| NVlabs/stylegan2 The official TensorFlow implementation of StyleGAN2, NVIDIA's improved style-based generative adversarial network for high-quality uncondit… | 32 | 11184 | maintenance |
| NVIDIA/FastPhotoStyle FastPhotoStyle is NVIDIA's official PyTorch implementation of the ECCV 2018 paper 'A Closed-form Solution to Photorealistic Image Stylizati… | 23 | 11177 | maintenance |
| kha-white/manga-ocr Manga OCR is a Python library providing optical character recognition for Japanese text, focused on Japanese manga. It uses a custom end-to… | 90 | 2758 | stable |
| apple/turicreate Turi Create is a Python library from Apple that simplifies building custom machine learning models for tasks like image classification, obj… | 10 | 11159 | maintenance |
| huggingface/swift-coreml-diffusers A native SwiftUI application demonstrating how to run Stable Diffusion text-to-image generation on-device using Apple's Core ML Stable Diff… | 54 | 2756 | active |
| weserv/images weserv/images is the source code of wsrv.nl, a self-hostable image cache and resize service that manipulates images on-the-fly via URL para… | 75 | 2755 | stable |
| barryvdh/laravel-snappy A Laravel service provider wrapping the KnpLabs Snappy library to generate PDFs and images from HTML via wkhtmltopdf/wkhtmltoimage binaries… | 79 | 2753 | active |
| jiangdongguo/AndroidUSBCamera AUSBC (AndroidUSBCamera) is a flexible UVC (USB video class) camera engine for Android, refactored in Kotlin with native C components. It s… | 23 | 2750 | active |
| voxelmorph/voxelmorph VoxelMorph is a Python library for learning-based image registration and alignment, using unsupervised deep learning to model deformations … | 76 | 2748 | active |
| napari/napari napari is a fast, interactive, multi-dimensional image viewer for Python, built on top of Qt and numpy. It is designed for browsing, annota… | 95 | 2739 | active |
| stevenlovegrove/Pangolin Pangolin is a lightweight, portable C++ utility library for rapid prototyping of 3D, numeric, and video-based programs, providing cross-pla… | 87 | 2739 | stable |
| suzuki-0000/SKPhotoBrowser SKPhotoBrowser is a Swift library for iOS that provides a simple photo browser/viewer inspired by Facebook and Twitter photo browsers. It s… | 86 | 2735 | active |
| hzeller/timg timg is a terminal-based image and video viewer written in C++ that renders images using terminal graphics protocols like Sixel, Kitty, and… | 72 | 2734 | active |
| IfcOpenShell/IfcOpenShell IfcOpenShell is an open source C++ and Python library for parsing and working with Industry Foundation Classes (IFC) building models, inclu… | 95 | 2733 | active |
| ModelTC/LightX2V LightX2V is a lightweight, high-performance inference framework for image and video generation, supporting tasks like text-to-video, image-… | 64 | 2733 | active |
| Audiveris/audiveris Audiveris is an open-source Optical Music Recognition (OMR) application that transcribes scanned sheet music images into symbolic music dat… | 97 | 2727 | active |
| CVCUDA/CV-CUDA CV-CUDA is an open-source GPU-accelerated library of computer vision and image processing operators built on CUDA, with C++ and Python APIs… | 93 | 2718 | active |
| hiukim/mind-ar-js MindAR is a web augmented reality library supporting image tracking and face tracking, written end-to-end in JavaScript with TensorFlow.js.… | 23 | 2718 | active |
| lengstrom/fast-style-transfer A TensorFlow implementation of fast neural style transfer that applies the style of famous paintings to photos and videos in real time. It … | 32 | 10962 | maintenance |
| teslamotors/react-native-camera-kit A high-performance React Native camera library providing cross-platform camera capture, QR/barcode scanning, and face detection for iOS and… | 96 | 2705 | active |
| yuweihao/MambaOut MambaOut is a PyTorch implementation of Gated CNN models from the CVPR 2025 paper 'MambaOut: Do We Really Need Mamba for Vision?', which qu… | 19 | 2704 | stable |
| magic-research/magic-animate MagicAnimate is the official implementation of a CVPR 2024 diffusion-based human image animation framework that animates a reference image … | 44 | 10897 | maintenance |
| xdit-project/xDiT xDiT is a scalable inference engine for Diffusion Transformers (DiTs) that enables parallel deployment across multiple GPUs and machines. I… | 77 | 2699 | active |
| IDEA-Research/T-Rex T-Rex is the official Python API client for T-Rex2, a generic open-set object detection model that combines text and visual prompts to dete… | 48 | 2699 | active |
| Nandaka/PixivUtil2 A Python command-line tool for bulk downloading images from Pixiv and Pixiv FANBOX, with support for downloading by member, tag, bookmark, … | 71 | 2698 | active |
| ComfyUI-Easy-Use ComfyUI-Easy-Use is an efficiency-focused custom nodes integration package for ComfyUI that optimizes and combines popular nodes for faster… | 78 | 2696 | active |
| dynobo/normcap NormCap is an OCR-powered screen-capture application that lets users select a region of the screen and extracts its text to the clipboard i… | 68 | 2695 | active |
| openai/DALL-E The official PyTorch package for the discrete VAE (dVAE) component of OpenAI's DALL·E model. It does not include the transformer that gener… | 10 | 10834 | maintenance |
| paulpacifico/shutter-encoder Shutter Encoder is a free, open-source GUI application for video, audio, and image transcoding built on FFmpeg, aimed at video editors and … | 92 | 2688 | active |
| bytedance/InfiniteYou InfiniteYou (InfU) is a research framework from ByteDance for identity-preserved text-to-image generation built on Diffusion Transformers l… | 37 | 2685 | active |
| IacobIonut01/ReFra ReFra is a free and open-source media gallery app for Android built with Jetpack Compose and Kotlin. It offers photo and video browsing and… | 97 | 2680 | active |
| MrGiovanni/UNetPlusPlus Official implementation of UNet++, a nested U-Net architecture for medical image segmentation, in both Keras and PyTorch. It redesigns skip… | 77 | 2679 | stable |
| tsayen/dom-to-image A JavaScript library that converts arbitrary DOM nodes into vector (SVG) or raster (PNG/JPEG) images using HTML5 canvas and SVG serializati… | 32 | 10779 | maintenance |
| civilblur/mazanoke MAZANOKE is a self-hosted, privacy-focused image optimizer that runs entirely in the browser, compressing and converting images on-device w… | 73 | 2677 | active |
| HiLab-git/SSL4MIS A benchmark and code collection of semi-supervised learning methods for medical image segmentation, re-implementing approaches like Mean Te… | 44 | 2676 | active |
| PowerHouseMan/ComfyUI-AdvancedLivePortrait A ComfyUI custom node implementing LivePortrait for fast facial expression editing and animation with real-time preview. It can edit expres… | 23 | 2675 | active |
| immich-power-tools/immich-power-tools An unofficial web client for Immich that adds bulk organization and management tools on top of the standard Immich UI. It helps users manag… | 86 | 2674 | active |
| IceClear/StableSR StableSR is a Python research library that leverages pre-trained Stable Diffusion priors for real-world blind image super-resolution. It pr… | 21 | 2668 | stable |
| aigc3d/LHM LHM is a PyTorch-based large reconstruction model that reconstructs high-fidelity animatable 3D human avatars from a single image in second… | 52 | 2664 | active |
| Nutlope/roomGPT RoomGPT is an open-source Next.js web application that lets users upload a photo of a room and generate redesigned variations using the Con… | 31 | 10671 | maintenance |
| hgmzhn/manga-translator-ui A desktop GUI application built on manga-image-translator that automatically translates text in manga/comic images across Japanese, Korean,… | 80 | 2651 | active |
| phillipi/pix2pix The original Torch (Lua) implementation of pix2pix, a conditional GAN for image-to-image translation tasks such as synthesizing photos from… | 32 | 10652 | maintenance |
| WSTxda/QP-Gallery-Releases A modernized mod of the classic QuickPic Gallery Android app with a refreshed Material 3 design, bug fixes, and compatibility updates for r… | 57 | 2650 | active |
| szTheory/exifcleaner ExifCleaner is a free, open-source cross-platform desktop GUI app that strips EXIF and other metadata from images, videos, and PDFs using E… | 99 | 2644 | active |
| colour-science/colour Colour is an open-source Python library providing a comprehensive collection of colour science algorithms and datasets, including colour sp… | 74 | 2642 | active |
| black-forest-labs/flux2 Official inference repository for Black Forest Labs' FLUX.2 family of open-weight image generation and editing models. It provides minimal … | 48 | 2642 | active |
| mirari/v-viewer v-viewer is an image viewer component and directive for Vue 2 and Vue 3, built on top of viewer.js. It supports rotation, scaling, zooming,… | 62 | 2638 | stable |
| thephpleague/glide Glide is a PHP library for on-demand image manipulation exposed via a simple HTTP-based API, similar to cloud services like Imgix and Cloud… | 90 | 2632 | stable |
| swz30/Restormer Restormer is an efficient Transformer architecture for high-resolution image restoration, published as a CVPR 2022 Oral paper. It provides … | 44 | 2625 | stable |
| huangserva/3DCellForge A React + Three.js web application for AI-powered interactive 3D model generation, inspection, and presentation. It turns uploaded referenc… | 51 | 2624 | active |
| gridaco/grida Grida is an open-source 2D canvas and SVG editor powered by a Rust/Skia/WASM graphics engine, with additional Database/CMS and Forms produc… | 98 | 2622 | active |