function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| lllyasviel/IC-Light IC-Light is a Python tool for manipulating the illumination of images using diffusion models, offering text-conditioned and background-cond… | 27 | 8508 | active |
| DImuthuUpe/AndroidPdfViewer An Android library for displaying PDF documents in apps, built on PdfiumAndroid for decoding. It supports animations, gestures, zoom, and d… | 55 | 8470 | active |
| Nutlope/logocreator An open-source AI logo generator web app that creates brand-ready logos using FLUX models on Together AI, with logo editing, style presets,… | 64 | 8435 | active |
| lucidrains/imagen-pytorch A PyTorch implementation of Imagen, Google's text-to-image neural network based on cascading DDPMs conditioned on T5 text embeddings. It pr… | 23 | 8424 | active |
| photopea/photopea Photopea is a free online image editor for raster and vector graphics that runs entirely in the browser, supporting PSD, AI, Sketch, and do… | 45 | 8405 | active |
| CASIA-LMC-Lab/FastSAM FastSAM is a CNN-based Segment Anything Model trained on only 2% of the SA-1B dataset, achieving comparable segmentation performance to SAM… | 19 | 8401 | active |
| XPixelGroup/BasicSR BasicSR is an open-source PyTorch toolbox for image and video restoration tasks such as super-resolution, denoising, deblurring, and JPEG a… | 23 | 8367 | stable |
| PDFCraftTool/pdfcraft PDFCraft is a free, privacy-focused PDF toolkit with 90+ tools (merge, split, compress, convert, edit, secure) that runs entirely in the br… | 75 | 8340 | active |
| bytedeco/javacv JavaCV is a Java library that wraps OpenCV, FFmpeg, and other computer vision and multimedia libraries via JavaCPP Presets, with utility cl… | 86 | 8335 | active |
| vietnh1009/ASCII-generator A Python tool that converts images and videos into ASCII art, outputting text files or image/video files in grayscale or color. It supports… | 32 | 8318 | stable |
| LibreSprite/LibreSprite LibreSprite is a free, open-source animated sprite editor and pixel art tool, forked from the last GPLv2 commit of Aseprite. It supports la… | 62 | 8298 | active |
| exadel-inc/CompreFace Exadel CompreFace is a free, open-source face recognition system that provides REST APIs for face recognition, verification, detection, lan… | 23 | 8273 | stable |
| zeux/meshoptimizer A C/C++ library providing algorithms to optimize triangle meshes for GPU rendering, including vertex cache optimization, overdraw reduction… | 88 | 8268 | stable |
| QwenLM/Qwen-Image Qwen-Image is a 20B MMDiT image generation foundation model from the Qwen team, with strong complex text rendering (especially Chinese) and… | 48 | 8265 | active |
| Viewer.js Viewer.js is a JavaScript library for viewing images in the browser with zoom, rotation, flip, move, and keyboard/touch support. It works i… | 95 | 8234 | stable |
| lltcggie/waifu2x-caffe A Windows GUI/CLI application that reimplements the waifu2x image upscaling and noise-reduction tool using the Caffe deep learning framewor… | 56 | 8221 | stable |
| carson-katri/dream-textures A Blender add-on that integrates Stable Diffusion for generating textures, concept art, and background assets directly inside Blender. It s… | 23 | 8195 | active |
| brycedrennan/imaginAIry A Python library and CLI tool (imaginairy/aimg) for generating images and videos with Stable Diffusion and Stable Video Diffusion models. I… | 63 | 8179 | active |
| soldair/node-qrcode A JavaScript QR code generator library that works on the server, in the browser, and in React Native, with support for PNG, SVG, and termin… | 32 | 8160 | stable |
| banchichen/TZImagePickerController TZImagePickerController is an Objective-C library for iOS that clones UIImagePickerController with support for picking multiple photos, ori… | 66 | 8149 | active |
| LibrePhotos/librephotos LibrePhotos is a self-hosted, open-source photo management service built with Django and React. It offers AI-powered features like face rec… | 96 | 8049 | active |
| zumerlab/snapdom SnapDOM is a high-performance, dependency-free browser library that captures DOM elements as self-contained SVG representations and exports… | 86 | 8040 | active |
| SixLabors/ImageSharp ImageSharp is a fully managed, high-performance 2D graphics and image processing library for .NET 8+. It provides loading, resizing, format… | 95 | 8034 | stable |
| Agents365-ai/drawio-skill An agent skill (SKILL.md + Python CLI) that generates professional draw.io diagrams from natural language, codebases, IaC configs, SQL sche… | 82 | 8022 | active |
| nadermx/backgroundremover A free, open-source command line tool that removes backgrounds from images and videos using the U2Net AI model built on PyTorch. It is inst… | 93 | 8020 | active |
| bingoogolapple/BGAQRCode-Android An Android library for scanning and generating QR codes and barcodes, offering both ZXing and ZBar engines behind customizable scan views. … | 64 | 8008 | stable |
| davemorrissey/subsampling-scale-image-view An Android library providing a highly configurable, extendable image view for displaying very large images with deep zoom support. It uses … | 23 | 8006 | stable |
| Lake1059/FFmpegFreeUI FFmpegFreeUI (3FUI) is a free, open-source Windows GUI front-end for FFmpeg aimed at advanced users who want full control over encoding par… | 81 | 8002 | active |
| software-mansion/react-native-svg A React Native library providing SVG rendering support on iOS, Android, macOS, Windows, and the web via React Native Web. It supports most … | 89 | 8000 | stable |
| PixiEditor/PixiEditor PixiEditor is a free, open-source, cross-platform desktop 2D graphics editor built in C# with Avalonia. It combines pixel art, painting, an… | 99 | 7999 | active |
| ikuaitu/vue-fabric-editor An open-source web-based image and graphic editor built on fabric.js and Vue 3 with a plugin-based architecture. It supports custom fonts, … | 65 | 7946 | active |
| bestony/logoly Logoly is a web application that generates parody logos in the style of Pornhub or OnlyFans, with customizable text, colors, and font sizes… | 73 | 7941 | active |
| MochiDiffusion/MochiDiffusion A native macOS app built with SwiftUI for running Stable Diffusion and FLUX.2 Klein image generation locally on Apple Silicon Macs. It uses… | 78 | 7926 | active |
| verlok/vanilla-lazyload A lightweight (2.4 kB) vanilla JavaScript library that speeds up websites by deferring the loading of below-the-fold images, backgrounds, v… | 62 | 7855 | stable |
| open-mmlab/mmpose MMPose is an open-source pose estimation toolbox and benchmark built on PyTorch as part of the OpenMMLab ecosystem. It provides implementat… | 39 | 7855 | active |
| TencentARC/GFPGAN GFPGAN is a Python library built on PyTorch that restores and enhances real-world degraded face photos using GAN-based priors. It provides … | 23 | 37657 | maintenance |
| TadasBaltrusaitis/OpenFace OpenFace is a facial behavior analysis toolkit that performs facial landmark detection, head pose estimation, facial action unit recognitio… | 23 | 7740 | active |
| geekyutao/Inpaint-Anything Inpaint Anything combines Segment Anything (SAM) with inpainting models like LaMa and Stable Diffusion to remove, fill, or replace objects … | 65 | 7703 | active |
| CeuiLiSA/Pixiv-Shaft PixShaft (Shaft) is an open-source third-party Pixiv client for Android built in Kotlin. It provides access to illustrations, manga, novels… | 95 | 7691 | active |
| lllyasviel/Omost Omost is a Python application that converts LLM coding capability into image composition by having pretrained LLMs (based on Llama3 and Phi… | 24 | 7610 | active |
| RapidAI/RapidOCR RapidOCR is an open-source, multi-language OCR toolkit that performs text detection and recognition using models converted to run on ONNX R… | 97 | 7599 | active |
| opentoonz/opentoonz OpenToonz is a free, open-source, full-featured 2D animation production application based on the Toonz software customized by Studio Ghibli… | 96 | 7594 | active |
| 1adrianb/face-alignment A Python library built on PyTorch that detects 2D and 3D facial landmarks in images using the FAN deep learning face alignment network. It … | 70 | 7538 | active |
| phoboslab/qoi QOI is the 'Quite OK Image Format', a fast, lossless image compression format with a single-file MIT-licensed C/C++ reference implementatio… | 70 | 7522 | stable |
| hooke007/mpv_PlayKit A curated mpv media player configuration package for Windows (x64), bundling Chinese-annotated config files, shader/filter integration, and… | 83 | 7500 | active |
| ApoorvSaxena/lozad.js Lozad.js is a lightweight (~1kb) pure JavaScript lazy-loading library with no dependencies, built on the IntersectionObserver and MutationO… | 46 | 7494 | stable |
| hybridgroup/gocv GoCV is a Go language binding for the OpenCV 4 computer vision library, supporting Linux, macOS, Windows, and Docker. It includes support f… | 76 | 7491 | active |
| google/draco Draco is an open-source C++ library from Google for compressing and decompressing 3D geometric meshes and point clouds. It improves the sto… | 67 | 7458 | stable |
| EutropicAI/Final2x Final2x is a cross-platform desktop application for image super-resolution (upscaling) built with Electron, Vue3, and a PyTorch-based Pytho… | 74 | 7323 | active |
| facebookresearch/sam-3d-objects SAM 3D Objects is a foundation model from Meta that reconstructs full 3D shape geometry, texture, and layout from a single image, with code… | 55 | 7322 | active |
| vladmandic/sdnext SD.Next is an open-source, self-hosted WebUI server application for AI generative image and video creation built on Stable Diffusion and Di… | 76 | 7320 | active |
| naver/dust3r DUSt3R is the official PyTorch implementation of a CVPR 2024 model that performs dense, unconstrained stereo and multi-view 3D reconstructi… | 45 | 7288 | active |
| AbdullahAlfaraj/Auto-Photoshop-StableDiffusion-Plugin A Photoshop plugin (built on Adobe UXP) that lets users generate Stable Diffusion images directly inside Photoshop, using Automatic1111 Web… | 22 | 7288 | active |
| imgly/background-removal-js An npm package (browser and Node.js variants) that removes image backgrounds using ONNX-based image segmentation/matting models running ent… | 43 | 7287 | active |
| lightningpixel/modly Modly is an open-source desktop application that generates textured 3D meshes from images or text prompts using open-source AI models (Huny… | 80 | 7275 | active |
| liuliu/ccv ccv is a modern, minimalist computer vision library written in C/C++ with an application-driven set of state-of-the-art algorithms includin… | 77 | 7243 | active |
| zetbaitsu/Compressor Compressor is a lightweight Android image compression library written in Kotlin. It lets developers shrink large photos into smaller files … | 54 | 7227 | stable |
| princeton-vl/infinigen Infinigen is a procedural generator of infinite photorealistic 3D worlds and scenes, built on Blender by the Princeton Vision & Learning La… | 94 | 7222 | active |
| bubkoo/html-to-image A TypeScript library that generates images (PNG, JPEG, SVG, blob, canvas, or pixel data) from DOM nodes using HTML5 canvas and SVG foreignO… | 72 | 7222 | stable |
| kohya-ss/sd-scripts A collection of Python training, generation, and utility scripts for Stable Diffusion and other image generation models, most widely used f… | 89 | 7210 | active |
| PeterL1n/BackgroundMattingV2 Official PyTorch implementation of the CVPR 2021 paper 'Real-Time High-Resolution Background Matting'. It produces state-of-the-art alpha m… | 23 | 7189 | stable |
| TagStudioDev/TagStudio TagStudio is a cross-platform desktop application for organizing photos and files using a tag-based system. It stores tags and metadata in … | 91 | 7146 | active |
| ControlNet ControlNet is a neural network architecture that adds conditional control (edges, poses, depth, etc.) to pretrained text-to-image diffusion… | 31 | 34091 | maintenance |
| zxing/zxing ZXing ('Zebra Crossing') is an open-source, multi-format 1D/2D barcode image processing library implemented in Java, with ports to other la… | 73 | 34077 | maintenance |
| yangchris11/samurai SAMURAI is the official implementation of a zero-shot visual object tracker built on top of Segment Anything Model 2 (SAM 2), using a motio… | 27 | 7112 | active |
| chenfei-wu/TaskMatrix TaskMatrix (Visual ChatGPT) is a Python framework that connects ChatGPT with a suite of visual foundation models like Stable Diffusion, Gro… | 30 | 34003 | maintenance |
| pixelfed/pixelfed Pixelfed is a free, open-source, decentralized photo-sharing social network built on Laravel, compatible with the ActivityPub federation pr… | 91 | 7071 | active |
| lightGallery lightGallery is a customizable, modular, dependency-free JavaScript lightbox gallery plugin for displaying images and videos on the web and… | 72 | 7049 | active |
| xushengfeng/eSearch eSearch is a cross-platform desktop application (Electron) combining screenshot capture, offline OCR based on PaddleOCR, screen search, tra… | 98 | 7036 | active |
| dwzhu-pku/PaperBanana PaperBanana is a reference-driven multi-agent framework that automatically generates publication-ready academic illustrations such as metho… | 55 | 7000 | active |
| hect0x7/JMComic-Crawler-Python A Python library providing an API client for the JMComic (18comic) site, supporting both web and mobile endpoints, with album downloading, … | 97 | 6997 | stable |
| latentcat/qrbtf QRBTF is an AI and parametric QR code generator that creates stylized, scannable QR codes, available as a hosted web app at qrbtf.com with … | 30 | 6996 | active |
| PaddlePaddle/models PaddlePaddle's officially maintained industry-grade model repository containing 600+ models across computer vision, NLP, speech, recommenda… | 23 | 6932 | active |
| mapbox/pixelmatch A tiny, dependency-free JavaScript library for pixel-level image comparison, originally built for comparing screenshots in tests. It works … | 89 | 6931 | stable |
| sczhou/ProPainter ProPainter is a PyTorch-based video inpainting model from ICCV 2023 that combines dual-domain propagation with a mask-guided sparse video T… | 22 | 6916 | stable |
| VAST-AI-Research/TripoSR TripoSR is an open-source model for fast feedforward 3D object reconstruction from a single image, developed by Tripo AI and Stability AI. … | 64 | 6888 | active |
| leejet/stable-diffusion.cpp A pure C/C++ inference engine for diffusion models (Stable Diffusion, FLUX, Wan, Qwen Image, Z-Image, and more) built on ggml in the style … | 91 | 6846 | active |
| Dooy/chatgpt-web-midjourney-proxy A unified web/desktop UI for ChatGPT plus AI image, music, and video generation services like Midjourney, Suno, Luma, Runway, and Flux. It … | 76 | 6785 | active |
| opengeos/GeoLibre GeoLibre is a free, open-source, lightweight cloud-native GIS platform for visualizing, exploring, and analyzing geospatial data. Built wit… | 81 | 6772 | active |
| mbrlabs/Lorien Lorien is an open-source infinite canvas drawing and note-taking application built with the Godot game engine. It stores brush strokes as v… | 42 | 6771 | active |
| visioncortex/vtracer VTracer is an open-source tool and library that converts raster images (JPG, PNG) into vector graphics (SVG), handling both colored images … | 98 | 6745 | active |
| niklasvh/html2canvas html2canvas is a JavaScript library that renders webpages or DOM elements as canvas images entirely in the browser, without server-side ren… | 23 | 31923 | maintenance |
| nayuki/QR-Code-generator A high-quality QR Code generator library implemented independently from the ISO/IEC 18004 (QR Code Model 2) specification, available in Jav… | 79 | 6728 | stable |
| ParthJadhav/app-store-screenshots An AI agent skill that scaffolds a production-ready Next.js screenshot editor for generating App Store and Google Play marketing screenshot… | 56 | 6706 | active |
| tencent-ailab/IP-Adapter IP-Adapter is a lightweight (22M parameter) adapter that adds image prompt capability to pretrained text-to-image diffusion models like Sta… | 28 | 6677 | stable |
| LiamGvchi/gc-minimal-zine-poster A Codex skill that turns themes, sentences, photos, or reference sets into minimal zine-style editorial poster prompts and generated images… | 78 | 6668 | active |
| FoundationVision/ByteTrack ByteTrack is a PyTorch-based multi-object tracking (MOT) library implementing the ECCV 2022 paper 'Multi-Object Tracking by Associating Eve… | 32 | 6654 | stable |
| op7418/guizang-social-card-skill A Claude Code / Codex agent skill that generates Xiaohongshu (Rednote) carousel images and WeChat 21:9 + 1:1 cover pairs from articles, not… | 54 | 6635 | active |
| halide/Halide Halide is an embedded DSL (in C++ and Python) for writing high-performance, data-parallel image and array processing pipelines. It separate… | 70 | 6590 | stable |
| 11cafe/jaaz Jaaz is an open-source multimodal creative assistant application that serves as a privacy-focused, locally usable alternative to Canva and … | 53 | 6589 | active |
| ml5js/ml5-library ml5.js is a friendly, beginner-oriented JavaScript machine learning library for the browser, built on top of TensorFlow.js. It provides acc… | 23 | 6587 | active |
| scikit-image/scikit-image scikit-image is a Python library providing a collection of peer-reviewed image processing algorithms built on NumPy and SciPy. It offers ro… | 82 | 6577 | stable |
| iperov/DeepFaceLive DeepFaceLive is a real-time face-swap application for PC streaming and video calls, using trained face models (DFM) applied to webcam or vi… | 10 | 31011 | maintenance |
| openMVG/openMVG OpenMVG is a C++ library for multiple view geometry and Structure from Motion (SfM), providing end-to-end 3D reconstruction from images. It… | 48 | 6542 | active |
| AILab-CVC/YOLO-World YOLO-World is a real-time open-vocabulary object detection model and Python toolkit from Tencent AI Lab and HUST, published at CVPR 2024. I… | 29 | 6529 | active |
| blueimp/jQuery-File-Upload A jQuery plugin providing a file upload widget with multiple file selection, drag & drop, progress bars, validation, and media previews. It… | 10 | 30717 | maintenance |
| photoview/photoview Photoview is a self-hosted photo gallery web application for personal servers, written in Go. It scans directories of photos and videos, ge… | 67 | 6515 | active |
| bashalarmistalt/decimen-optical-transfer A web application that transfers files between devices using animated QR codes displayed on one screen and captured by another device's cam… | 79 | 6504 | active |
| open-mmlab/mmcv MMCV is the foundational computer vision library for the OpenMMLab ecosystem, providing image/video I/O, data transformations, and CUDA ope… | 52 | 6470 | stable |
| HVision-NKU/StoryDiffusion StoryDiffusion is the official implementation of a NeurIPS 2024 Spotlight paper introducing Consistent Self-Attention for character-consist… | 24 | 6452 | active |