function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| SatDump/SatDump SatDump is a general-purpose satellite data processing application that receives, records, demodulates, and decodes signals from weather sa… | 74 | 2088 | active |
| tumuyan/RealSR-NCNN-Android An Android application for image super-resolution and upscaling built on NCNN and MNN inference engines, bundling models like RealSR, Real-… | 84 | 2087 | active |
| WebAV-Tech/WebAV WebAV is a TypeScript SDK for creating and editing audio/video files entirely in the browser, built on the WebCodecs API. It provides compo… | 58 | 2087 | active |
| 1038lab/ComfyUI-RMBG A ComfyUI custom node package for advanced image background removal and segmentation of objects, faces, clothing, and fashion elements. It … | 66 | 2086 | active |
| ali-vilab/In-Context-LoRA Official repository for In-Context LoRA (IC-LoRA), a framework for adapting Diffusion Transformers to diverse visual generation tasks via L… | 22 | 2083 | active |
| alex-damian/pulse PULSE is a Python research implementation of a CVPR 2020 paper that upscales low-resolution face photos by searching the latent space of a … | 32 | 8023 | maintenance |
| vapoursynth/vapoursynth VapourSynth is a video processing framework with a C++ core library and a Python module for writing video processing scripts. It supports m… | 94 | 2079 | active |
| hilongjw/vue-lazyload A lightweight Vue.js plugin that lazy-loads images and components via a v-lazy directive, with loading/error placeholders and CSS state cla… | 23 | 7993 | maintenance |
| ascorbic/unpic-img unpic-img is a cross-framework responsive image component library for React, Vue, Svelte, Astro, Angular, SolidJS, and more. It generates c… | 84 | 2075 | active |
| DanBloomberg/leptonica Leptonica is an open-source C library providing a broad set of image processing and image analysis operations, with a focus on document ima… | 74 | 2074 | stable |
| MiniMax-AI/cli The official CLI for the MiniMax AI Platform, written in TypeScript, that generates text, images, video, speech, and music from the termina… | 81 | 2073 | active |
| GPUOpen-Effects/FidelityFX-FSR2 AMD FidelityFX Super Resolution 2 (FSR 2) is an open-source, high-quality temporal upscaling solution that reconstructs high-resolution fra… | 23 | 2073 | stable |
| esimov/triangle A Go CLI tool and library that converts images into abstract computer-generated art using Delaunay triangulation. It blurs, grayscales, and… | 23 | 2069 | active |
| facebookresearch/ConvNeXt-V2 Official PyTorch implementation of ConvNeXt V2, a family of pure convolutional neural network models co-designed with a fully convolutional… | 10 | 2069 | stable |
| Flipboard/FLAnimatedImage FLAnimatedImage is a performant animated GIF engine for iOS that plays GIFs with desktop-browser-like speed while handling variable frame d… | 23 | 7949 | maintenance |
| dai-shi/excalidraw-animate A web-based tool and npm package that converts Excalidraw drawings into animations, with configurable per-element animation order and durat… | 72 | 2066 | active |
| 3rd/image.nvim A Neovim plugin written in Lua that adds image display support inside the editor using Kitty's Graphics Protocol, ueberzugpp, or Sixel back… | 81 | 2064 | active |
| rlxone/Equinox Equinox is a free, open-source native macOS app for creating dynamic wallpapers such as Dynamic Desktop and Light & Dark Desktop. Users cho… | 75 | 2054 | active |
| marcoslucianops/DeepStream-Yolo A collection of configuration files, parsers, and conversion utilities for running YOLO-family object detection models on NVIDIA DeepStream… | 61 | 2054 | active |
| visomaster/VisoMaster VisoMaster is a Python-based desktop application for AI-powered face swapping and face editing in images and videos. It supports multiple s… | 27 | 2052 | active |
| extesy/hoverzoom Hover Zoom+ is an open-source browser extension that enlarges images and videos to full size when you hover your mouse over them on support… | 93 | 2051 | active |
| discord/lilliput A Go library for resizing and transcoding images, backed by mature C libraries (JPEG, PNG, WebP, AVIF, animated GIF) via cgo. It minimizes … | 74 | 2051 | active |
| Sygil-Dev/sygil-webui A browser-based web UI for generating images with Stable Diffusion, built in Python with Gradio and Streamlit frontends. It supports text-t… | 74 | 7870 | maintenance |
| alganzory/HaramBlur HaramBlur is a browser extension that automatically detects and blurs inappropriate images and videos on web pages using on-device machine … | 18 | 2048 | active |
| storytold/artcraft ArtCraft is an open-source desktop application for interactive AI image and video creation, described as 'the IDE for artists'. It provides… | 94 | 2044 | active |
| PRIS-CV/DemoFusion DemoFusion is a CVPR 2024 framework that extends open-source latent diffusion models like SDXL to generate high-resolution images without a… | 48 | 2041 | stable |
| emilianavt/OpenSeeFace OpenSeeFace is a robust realtime face and facial landmark tracking library that runs on CPU at 30-60 fps using ONNX-converted MobileNetV3 m… | 49 | 2038 | active |
| deep-floyd/IF DeepFloyd IF is an open-source text-to-image model library implementing a cascaded pixel diffusion architecture with a frozen T5 text encod… | 22 | 7804 | maintenance |
| YUZU-Hub/appscreen A free, open-source browser-based tool for creating App Store screenshots with customizable backgrounds, text overlays, and 2D/3D device mo… | 52 | 2032 | active |
| jaywcjlove/DevHub DevHub is an offline, local-first developer toolbox application for macOS built with SwiftUI, bundling 100+ everyday utilities such as JSON… | 71 | 2027 | active |
| serengil/retinaface RetinaFace is a Python library for deep learning based face detection, built on TensorFlow and derived from the insightface project's Retin… | 61 | 2027 | active |
| renzhezhilu/webp2jpg-online A browser-based, pure front-end image format converter that converts between jpeg, png, gif, webp, svg, ico, bmp, psd, heic and more withou… | 23 | 2024 | stable |
| TianZerL/Anime4KCPP Anime4KCPP is a high-performance anime image and video upscaler built on CNN-based algorithms, written in C++. It ships as a library plus V… | 78 | 2022 | active |
| RexanWONG/text-behind-image An open-source web application for creating text-behind-image designs, where text appears layered behind the subject of a photo. It is avai… | 62 | 2020 | active |
| XavierXiao/Dreambooth-Stable-Diffusion An implementation of Google's Dreambooth fine-tuning method applied to Stable Diffusion, enabling personalization of a text-to-image diffus… | 32 | 7738 | maintenance |
| instantX-research/InstantStyle InstantStyle is a framework for style-preserving text-to-image generation that disentangles style and content from reference images using f… | 26 | 2018 | active |
| NVlabs/SPADE Official PyTorch implementation of SPADE (GauGAN), a CVPR 2019 method for synthesizing photorealistic images from semantic segmentation map… | 32 | 7717 | maintenance |
| sindresorhus/capture-website A Node.js library for capturing screenshots of websites using Puppeteer (headless Chrome) under the hood. It supports saving screenshots to… | 61 | 2013 | active |
| julyx10/lap Lap is an open-source, local-first desktop photo manager for macOS, Windows, and Linux built for large personal photo libraries. It offers … | 89 | 2012 | active |
| webp-sh/webp_server_go A Go-based HTTP server that serves JPEG, PNG, BMP, GIF, SVG and other images as WebP/AVIF/JXL on the fly, without changing the original URL… | 95 | 2006 | active |
| nhn/tui.image-editor TOAST UI Image Editor is a full-featured photo image editor built on HTML5 Canvas, providing crop, flip, rotate, draw, shape, text, mask, a… | 23 | 7669 | maintenance |
| AntixK/PyTorch-VAE A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f… | 38 | 7665 | maintenance |
| flytkgl/PDFQFZ PDFQFZ is a small desktop tool for adding cross-page (riding) seals to PDF documents. It takes a full seal image, randomly splits it across… | 90 | 2003 | active |
| Badge Magic Badge Magic is a cross-platform mobile and desktop app for creating text, drawings, and animations on LED name badges and transferring them… | 80 | 2001 | active |
| WhatDreamsCost/WhatDreamsCost-ComfyUI A collection of free custom ComfyUI nodes and workflows, centered on LTX Director, a timeline-based tool for directing LTX video generation… | 57 | 2001 | active |
| bytetriper/RAE Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u… | 48 | 2001 | active |
| fogleman/sdf A Python library for generating 3D meshes from signed distance functions (SDFs) with a simple, operator-based API supporting constructive s… | 32 | 2000 | stable |
| DEIM DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp… | 62 | 1999 | active |
| whomwah/rqrcode RQRCode is a Ruby library for generating QR codes with a simple interface exposing standard QR code options like error correction level, si… | 77 | 1998 | stable |
| bluefireteam/photo_view A Flutter library providing a customizable zoomable image widget with gesture support for pinch, rotate, and drag. It can also display arbi… | 23 | 1997 | stable |
| fex-team/webuploader WebUploader is a JavaScript file upload component that uses HTML5 as its primary runtime with a Flash fallback for legacy browsers like IE6… | 10 | 7634 | maintenance |
| facebookresearch/dino PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th… | 10 | 7611 | maintenance |
| NVIDIA/Stable-Diffusion-WebUI-TensorRT An NVIDIA extension for the Stable Diffusion Web UI (Automatic1111) that accelerates image generation using TensorRT-optimized engines on R… | 18 | 1989 | active |
| diana7127/mpv.net-DW A personally customized Windows build of mpv.net (based on mpv.net_CM and mpv_lazy) that bundles a ModernX playback UI, thumbfast seekbar t… | 23 | 1988 | active |
| zxing-cpp/zxing-cpp ZXing-C++ is an open-source, multi-format 1D/2D barcode image processing library written in pure C++20, ported from the Java ZXing library … | 97 | 1987 | active |
| svg-sprite/svg-sprite A low-level Node.js module that takes a batch of SVG files, optimizes them with SVGO, and generates SVG sprites of several types (CSS sprit… | 54 | 1986 | stable |
| LinwoodDev/Butterfly Butterfly is a powerful, minimalistic, open-source note-taking app built with Flutter that supports drawing, handwriting, and rich text on … | 98 | 1984 | active |
| manisandro/gImageReader gImageReader is a graphical GTK/Qt front-end to the tesseract-ocr engine for recognizing text in images, PDFs, scans, and screenshots. It s… | 51 | 1984 | active |
| op7418/logo-generator-skill A Claude Code skill that generates professional SVG logos in 6+ design variants and produces high-end showcase images using Gemini 3.1 Flas… | 49 | 1984 | active |
| xingyizhou/CenterNet CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi… | 32 | 7573 | maintenance |
| antirez/iris.c Iris is a pure C inference pipeline that generates images from text prompts using open-weights diffusion transformer models like FLUX.2 Kle… | 46 | 1983 | active |
| hkchengrex/XMem XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in… | 23 | 1983 | stable |
| Belval/pdf2image A Python module that wraps the pdftoppm and pdftocairo command-line utilities (from Poppler) to convert PDF pages into PIL Image objects. I… | 23 | 1982 | stable |
| laravolt/avatar A PHP/Laravel package that generates avatar images from names, emails, or arbitrary strings, producing initials-based avatars as base64, PN… | 92 | 1981 | active |
| blend2d/blend2d Blend2D is a high-performance 2D vector graphics engine written in C++ that uses a built-in JIT compiler (via AsmJit) to generate optimized… | 57 | 1981 | active |
| thx/resvg-js A high-performance SVG renderer and toolkit for Node.js, Deno, and browsers, built on the Rust resvg library via napi-rs with a WebAssembly… | 82 | 1980 | active |
| patrikhuber/eos A lightweight, header-only 3D Morphable Face Model (3DMM) fitting library written in modern C++11/14, with Python bindings. It provides mod… | 31 | 1980 | active |
| gcui-art/markdown-to-image A React component library that renders Markdown into visually appealing poster images optimized for social media sharing, with support for … | 28 | 1980 | active |
| JIA-Lab-research/DreamOmni2 DreamOmni2 is the official PyTorch implementation of a CVPR 2026 Highlight model for multimodal instruction-based image editing and generat… | 51 | 1978 | active |
| Linzaer/Ultra-Light-Fast-Generic-Face-Detector-1MB An ultra-lightweight face detection model (~1MB FP32, ~300KB quantized) designed for edge computing devices, with slim and RFB variants tra… | 32 | 7542 | maintenance |
| tandpfun/wardrobe A self-hosted web application that detects garments in photos, extracts clean product cutouts, and generates modeled editorial previews usi… | 54 | 1975 | active |
| showlab/Show-o Show-o is a research repository implementing a unified transformer model that combines autoregressive and discrete diffusion modeling for m… | 50 | 1973 | active |
| mborgerding/kissfft KISS FFT is a mixed-radix Fast Fourier Transform library written in C that supports fixed-point and floating-point data types. It is design… | 71 | 1972 | stable |
| android/androidify An open-source Android sample app from Google that lets users create custom Android bot avatars using AI image generation via the Gemini AP… | 68 | 1972 | active |
| crystian/ComfyUI-Crystools A collection of utility custom nodes and UI extensions for ComfyUI, including real-time CPU/GPU/RAM resource monitors, progress bars, and m… | 39 | 1972 | active |
| theamusing/perfectPixel A Python library that automatically detects the optimal grid size in AI-generated pixel art images and refines them into clean, perfectly a… | 45 | 1971 | active |
| astrofox-io/astrofox Astrofox is a free, open-source motion graphics application that turns audio into audio-reactive visuals, available as both a web app and a… | 67 | 1970 | active |
| Netflix/void-model VOID (Video Object and Interaction Deletion) is a research model from Netflix that removes objects from videos along with the physical inte… | 54 | 1965 | active |
| boycy815/PinchImageView PinchImageView is a lightweight Android image gesture control that extends ImageView with pinch-to-zoom, swipe inertia, double-tap zoom, an… | 32 | 1961 | stable |
| open-mmlab/mmagic MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri… | 23 | 7457 | maintenance |
| SizheAn/PanoHead PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis… | 29 | 1956 | active |
| tpaviot/pythonocc-core pythonocc-core is a Python package providing 3D modeling and data exchange features based on the OpenCascade Technology (OCCT) CAD kernel. … | 74 | 1955 | active |
| alibaba/EasyCV EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo… | 32 | 1954 | active |
| eriklindernoren/PyTorch-YOLOv3 A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights… | 32 | 7440 | maintenance |
| xdan/jodit Jodit is an open-source WYSIWYG rich text editor written in pure TypeScript with zero dependencies, offering a built-in file browser and im… | 94 | 1953 | stable |
| ollm/OpenComic OpenComic is a cross-platform comic and manga reader desktop application built with Node.js and Electron. It supports a wide range of image… | 83 | 1953 | active |
| Fafa-DL/Awesome-Backbones A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin … | 33 | 1953 | active |
| 66HEX/frame Frame is a native desktop GUI for FFmpeg built in Rust with GPUI-CE, providing media conversion for video, audio, and image files with gran… | 77 | 1952 | active |
| chn-lee-yumi/MaterialSearch MaterialSearch is a self-hosted semantic search tool that indexes local photos and videos using a CLIP multimodal model, letting users find… | 73 | 1949 | active |
| LTH14/mar Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive … | 54 | 1949 | stable |
| openai/guided-diffusion OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio… | 32 | 7419 | maintenance |
| multiavatar/Multiavatar Multiavatar is an open-source multicultural avatar generator library that converts any input string into a unique SVG avatar, capable of pr… | 32 | 1946 | stable |
| ruyo/VRM4U VRM4U is an Unreal Engine (UE4/UE5) plugin that imports VRM 3D avatar files at runtime and in-editor. It generates skeletal rigs, morph tar… | 97 | 1945 | active |
| lgarron/folderify A Rust CLI tool that generates pixel-perfect macOS folder icons in the native style from a PNG mask file, producing .icns and .iconset file… | 80 | 1941 | active |
| clawsoftware/clawPDF clawPDF is an open-source virtual printer for Windows that converts printed output into PDF, PDF/A, OCR text, SVG, and various image format… | 23 | 1940 | active |
| PixArt-alpha/PixArt-sigma PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with… | 25 | 1939 | active |
| starik222/BooruDatasetTagManager A desktop tag editor for managing booru-style tagged image and video datasets used to train Stable Diffusion models such as LoRAs, embeddin… | 76 | 1938 | active |
| NVlabs/RADIO Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro… | 64 | 1933 | active |
| wysaid/android-gpuimage-plus A C++ and Java library for Android that applies GPU-accelerated image, camera, and video filters using OpenGL shaders. It supports rule-str… | 88 | 1927 | active |
| riddleling/iOS-OCR-Server An iOS app that turns an iPhone into a local OCR server using Apple's Vision Framework, exposing an HTTP API and web interface for image te… | 74 | 1927 | active |