function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| manycore-maas/Painter Painter is a JSON-driven canvas drawing library for WeChat mini programs (also usable in Node and HTML5) that renders images from a declara… | 23 | 4475 | maintenance |
| AravisProject/aravis Aravis is a C library based on GLib/GObject for video acquisition from Genicam-compliant industrial cameras, implementing GigE Vision and U… | 81 | 1268 | active |
| BachiLi/diffvg diffvg is a differentiable rasterizer for 2D vector graphics that bridges the raster and vector domains via backpropagation. It computes pi… | 42 | 1268 | active |
| Unity-Technologies/VFXToolbox A Unity package from Unity Technologies providing additional tools for Visual Effect artists, including an Image Sequencer for authoring fl… | 33 | 1268 | active |
| nv-tlabs/Difix3D Difix3D+ is a research codebase from NVIDIA implementing a single-step diffusion model pipeline that removes artifacts from NeRF and 3D Gau… | 32 | 1266 | active |
| blueimp/JavaScript-Load-Image A JavaScript library that loads images from File/Blob objects or URLs and returns optionally scaled, cropped, or rotated HTML img or canvas… | 32 | 4456 | maintenance |
| brendan-duncan/image A pure-Dart library for decoding, encoding, and manipulating images in many formats (PNG, JPEG, GIF, WebP, TIFF, BMP, and more). It works w… | 76 | 1263 | active |
| bryandlee/animegan2-pytorch A PyTorch implementation of AnimeGANv2, a GAN-based image-to-image style transfer model that converts photos into anime-style images. It pr… | 32 | 4452 | maintenance |
| GraphiteEditor/Graphite Graphite is a free, open source 2D graphics editor built in Rust that combines layer-based compositing with a node-based procedural graphic… | 67 | 26939 | experimental |
| ashuoAI/SHUO-Canvas SHUO Canvas (formerly AI-CanvasPro) is an AI multimodal creation canvas application that lets users combine text, images, video, and audio … | 82 | 1261 | active |
| rlawjdghek/StableVITON StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent… | 47 | 1261 | stable |
| Jack000/Expose Expose is a Bash-based static site generator that turns folders of images and videos into photoessay-style websites. It requires only Image… | 32 | 4437 | maintenance |
| ShiftHackZ/Stable-Diffusion-KMP SDAI is an open-source, cross-platform Stable Diffusion client app for Android and iOS built with Kotlin Multiplatform and Jetpack Compose.… | 91 | 1259 | active |
| nv-tlabs/GET3D GET3D is NVIDIA's PyTorch implementation of a generative model that synthesizes high-quality 3D textured meshes (cars, chairs, animals, bui… | 32 | 4435 | maintenance |
| skydoves/Cloudy Cloudy is a Kotlin Multiplatform library for Compose that provides blur, liquid glass, and shader-driven surface effects via simple Modifie… | 96 | 1258 | active |
| dmrschmidt/DSWaveformImage A Swift library that generates waveform images from audio files on iOS, iPadOS, macOS, visionOS, and Mac Catalyst. It provides native Swift… | 82 | 1258 | active |
| stardist/stardist StarDist is a Python library for object detection and instance segmentation in 2D and 3D microscopy images using star-convex shapes, built … | 60 | 1255 | stable |
| lucasb-eyer/go-colorful A Go library for storing, converting, and manipulating colors across many color spaces including RGB, HSL, HSV, hex, CIE-Lab, CIE-Luv, HCL,… | 88 | 1254 | stable |
| sassman/t-rec-rs t-rec is a fast, offline terminal recorder written in Rust that captures terminal (or arbitrary window) sessions and generates animated GIF… | 85 | 1253 | active |
| VainF/pytorch-msssim A PyTorch library providing fast, differentiable SSIM and MS-SSIM image quality metrics using separable Gaussian filtering for speed. It ca… | 23 | 1253 | stable |
| FaceAISDK/FaceAISDK_Android An Android SDK for fully on-device, offline face detection, recognition, liveness detection (anti-spoofing), and 1:1, 1:N, and M:N face sea… | 98 | 1252 | active |
| meodai/poline Poline is a tiny TypeScript micro-library that generates color palettes by interpolating HSL colors in polar/cartesian space between anchor… | 71 | 1252 | active |
| meowtec/Imagine Imagine is a cross-platform desktop GUI app for optimizing PNG, JPEG, and WebP images, built on pngquant, mozjpeg, and WebP encoders with a… | 23 | 4405 | maintenance |
| antfu/qrcode-toolkit A web-based toolkit for generating base QR codes and refining AI-generated QR codes by comparing outputs to find misaligned pixels. It also… | 29 | 1250 | active |
| KnpLabs/KnpSnappyBundle A Symfony bundle integrating the Snappy PHP wrapper around wkhtmltopdf/wkhtmltoimage, letting Symfony apps convert HTML documents or URLs i… | 70 | 1247 | active |
| sentinel-hub/eo-learn eo-learn is a collection of open-source Python packages for accessing and processing spatio-temporal satellite imagery, built around modula… | 51 | 1247 | active |
| MikeKovarik/exifr exifr is a fast, dependency-free JavaScript library for reading EXIF and other image metadata (TIFF, XMP, ICC, IPTC, GPS, JFIF) from JPG, P… | 23 | 1247 | active |
| peterbraden/node-opencv Native Node.js bindings for the OpenCV computer vision library, exposing Matrices, image reading/writing, and cascades like face detection … | 23 | 4384 | maintenance |
| ultranity/Pix-EzViewer PixEz is a third-party Pixiv client for Android built with Kotlin and Jetpack, offering a modern Material Design UI and many enhancements o… | 87 | 1246 | active |
| withoutbg/withoutbg-python A Python SDK (pip install withoutbg) for removing image backgrounds, offering a free local open-weights ONNX model and an optional paid clo… | 80 | 1246 | active |
| lpiccinelli-eth/UniDepth UniDepth is a Python library and research codebase for universal monocular metric depth estimation from single images, based on CVPR 2024 a… | 35 | 1246 | active |
| jark006/JarkViewer JarkViewer is a minimalist, fast image viewer for 64-bit Windows supporting a huge range of formats including AVIF, HEIC, JPEG XL, RAW came… | 83 | 1245 | active |
| unum-cloud/UForm UForm is a compact multimodal AI library providing tiny image-text embedding models (64-768 dimensions, Matryoshka-style) and small generat… | 55 | 1244 | active |
| fpgaminer/joycaption JoyCaption is an open, free, and uncensored image captioning Visual Language Model (VLM) with released weights and training scripts. It gen… | 53 | 1244 | active |
| LTH14/fractalgen A PyTorch implementation of Fractal Generative Models (FractalGen), enabling pixel-by-pixel high-resolution image generation. It includes p… | 24 | 1244 | active |
| 3DTopia/OpenLRM OpenLRM is an open-source PyTorch implementation of Large Reconstruction Models (LRM) that reconstruct 3D objects (meshes and rendered vide… | 17 | 1244 | active |
| abewley/sort SORT is a barebones Python implementation of a simple online and realtime multiple object tracking algorithm for 2D video sequences, based … | 32 | 4373 | maintenance |
| DDULDDUCK/every-pdf Every PDF is an all-in-one open-source desktop PDF toolkit built with Electron, Next.js, and a Python (FastAPI) backend. It lets users edit… | 71 | 1243 | active |
| cvg/depthsplat DepthSplat is a PyTorch research library implementing a CVPR 2025 model that connects Gaussian splatting with single/multi-view depth estim… | 56 | 1242 | active |
| raspberrypi/picamera2 Picamera2 is a Python library providing an interface to Raspberry Pi cameras via the libcamera stack, replacing the legacy Picamera library… | 94 | 1241 | active |
| Roblox/cube Cube is Roblox's open-source family of foundation models for 3D intelligence, including text-to-3D shape generation and part-controllable m… | 58 | 1241 | active |
| XPixelGroup/HYPIR Official PyTorch implementation of HYPIR, a SIGGRAPH 2025 method that harnesses diffusion-yielded score priors for image restoration. It pr… | 39 | 1241 | active |
| jcjohnson/fast-neural-style A Torch (Lua) implementation of feedforward neural style transfer from the ECCV 2016 paper 'Perceptual Losses for Real-Time Style Transfer … | 32 | 4359 | maintenance |
| TransparentLC/WechatMomentScreenshot A web-based tool that generates fake WeChat Moments (朋友圈) share screenshots, including text posts, shared articles, images, and nine-grid l… | 32 | 4358 | maintenance |
| ZHKKKe/MODNet MODNet is a deep learning model for real-time portrait matting (background removal) that requires only an RGB image as input, with no trima… | 32 | 4355 | maintenance |
| facebookresearch/deit Official PyTorch repository for DeiT and related vision transformer architectures (CaiT, ResMLP, PatchConvnet, DeiT III), providing trainin… | 10 | 4355 | maintenance |
| itgalaxy/favicons A Node.js library that generates favicons and associated web app icon files (including manifests and HTML snippets) from a source image, us… | 92 | 1238 | active |
| hymbz/ComicReadScript A userscript (Tampermonkey/Violentmonkey) that adds a two-page spread reading mode and various UX enhancements to popular manga/comic websi… | 96 | 1237 | active |
| amerkoleci/Vortice.Windows Vortice.Windows is a set of .NET libraries providing modern C# bindings for DirectX APIs including Direct3D 9/11/12, Direct2D, DirectWrite,… | 54 | 1237 | active |
| open-mmlab/playground OpenMMLab Playground is a central hub collecting and showcasing community projects that extend OpenMMLab libraries with Segment Anything Mo… | 30 | 1236 | active |
| apache/incubator-pagespeed-ngx ngx_pagespeed is an Apache-licensed Nginx module (originally from Google, incubated at Apache) that automatically applies web performance b… | 10 | 4339 | maintenance |
| exelix11/SwitchThemeInjector A suite of tools for creating and installing custom home menu themes on a modded Nintendo Switch, including the NXThemes Installer homebrew… | 94 | 1235 | active |
| bytedance/USO USO is ByteDance's open-source unified style- and subject-driven image generation model based on diffusion (FLUX), combining any subject wi… | 36 | 1235 | active |
| xtreme1-io/xtreme1 Xtreme1 is an open-source, self-hosted data labeling and annotation platform for multimodal training data, supporting images, 3D LiDAR poin… | 62 | 1234 | active |
| mbrevda/react-image A React library providing an <img> tag replacement and useImage hook that supports fallback to alternate image sources on load failure. It … | 77 | 1232 | active |
| Acly/comfyui-inpaint-nodes A set of custom nodes for ComfyUI that improve image inpainting and outpainting workflows. It integrates the Fooocus inpaint model for SDXL… | 64 | 1232 | active |
| VAST-AI-Research/TripoSplat TripoSplat is an inference-only Python library from TripoAI that converts a single 2D image into high-quality 3D Gaussian splats with a var… | 57 | 1232 | active |
| lucidrains/deep-daze Deep Daze is a simple command line tool for text-to-image generation that combines OpenAI's CLIP with a Siren implicit neural representatio… | 23 | 4315 | maintenance |
| stereolabs/zed-sdk The ZED SDK is a cross-platform spatial perception library for Stereolabs ZED stereo cameras, providing depth sensing, SLAM, 3D reconstruct… | 91 | 1229 | active |
| hundredrabbits/Ronin Ronin is an experimental procedural graphics terminal that interprets a minimal LISP dialect to automate graphical tasks like resizing, cro… | 39 | 1229 | active |
| florestefano1975/comfyui-portrait-master A ComfyUI custom node suite that helps AI image creators generate detailed, professional prompts for human portraits. It provides modular n… | 56 | 1228 | active |
| diffusionstudio/core Diffusion Studio Core is a TypeScript video compositing engine that runs in the browser, built on WebCodecs and Canvas2D for hardware-accel… | 48 | 1228 | active |
| abo-abo/org-download An Emacs Lisp extension that makes it easy to insert images into org-mode buffers. It supports dragging images from browsers or the file sy… | 32 | 1228 | active |
| SideFX Labs SideFX Labs is a free, open-source, artist-friendly toolset for Houdini containing hundreds of Houdini Digital Assets (HDAs), Python module… | 91 | 1226 | active |
| pythongosssss/ComfyUI-WD14-Tagger A ComfyUI custom node extension that interrogates images to extract booru-style tags using WD 1.4 tagger models (ONNX-based). It integrates… | 43 | 1226 | active |
| b-editor/beutl Beutl is a free, open-source, cross-platform video editing and compositing application built on .NET. It offers a timeline with layers, nod… | 99 | 1225 | active |
| PatilShreyas/Capturable Capturable is a Jetpack Compose utility library for Android that captures Composable UI content and converts it into a Bitmap image. It pro… | 51 | 1224 | active |
| mcmonkeyprojects/sd-dynamic-thresholding A Stable Diffusion extension that enables using higher CFG scale values without color artifacts by clamping latents between sampling steps.… | 36 | 1224 | active |
| flyimg/flyimg Flyimg is a Dockerized, self-hosted application that resizes, crops, and compresses images on the fly via URL parameters, serving optimized… | 100 | 1223 | active |
| benrugg/AI-Render A Blender add-on that renders AI-generated images with Stable Diffusion based on a text prompt and the user's 3D scene. It supports cloud r… | 57 | 1223 | active |
| mrousavy/react-native-fast-tflite A high-performance TensorFlow Lite library for React Native built on Nitro Modules, using the low-level C/C++ TFLite core API with zero-cop… | 84 | 1222 | active |
| maxritter/diy-thermocam DIY-Thermocam is an open-source, self-assembly thermal imaging camera based on the FLIR Lepton sensor and a Teensy 4.1 microcontroller, wit… | 48 | 1222 | active |
| vansh-nagar/ascii-studio ASCII Studio is a browser-based tool that converts videos into real-time ASCII character-based animations. It offers customizable character… | 56 | 1221 | active |
| fastgs/FastGS FastGS is a general acceleration framework for 3D Gaussian Splatting that trains scenes in roughly 100 seconds using multi-view consistent … | 49 | 1221 | active |
| rvanwijnen/spectral.js Spectral.js is a lightweight JavaScript library for realistic paint-like color mixing based on the Kubelka-Munk theory, simulating how pigm… | 48 | 1220 | active |
| Project-MONAI/research-contributions A collection of peer-reviewed research prototype implementations built on the MONAI framework for medical imaging AI. It serves as a fast-t… | 43 | 1220 | active |
| ststeiger/PdfSharpCore PdfSharpCore is a .NET Standard port of the PdfSharp and MigraDoc libraries for creating and manipulating PDF documents, with GDI+ dependen… | 39 | 1220 | active |
| tapmodo/Jcrop Jcrop is a JavaScript image cropping engine that lets developers add interactive crop-selection functionality to images in web applications… | 32 | 4273 | maintenance |
| bowang-lab/MedRAX MedRAX is a medical reasoning agent framework that integrates chest X-ray analysis tools (segmentation, grounding, report generation, disea… | 42 | 1218 | active |
| Apparence-io/CamerAwesome CamerAwesome is a Flutter plugin that embeds a fully customizable camera experience into Android and iOS apps. It exposes native camera fea… | 57 | 1215 | active |
| tonyqinatcmu/SlideBot-AI SlideBot AI is an AI-powered presentation generator that turns a topic, outline, or uploaded materials (documents, spreadsheets, meeting re… | 44 | 1214 | active |
| kijai/ComfyUI-segment-anything-2 A set of ComfyUI custom nodes that bring Meta's Segment Anything 2 (SAM2) models into ComfyUI workflows for promptable image and video segm… | 43 | 1214 | active |
| MapServer/MapServer MapServer is an open-source geographic data rendering engine written in C for publishing spatial data and interactive mapping applications … | 97 | 1213 | stable |
| fo-dicom/fo-dicom Fellow Oak DICOM (fo-dicom) is a C#/.NET library implementing the DICOM standard for medical imaging, including parsing, image codecs, anon… | 86 | 1210 | active |
| Jaysmito101/TerraForge3D TerraForge3D is a free, open-source, cross-platform procedural terrain generation and texturing tool built in C++ with GPU acceleration. It… | 66 | 1209 | active |
| DachunKai/EvTexture Official PyTorch implementation of EvTexture and EvTexture++, event-driven video super-resolution models that use event-camera signals to e… | 54 | 1207 | active |
| gali8/Tesseract-OCR-iOS An iOS framework wrapping the Tesseract OCR engine (with Leptonica and image libraries) for use in Objective-C or Swift apps on iOS 9.0+. I… | 23 | 4221 | maintenance |
| willisma/SiT Official PyTorch implementation of Scalable Interpolant Transformers (SiT), a family of generative models built on Diffusion Transformers t… | 53 | 1206 | active |
| pytroll/satpy Satpy is a Python library for reading, manipulating, and writing meteorological remote sensing data from earth-observing satellites. It sup… | 86 | 1205 | active |
| frotms/PaddleOCR2Pytorch A PyTorch port of PaddleOCR that lets you run PaddleOCR-trained models (detection, recognition, and document structure parsing) without the… | 73 | 1205 | active |
| emoon/rust_minifb minifb is a cross-platform Rust crate for creating windows and displaying 32-bit pixel buffers, with keyboard and mouse input support. It i… | 76 | 1204 | active |
| sammycage/lunasvg LunaSVG is a lightweight, portable C++ library for rendering and manipulating SVG files, built on PlutoVG. It can rasterize SVG documents t… | 77 | 1203 | active |
| JuneYaooo/gpt-image2-ppt-skills A Claude Code / OpenClaw skill that generates polished 16:9 presentations using OpenAI's gpt-image-2, rendering each slide as a complete vi… | 59 | 1203 | active |
| pyang5166/gbro-collage-broll An agent skill that turns short voiceover lines into editorial halftone paper-collage B-roll videos using Gemini Omni Flash first/last-fram… | 54 | 1202 | active |
| dendenxu/fast-gaussian-rasterization A drop-in replacement for diff-gaussian-rasterization that renders 3D Gaussian Splatting scenes using a geometry-shader-based GPU pipeline … | 19 | 1202 | active |
| jantimon/favicons-webpack-plugin A webpack plugin that generates favicons and app icons in dozens of formats from a single logo file, leveraging the favicons library. It al… | 76 | 1200 | stable |
| GoogleCloudPlatform/vertex-ai-creative-studio GenMedia Creative Studio is a web application showcasing Google Cloud's generative media APIs including Gemini Image (Nano Banana), Veo vid… | 90 | 1199 | active |
| cleanlab/cleanvision CleanVision is a Python library that automatically detects issues in image datasets, such as blurry, dark, over-exposed, or near-duplicate … | 59 | 1199 | active |
| Morgensonne/EditDeck EditDeck is an end-to-end pipeline that turns plain-language requirements into structured slide outlines, rendered slide images, standard P… | 50 | 1199 | active |
| guoyingtao/Mantis Mantis is a Swift image cropping library for iOS and Mac Catalyst offering UIKit and SwiftUI APIs with an Apple Photos-style crop experienc… | 94 | 1198 | active |