function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| AIGODLIKE/ComfyUI-BlenderAI-node A Blender addon that integrates ComfyUI into Blender by converting ComfyUI nodes into Blender nodes, enabling AI image generation, material… | 47 | 1531 | active |
| msterzhang/onelist onelist is a self-hosted media library application written in Go, similar to Emby, that scrapes metadata from TheMovieDb for files stored o… | 21 | 1531 | active |
| JingyunLiang/SwinIR Official PyTorch implementation of SwinIR, a Swin Transformer-based model for image restoration tasks including super-resolution, denoising… | 23 | 5580 | maintenance |
| hustvl/LightningDiT LightningDiT is a research codebase for latent diffusion models implementing VA-VAE and LightningDiT, achieving FID 1.35 on ImageNet-256 wi… | 47 | 1529 | active |
| luciddreamer-cvlab/LucidDreamer LucidDreamer is the official implementation of a research method that generates 3D Gaussian Splatting scenes from text prompts, published i… | 69 | 1528 | active |
| DavBfr/dart_pdf A Dart/Flutter library for creating PDF documents, offering both a low-level PDF generation API and a Flutter-like widget system for high-l… | 74 | 1527 | stable |
| cchen156/Learning-to-See-in-the-Dark TensorFlow implementation of 'Learning to See in the Dark' (CVPR 2018), a deep learning model that brightens very dark, short-exposure RAW … | 51 | 5565 | maintenance |
| vinceglb/FileKit FileKit is a Kotlin Multiplatform library providing a unified API for file operations such as picking files, saving documents, selecting di… | 86 | 1526 | active |
| AlekPet/ComfyUI_Custom_Nodes_AlekPet A collection of custom nodes for ComfyUI that extend its capabilities with painting, pose control, prompt translation, and speech recogniti… | 72 | 1524 | active |
| TransparentLC/realesrgan-gui A cross-platform graphical interface for the Real-ESRGAN AI image upscaler (with Real-CUGAN support), built in Python with tkinter. It wrap… | 59 | 1524 | active |
| anishathalye/neural-style A Python command-line tool implementing the neural style transfer algorithm (Gatys et al.) in TensorFlow, applying the style of one image t… | 67 | 5541 | maintenance |
| caiyuanhao1998/Retinexformer Retinexformer is a one-stage Retinex-based Transformer model and toolbox for low-light image enhancement, published at ICCV 2023. It suppor… | 66 | 1518 | active |
| elringus/sprite-dicing A cross-engine tool written in Rust for losslessly compressing sprite textures by splitting them into units, discarding identical regions, … | 69 | 1514 | active |
| rgab1508/PixelCraft PixelCraft is a browser-based pixel art editor built with JavaScript and the HTML5 Canvas API, installable as a progressive web app for off… | 37 | 1514 | active |
| NVlabs/describe-anything Describe Anything Model (DAM) is a vision-language model that generates detailed descriptions of user-specified regions in images and video… | 32 | 1514 | active |
| BrokenSource/DepthFlow DepthFlow is a free, open-source Python application and library that converts still images into 3D parallax effect videos using monocular d… | 85 | 1512 | active |
| HiDream-ai/HiDream-O1-Image HiDream-O1-Image is an open-weights 8B image generation foundation model built on a Pixel-level Unified Transformer (UiT) that natively enc… | 53 | 1510 | active |
| hkchengrex/Tracking-Anything-with-DEVA DEVA is a decoupled video segmentation framework that combines task-specific image-level segmentation models with a universal bi-directiona… | 27 | 1508 | stable |
| koide3/direct_visual_lidar_calibration A C++ toolbox for target-less, single-shot extrinsic calibration between LiDAR sensors and cameras, supporting spinning and non-repetitive … | 74 | 1507 | stable |
| EricLengyel/Slug Reference shader implementations of the Slug algorithm for rendering vector fonts directly on the GPU using quadratic Bézier curve evaluati… | 49 | 1506 | stable |
| kenglxn/QRGen QRGen is a simple Java QR code generation library built as a wrapper on top of ZXING, with a fluent API for creating QR codes as files or s… | 52 | 1505 | stable |
| CanHub/Android-Image-Cropper A Kotlin image cropping library for Android, optimized for images picked from the camera or gallery. It provides a customizable CropImageVi… | 65 | 1504 | active |
| pmp-library/pmp-library A modern C++ open-source library for processing and visualizing polygon surface meshes. It provides an efficient mesh data structure, stand… | 67 | 1503 | stable |
| czczup/ViT-Adapter Official PyTorch implementation of ViT-Adapter, an ICLR 2023 Spotlight paper introducing a pre-training-free adapter that lets plain Vision… | 34 | 1503 | stable |
| charlesq34/pointnet Reference implementation of PointNet, a neural network architecture that directly consumes unordered 3D point clouds for classification and… | 32 | 5459 | maintenance |
| daavoo/pyntcloud pyntcloud is a Python library for working with 3D point clouds, built on the Python scientific stack (pandas, numpy). It supports loading/s… | 64 | 1500 | active |
| meiqua/shape_based_matching A C++ library implementing Halcon-style shape-based matching (equivalent to LINE-MOD) using gradient orientation templates for robust 2D ob… | 32 | 1500 | active |
| ANTsX/ANTs Advanced Normalization Tools (ANTs) is a C++ command-line library for high-dimensional medical image registration and segmentation, built o… | 83 | 1497 | active |
| microsoft/ai-dev-gallery AI Dev Gallery is a Windows application from Microsoft that lets developers explore over 25 interactive samples powered by local AI models … | 66 | 1497 | active |
| jtheoof/swappy Swappy is a Wayland-native screenshot annotation and editing tool inspired by Snappy on macOS. It accepts images from stdin or files (e.g.,… | 61 | 1496 | active |
| piddnad/DDColor DDColor is the official PyTorch implementation of an ICCV 2023 paper on photo-realistic automatic image colorization using dual decoders an… | 59 | 1496 | active |
| NJU-PCALab/STAR STAR is a research implementation of an ICCV 2025 paper performing real-world video super-resolution using spatial-temporal augmentation wi… | 34 | 1495 | active |
| ecomfe/fonteditor Fonteditor is an online font editor for editing, transforming, and previewing fonts directly in the browser. It supports importing and edit… | 37 | 1494 | active |
| Harley-xk/MaLiang MaLiang is an iOS painting and drawing framework built on Apple's Metal, supporting textured strokes, handwriting, and Apple Pencil input. … | 62 | 1493 | active |
| RobertBroersma/beanheads Bean Heads is a React component library for generating customizable cartoon-style avatar illustrations via SVG. It exposes props for hair, … | 60 | 1489 | active |
| una/CSSgram CSSgram is a tiny CSS/Sass library that recreates Instagram-style photo filters using CSS filters and blend modes. Filters are applied by a… | 23 | 5395 | maintenance |
| huiyadanli/PasteEx PasteEx is a portable Windows utility that pastes clipboard contents directly into files, with automatic image extension detection and cust… | 34 | 1485 | stable |
| CameraKit/camerakit-android CameraKit is an Android library that wraps the Camera 1 and Camera 2 APIs behind a reliable CameraView component for photo and video captur… | 23 | 5387 | maintenance |
| amandaghassaei/gpu-io gpu-io is a TypeScript WebGL library for composing GPU-accelerated computing workflows in the browser. It handles WebGL state management, s… | 32 | 1482 | active |
| atteneder/glTFast glTFast is a Unity package for efficiently importing and exporting glTF 3D files at runtime or in the Unity Editor. It supports the glTF 2.… | 91 | 1480 | active |
| emcconville/wand Wand is a ctypes-based Python binding for the ImageMagick MagickWand API, supporting Python 3.8+ and PyPy. It exposes the full MagickWand f… | 90 | 1478 | active |
| milon/barcode A Laravel package that generates 1D and 2D barcodes (including QR codes, Datamatrix, and PDF417) by wrapping TCPDF's barcode classes. It pr… | 65 | 1478 | active |
| dcalsky/zzkia A web application that generates fake Nokia phone message screenshots for fun. Users can input text and get back an image styled like an ol… | 44 | 1473 | active |
| KhronosGroup/glTF-Sample-Viewer The official Khronos glTF 2.0 Sample Viewer, a web application that renders glTF 2.0 3D models with physically-based rendering using WebGL … | 66 | 1472 | active |
| Jounce/Surge Surge is a Swift library that wraps Apple's Accelerate framework to expose high-performance SIMD-accelerated functions for matrix math, dig… | 23 | 5325 | maintenance |
| jonbhanson/flutter_native_splash A Flutter/Dart package that automatically generates native code to customize the default white splash screen shown while the native app loa… | 84 | 1470 | active |
| javakam/FileOperator An Android file operations library covering file creation/deletion/copying, opening files and directories, MIME types, MediaStore and SAF h… | 32 | 1469 | active |
| aws-solutions/dynamic-image-transformation-for-amazon-cloudfront An official AWS Solutions implementation that deploys a serverless architecture for real-time image transformation, optimization, and deliv… | 97 | 1467 | stable |
| unlimitedbacon/stl-thumb A fast, lightweight thumbnail generator for 3D model files (STL, OBJ, 3MF) written in Rust using OpenGL. It integrates with file managers o… | 27 | 1465 | active |
| hojonathanho/diffusion The official reference implementation of Denoising Diffusion Probabilistic Models (DDPM) from the 2020 paper by Jonathan Ho et al., written… | 32 | 5300 | maintenance |
| yunjey/stargan Official PyTorch implementation of StarGAN, a unified generative adversarial network for multi-domain image-to-image translation (CVPR 2018… | 32 | 5296 | maintenance |
| fossasia/magic-epaper-firmware Magic ePaper Firmware is open-source C firmware for controlling FOSSASIA's e-ink (e-paper) displays on microcontrollers. It acts as a bridg… | 45 | 1462 | active |
| guardian/grid Grid is the Guardian's open-source image management system (digital asset management) built as a set of Scala/Play microservices with an An… | 77 | 1461 | active |
| microlinkhq/unavatar unavatar.io is a unified avatar API that retrieves user profile images from a single URL by username, email, or domain across 72 platforms … | 95 | 1460 | active |
| FeiYull/TensorRT-Alpha A C++/CUDA library providing TensorRT-accelerated deployment for 30+ popular computer vision models including YOLOv3-v8, YOLOv8-Pose/Seg/Cl… | 32 | 1460 | active |
| fholger/vrperfkit A collection of performance-oriented mods for VR games, delivered as a DLL injected next to a game executable. It provides upscaling (AMD F… | 23 | 1458 | active |
| mikepenz/Android-Iconics An Android library that lets developers use any icon font or vector (.svg) as a Drawable in their app, with full customization of size, col… | 87 | 5272 | maintenance |
| spatie/pdf-to-image A PHP library that converts PDF files to images (jpg, png, webp) using Imagick and Ghostscript. It supports rendering single or multiple pa… | 88 | 1457 | active |
| Hello-hao/Tbed Hellohao Image Hosting (Tbed) is an open-source, self-hosted image hosting application built with Java and SpringBoot using a front-end/bac… | 67 | 1457 | active |
| 869413421/ai-moive-studio AICON is a full-stack, self-hostable AI video creation studio that pairs a natural-language-driven agent with an infinite-canvas node edito… | 60 | 1456 | active |
| p2r3/beheader A command-line tool that generates polyglot files - single files that are simultaneously valid images, videos, PDFs, ZIP archives, and HTML… | 48 | 1455 | active |
| erweixin/RaTeX RaTeX is a KaTeX-compatible LaTeX math rendering engine written in pure Rust, with no JavaScript, WebView, or DOM dependency. A single Rust… | 80 | 1451 | active |
| dcm4che/dcm4che dcm4che is a Java toolkit and library implementing the DICOM standard for medical imaging, including data set handling, network communicati… | 91 | 1450 | active |
| psd-tools/psd-tools psd-tools is a Python package for reading and manipulating Adobe Photoshop PSD/PSB files. It parses the low-level file structure, exports l… | 98 | 1449 | active |
| showmewebcam/showmewebcam A firmware image that turns a Raspberry Pi Zero (or Pi 4) with a Raspberry Pi camera module into a high-quality USB webcam using UVC gadget… | 23 | 1449 | active |
| GPUOpen-Tools/compressonator Compressonator is a tool suite from AMD GPUOpen for compressing, optimizing, and analyzing texture assets and 3D model meshes using CPUs, G… | 23 | 1447 | active |
| phonowell/genshin-impact-script A Genshin Impact automation script written in AutoHotkey that provides features like automatic fishing, item pickup, and dialogue skipping.… | 63 | 1446 | active |
| morsoli/aimangastudio AIMangaStudio is a web application that uses AI (Google GenAI) to create manga/comics end-to-end, including script generation, storyboard l… | 62 | 1446 | active |
| amdegroot/ssd.pytorch A PyTorch implementation of the Single Shot MultiBox Detector (SSD) object detection model from the 2016 paper by Wei Liu et al. It include… | 32 | 5221 | maintenance |
| xmikos/qspectrumanalyzer QSpectrumAnalyzer is a PyQtGraph-based GUI spectrum analyzer for software-defined radio devices. It wraps multiple measurement backends suc… | 23 | 1445 | stable |
| raoxwup/haka_comic HaKa Comic is a third-party cross-platform client for the PicACG (Bika/Pica) comic platform, built with Flutter. It provides a clean, ad-fr… | 87 | 1444 | active |
| zsyOAOA/InvSR InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff… | 51 | 1443 | active |
| serratus/quaggaJS QuaggaJS is a barcode-scanner library written entirely in JavaScript that supports real-time localization and decoding of barcode types suc… | 23 | 5207 | maintenance |
| sdcb/PaddleSharp A .NET/C# wrapper around Baidu's PaddleInference C API, providing PaddleOCR, PaddleDetection, rotation detection, Chinese segmentation, and… | 72 | 1441 | active |
| xhongc/ai_story AI Story is a self-hosted platform that automates story video production: given a topic, it generates scripts, storyboards, AI images, came… | 60 | 1441 | active |
| AIM-Harvard/pyradiomics PyRadiomics is an open-source Python package for extracting radiomics features from 2D and 3D medical images and binary masks. It provides … | 45 | 1441 | active |
| ThoughtfulDev/EagleEye EagleEye is a Python-based OSINT tool that identifies social media profiles (Instagram, Facebook, Twitter, YouTube) of a person using face … | 32 | 5197 | maintenance |
| txperl/PixivBiu PixivBiu is a self-hosted Pixiv client written in Go that provides a browser-based interface for searching, filtering, browsing, and downlo… | 87 | 1440 | active |
| 027xiguapi/pear-rec pear-rec is a free, open-source, cross-platform desktop application for taking screenshots and recording screen video, webcam video, and au… | 27 | 1440 | active |
| neuralchen/SimSwap SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It… | 23 | 5188 | maintenance |
| natemoo-re/astro-icon Astro Icon is an Astro integration that provides an Icon component for inlining local SVGs and icons from any Iconify icon set, with suppor… | 95 | 1436 | active |
| xuanyustudio/LocalMiniDrama LocalMiniDrama is an open-source Electron desktop application that generates AI short dramas and animated dramas end-to-end, from story and… | 78 | 1436 | active |
| Jonathan-LeRoux/IguanaTex IguanaTex is a free PowerPoint add-in that lets users insert and edit LaTeX equations directly in presentations on Windows and Mac. It comp… | 64 | 1436 | active |
| jakowenko/double-take Double Take is a self-hosted Docker application providing a unified UI and API for facial recognition. It abstracts multiple face detection… | 40 | 1436 | active |
| potatameister/PaperKnife PaperKnife is a privacy-first PDF utility that merges, splits, compresses, encrypts, signs, and sanitizes PDFs entirely on-device with no s… | 67 | 1435 | active |
| alelievr/HDRP-Custom-Passes A collection of custom render passes and shader effects for Unity's High Definition Render Pipeline (HDRP). It includes ready-to-use effect… | 32 | 1434 | active |
| harlan-zw/nuxt-seo Nuxt SEO (@nuxtjs/seo) is a bundle of Nuxt modules providing technical SEO and answer-engine optimization: robots.txt generation, XML sitem… | 98 | 1433 | active |
| LingGuoAI/LingGuo-Drama LingGuo-Drama is a self-hosted, AI-powered one-stop platform for generating mini-dramas and motion comics, covering script writing, charact… | 59 | 1433 | active |
| JiongXing/PhotoBrowser JXPhotoBrowser is a lightweight, highly customizable iOS photo and video browser library written in Swift. It provides zoom transitions, dr… | 86 | 1432 | stable |
| Haneke Haneke is a lightweight generic cache library for iOS and tvOS written in Swift, with memory and LRU disk caching for UIImage, NSData, JSON… | 23 | 5155 | maintenance |
| qupath/qupath QuPath is an open-source desktop application for bioimage analysis, aimed especially at digital pathology and whole-slide imaging. It provi… | 79 | 1430 | active |
| Francis-Rings/StableAnimator StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po… | 41 | 1430 | active |
| jasonmayes/Real-Time-Person-Removal A browser-based demo that removes people from complex video backgrounds in real time using TensorFlow.js. It learns the static background o… | 23 | 5154 | maintenance |
| zsyOAOA/ResShift ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi… | 62 | 1427 | active |
| zuruoke/watermark-removal A machine learning tool that removes watermarks from images using deep learning image inpainting, based on Contextual Attention and Gated C… | 83 | 5139 | maintenance |
| Tom94/tev tev is a high dynamic range (HDR/EDR) image viewer written in C++ for people who care about accurate colors. It supports many formats (EXR,… | 99 | 1426 | active |
| mdbloice/Augmentor Augmentor is a standalone Python library for image augmentation in machine learning, providing a pipeline of stochastic operations like rot… | 32 | 5133 | maintenance |
| nilearn/nilearn Nilearn is a Python library providing statistical and machine-learning tools for analyzing brain imaging data such as fMRI and MRI volumes … | 88 | 1425 | active |
| ctrl-freaks/freezeframe.js Freezeframe.js is a TypeScript library that pauses animated .gif images by rendering their first frame to a canvas, then animating them on … | 32 | 1425 | active |
| numz/sd-wav2lip-uhq A Wav2Lip Studio extension for the Stable Diffusion WebUI (Automatic1111) that generates high-quality lip-synced talking-face videos from a… | 28 | 1424 | active |