function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| steffest/DPaint-js DPaint.js is a browser-based image editor modeled after Deluxe Paint, with strong support for retro Amiga file formats like IFF ILBM images… | 77 | 1021 | active |
| TencentARC/SEED-Voken SEED-Voken is a collection of visual tokenizers (Open-MAGVIT2 and IBQ) that convert images and videos into discrete tokens for autoregressi… | 48 | 1021 | active |
| yangxy/PASD PASD (Pixel-Aware Stable Diffusion) is a Python research codebase implementing an ECCV 2024 method for realistic image super-resolution and… | 28 | 1021 | active |
| JiahuiYu/generative_inpainting An open-source implementation of DeepFill v1/v2 generative image inpainting models, featuring Contextual Attention (CVPR 2018) and Gated Co… | 32 | 3466 | maintenance |
| tensorlayer/SRGAN Reference implementation of SRGAN, a generative adversarial network for photo-realistic single image super-resolution, built on TensorLayer… | 23 | 3466 | maintenance |
| ocropus-archive/DUP-ocropy OCRopy is a collection of Python-based tools for document analysis and OCR, covering binarization, page layout analysis, and text line reco… | 10 | 3465 | maintenance |
| richzhang/colorization A Python library implementing automatic colorization of grayscale photos using deep neural networks from the ECCV 2016 'Colorful Image Colo… | 32 | 3461 | maintenance |
| zai-org/GLM-Image GLM-Image is an open-source image generation model combining a 9B autoregressive generator with a 7B diffusion decoder, excelling at text r… | 48 | 1019 | active |
| bowang-lab/U-Mamba U-Mamba is a hybrid CNN-state-space-model (Mamba) network for biomedical image segmentation, built on top of the nnU-Net framework. It comb… | 26 | 1019 | active |
| gmattie/Data-Pixels DataPixels.js is a JavaScript library for generating pixel art programmatically at runtime from RGB/RGBA data arrays, producing HTMLCanvasE… | 23 | 3456 | maintenance |
| google/prompt-to-prompt Google's official implementation of the Prompt-to-Prompt paper, which enables text-driven image editing in Latent Diffusion and Stable Diff… | 10 | 3456 | maintenance |
| SnapXL/SnapX SnapX is a free, open-source, cross-platform screenshot and screen recording tool forked from ShareX, built with C# and Avalonia. It lets u… | 76 | 1018 | active |
| linjc/smooth-signature A TypeScript canvas library for capturing handwritten signatures in the browser with a natural pen-brush (variable stroke width) effect. It… | 40 | 1018 | active |
| autonomousvision/gaussian-opacity-fields Gaussian Opacity Fields (GOF) is a Python/CUDA research implementation for efficient, adaptive surface reconstruction in unbounded scenes u… | 25 | 1017 | active |
| evanw/glfx.js glfx.js is a JavaScript library for applying real-time image effects and photo adjustments in the browser using WebGL. It leverages the GPU… | 32 | 3451 | maintenance |
| Sxela/WarpFusion WarpFusion is a Stable Diffusion-based video-to-video style transfer tool distributed as a Jupyter/Colab notebook. It applies AI animation … | 31 | 1016 | active |
| eliukblau/pixterm PIXterm is a Go CLI tool that renders images directly in ANSI terminals using true color escape codes and unicode half-block characters. It… | 85 | 1015 | active |
| Alpha-VLLM/Lumina-DiMOO Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outpu… | 55 | 1015 | active |
| alessandrofrancesconi/gimp-plugin-bimp BIMP is a GIMP plugin that applies a set of image manipulations (resize, crop, rotate, watermark, color correction, format conversion, etc.… | 23 | 1015 | active |
| fallenshock/FlowEdit Official PyTorch implementation of FlowEdit, an ICCV 2025 method for inversion-free, text-based editing of real images using pre-trained fl… | 66 | 1014 | active |
| Kuingsmile/PicHoro PicHoro is a Flutter-based Android app for managing cloud storage platforms and image hosting services, with file upload/download and multi… | 61 | 1014 | active |
| Gregwar/Image A PHP library providing a simple object-oriented API for image handling, resizing, cropping, and applying filters, with built-in caching of… | 32 | 1014 | stable |
| maximeraafat/BlenderNeRF BlenderNeRF is a Blender add-on that generates synthetic NeRF and Gaussian Splatting datasets with a single click, exporting renders and ca… | 23 | 1014 | active |
| addyosmani/bg-remove A React + Vite web application that removes image backgrounds entirely in the browser using Transformers.js with the RMBG-1.4 model (and op… | 22 | 1014 | active |
| fiji/fiji Fiji is a batteries-included distribution of ImageJ, bundling thousands of plugins for scientific image processing and analysis into a port… | 74 | 1013 | active |
| Alpha-VLLM/Lumina-Image-2.0 Lumina-Image 2.0 is an open-source 2.6B-parameter text-to-image generation framework built on a unified Next-DiT architecture with a unifie… | 58 | 1013 | active |
| eragonruan/text-detection-ctpn A TensorFlow implementation of the Connectionist Text Proposal Network (CTPN) for detecting horizontal scene text in images. It includes pr… | 23 | 3429 | maintenance |
| Pjbomb2/TrueTrace-Unity-Pathtracer A high-performance compute shader based path tracer for Unity3D that works without RT cores, using compressed wide BVH for software ray tra… | 81 | 1011 | active |
| chrissimpkins/Crunch Crunch is a lossy PNG image optimization tool that combines bit depth, color type, and palette reduction with zopfli DEFLATE compression vi… | 23 | 3425 | maintenance |
| marwin1991/profile-technology-icons A web-based generator that provides a curated collection of technology icons for decorating GitHub profile READMEs. Users search for techno… | 77 | 1010 | active |
| jabcode/jabcode JAB Code (Just Another Bar Code) is a high-capacity 2D color bar code that encodes more data than traditional black-and-white barcodes. The… | 67 | 1010 | active |
| sail-sg/EditAnything Edit Anything is a Python application for text-guided image editing and generation, combining Segment Anything, ControlNet, BLIP2, and Stab… | 33 | 3422 | maintenance |
| RupertAvery/DiffusionToolkit Diffusion Toolkit is a Windows desktop application that indexes and views metadata (prompts, models, settings) embedded in AI-generated ima… | 65 | 1009 | active |
| ivandokov/phockup Phockup is a Python command-line media sorting tool that organizes photos and videos from a camera into year/month/day folder structures ba… | 32 | 1009 | active |
| waifu2x (nunif) waifu2x is an image super-resolution and noise-reduction tool for anime-style art (and photos) using deep convolutional neural networks, or… | 78 | 3418 | maintenance |
| shaoanlu/faceswap-GAN A Jupyter Notebook-based implementation of face swapping using a denoising autoencoder architecture enhanced with adversarial losses, VGGFa… | 32 | 3416 | maintenance |
| facebookresearch/Mask2Former Mask2Former is the official PyTorch implementation of the CVPR 2022 paper 'Masked-attention Mask Transformer for Universal Image Segmentati… | 10 | 3416 | maintenance |
| nashaofu/xcap XCap is a cross-platform screen capture library written in Rust supporting Linux (X11, Wayland), macOS, Windows, and HarmonyOS. It provides… | 98 | 1007 | active |
| tomohiron907/Strecs3D Strecs3D is a desktop preprocessor that optimizes 3D printing infill using built-in FEM stress analysis, assigning dense infill to high-str… | 59 | 1007 | active |
| MeiGen-AI/PosterCraft PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight… | 48 | 1007 | active |
| libtv-labs/libtv-skills A collection of AI agent skill packages that expose LibLib.tv's AIGC capabilities (AI image and video generation) via its OpenAPI. It follo… | 47 | 1007 | active |
| jhfmat/ISP-pipeline-hdrplus A C/C++ image processing library (Matlib) implementing a fast ISP pipeline with HDR+ multi-frame denoising, super-low-light processing, and… | 32 | 1005 | active |
| meetps/pytorch-semseg A PyTorch library implementing popular semantic segmentation architectures such as FCN, U-Net, SegNet, PSPNet, ICNet, FRRN, and LinkNet, wi… | 23 | 3402 | maintenance |
| neelabo/NeeView NeeView is a free, open-source Windows image viewer that lets you browse images in folders and compressed archives like flipping through a … | 85 | 1004 | active |
| Kosinkadink/ComfyUI-Advanced-ControlNet A set of ComfyUI custom nodes providing advanced ControlNet scheduling, weighting, and masking for Stable Diffusion workflows. It supports … | 71 | 1004 | active |
| lumina-layer-studio/Lumina-Layers Lumina Studio is a Python/Gradio application that converts images into slicer-ready multi-material 3D models for full-color FDM printing. I… | 67 | 1004 | active |
| dlbeer/quirc Quirc is a small, dependency-free C library for extracting and decoding QR codes from images, fast enough for realtime video. It handles ro… | 42 | 1004 | active |
| clovaai/CRAFT-pytorch Official PyTorch implementation of CRAFT (Character Region Awareness for Text Detection), a scene text detector that localizes text by pred… | 32 | 3398 | maintenance |
| AkiraBit/PicSharp PicSharp is a modern, cross-platform desktop application for high-performance image compression, built with Tauri and TypeScript. It suppor… | 59 | 1003 | active |
| aras-p/UnityGaussianSplatting A Unity package implementing real-time visualization of 3D Gaussian Splatting models from the SIGGRAPH 2023 paper. It imports PLY and SPZ s… | 47 | 3389 | maintenance |
| fogleman/ln A vector-based 3D rendering engine written in Go that produces 2D vector graphics (SVG/PNG) depicting 3D scenes as line art. It was created… | 32 | 3373 | maintenance |
| mit-han-lab/efficientvit A collection of efficient vision foundation models from MIT Han Lab, including EfficientViT backbones for perception, EfficientViT-SAM for … | 48 | 3354 | maintenance |
| eladrich/pixel2style2pixel Official PyTorch implementation of pixel2style2pixel (pSp), a StyleGAN encoder from CVPR 2021 that maps real images directly into the W+ la… | 32 | 3350 | maintenance |
| shelhamer/fcn.berkeleyvision.org Reference implementation of Fully Convolutional Networks (FCN) for semantic segmentation from the CVPR 2015 / PAMI 2016 papers, built on Ca… | 32 | 3350 | maintenance |
| tamarott/SinGAN Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat… | 32 | 3344 | maintenance |
| NVlabs/eg3d Official PyTorch implementation of EG3D, an efficient geometry-aware 3D generative adversarial network from NVIDIA Research (CVPR 2022). It… | 32 | 3338 | maintenance |
| run-youngjoo/SC-FEGAN SC-FEGAN is a GUI application that uses a generative adversarial network (SN-patchGAN discriminator with a U-Net generator) to edit face im… | 32 | 3329 | maintenance |
| pytorch-yolo-v3 A minimal PyTorch implementation of the YOLO v3 object detection algorithm, supporting detection on images and video with configurable reso… | 32 | 3312 | maintenance |
| PixArt-alpha/PixArt-alpha PixArt-α is a Transformer-based text-to-image diffusion model with PyTorch model definitions, pre-trained weights, and inference/training c… | 27 | 3304 | maintenance |
| chenBingX/SuperTextView SuperTextView is an Android UI library providing an enhanced TextView/View widget with built-in support for rounded corners, borders, gradi… | 23 | 3304 | maintenance |
| ksnip/ksnip ksnip is a Qt-based cross-platform screenshot tool with rich annotation features for Linux (X11 and Wayland), Windows, and macOS. It suppor… | 81 | 3302 | maintenance |
| aserbao/AndroidCamera An Android library and demo app implementing a TikTok-style custom camera with video and audio editing features such as segment recording, … | 23 | 3297 | maintenance |
| facebookresearch/vissl VISSL is Facebook AI Research's extensible, modular and scalable PyTorch library for state-of-the-art self-supervised learning with images.… | 10 | 3293 | maintenance |
| jathu/UIImageColors A Swift library that extracts the most dominant and prominent colors from UIImage and NSImage instances, in the style of iTunes artwork col… | 10 | 3275 | maintenance |
| morkt/GARbro GARbro is a Windows GUI application for browsing, extracting, and converting resources (archives, images, audio) from visual novel games. I… | 23 | 3263 | maintenance |
| zhanghang1989/ResNeSt ResNeSt is a PyTorch implementation of the Split-Attention Network, a ResNet variant that applies channel-wise attention across network bra… | 23 | 3261 | maintenance |
| kennethcachia/background-check A small vanilla JavaScript library that detects the brightness of images behind overlapping elements and applies `.background--dark` or `.b… | 23 | 3250 | maintenance |
| anandpawara/Real_Time_Image_Animation A real-time Python application that animates a still image (e.g., a portrait) using facial motion from a live camera or video file, built o… | 32 | 3248 | maintenance |
| stereobooster/react-ideal-image An adaptive React image component that lazily loads images based on viewport visibility and network conditions, with placeholders (LQIP) to… | 32 | 3244 | maintenance |
| LBXScan LBXScan is an iOS barcode and QR code scanning library that wraps the native AVFoundation API, ZXing, and ZBar engines behind a unified int… | 32 | 3238 | maintenance |
| chjj/ttystudio ttystudio is a Node.js CLI tool that records terminal sessions and compiles them directly to GIF or APNG animations without external depend… | 32 | 3237 | maintenance |
| thearn/webcam-pulse-detector A Python desktop application that estimates a person's heart rate in real time using only a webcam, by analyzing subtle color intensity cha… | 42 | 3232 | maintenance |
| Rush/Font-Awesome-SVG-PNG A Node.js CLI tool that splits Font Awesome into individual SVG and PNG icon files at various sizes and colors, with pre-generated black an… | 23 | 3227 | maintenance |
| JoePenna/Dreambooth-Stable-Diffusion A Jupyter Notebook-based implementation of Dreambooth fine-tuning for Stable Diffusion, adapted from XavierXiao's repo with tweaks for trai… | 32 | 3211 | maintenance |
| MaximeBeasse/KeyDecoder KeyDecoder is a Flutter mobile app that lets pentesters and security enthusiasts measure the bitting of a mechanical key from a photo, usin… | 23 | 3193 | maintenance |
| kciter/qart.js qart.js is a JavaScript library that merges pictures with QR codes to generate artistic, scannable QR codes rendered on canvas. It works in… | 32 | 3188 | maintenance |
| jasondu/wxa-plugin-canvas A WeChat Mini Program component that generates shareable poster images (e.g., for Moments sharing) rendered on canvas from simple JSON conf… | 23 | 3186 | maintenance |
| JuanPotato/Legofy Legofy is a Python CLI program that transforms static images and GIFs so they appear to be built out of 1x1 LEGO bricks, using palettes bas… | 32 | 3179 | maintenance |
| tilemill-project/tilemill TileMill is an open-source map design studio powered by Node.js and Mapnik, used to style and render custom maps. It runs in server mode wi… | 23 | 3152 | maintenance |
| google-research/frame-interpolation FILM is the official TensorFlow 2 implementation of a state-of-the-art frame interpolation neural network from Google Research, presented a… | 10 | 3150 | maintenance |
| amulyakhare/TextDrawable A lightweight Android library that generates letter/text-based drawable images similar to Gmail's contact avatars. It extends the Drawable … | 32 | 3141 | maintenance |
| layervault/psd.rb PSD.rb is a Ruby library for parsing Adobe Photoshop (PSD) files. It exposes the document as a manageable tree structure with access to lay… | 23 | 3116 | maintenance |
| lucasjinreal/yolov7_d2 A detectron2-based implementation of YOLOv7 that extends YOLO-style detection to instance segmentation, keypoint detection, and multi-head … | 23 | 3109 | maintenance |
| jonom/jquery-focuspoint A jQuery plugin for responsive cropping that dynamically crops images to fill available space while keeping the image's focal point visible… | 23 | 3108 | maintenance |
| madebybowtie/FlagKit FlagKit is a collection of over 250 country flag icons provided as PNG and SVG assets, plus a Swift framework and Asset Catalog for Apple p… | 23 | 3105 | maintenance |
| cysmith/neural-style-tf A TensorFlow implementation of neural style transfer based on Gatys et al.'s convolutional neural network approach, with support for video … | 32 | 3104 | maintenance |
| Tramac/awesome-semantic-segmentation-pytorch A PyTorch library providing concise, modifiable reference implementations of many semantic segmentation models such as FCN, PSPNet, DeepLab… | 32 | 3069 | maintenance |
| cvlab-columbia/zero123 Zero-1-to-3 is a research codebase and pretrained diffusion model from Columbia CVLab that changes the camera viewpoint of an object from a… | 30 | 3058 | maintenance |
| rFlex/SCRecorder SCRecorder is an Objective-C framework for iOS that provides a Vine/Instagram-style camera engine built on AVCaptureSession. It supports ta… | 32 | 3044 | maintenance |
| AlloyTeam/AlloyImage AlloyImage is a JavaScript image processing library built on HTML5 Canvas, offering multi-layer editing, 17 Photoshop-compatible blend mode… | 32 | 3022 | maintenance |
| schmich/instascan Instascan is a JavaScript library that provides real-time QR code scanning from a webcam feed in the browser, built on top of ZXing compile… | 23 | 3022 | maintenance |
| HuanTanSheng/EasyPhotos An Android photo and video picker library supporting single/multi selection, camera capture, GIF/video filtering, custom UI theming, and ad… | 23 | 3021 | maintenance |
| skip2/go-qrcode A Go library implementing a QR Code encoder that can generate QR codes as PNG images with configurable error recovery levels and colors. It… | 32 | 3012 | maintenance |
| divamgupta/image-segmentation-keras A Keras library implementing popular deep learning semantic image segmentation models including SegNet, FCN, U-Net, and PSPNet. It provides… | 23 | 3003 | maintenance |
| jfzhang95/pytorch-deeplab-xception A PyTorch implementation of the DeepLab v3+ semantic segmentation model with support for multiple backbones (Xception, ResNet, MobileNet, D… | 32 | 3000 | maintenance |
| BeauNouvelle/FaceAware A Swift extension for UIImageView on iOS that detects faces in an image and adjusts the view's focus so faces stay visible when aspect-fill… | 10 | 2996 | maintenance |
| replicate/scribble-diffusion Scribble Diffusion is an open-source Next.js web application that turns rough sketches into refined images using the ControlNet scribble mo… | 48 | 2977 | maintenance |
| biubug6/Pytorch_Retinaface A PyTorch implementation of the RetinaFace single-stage face detection model, supporting mobilenet0.25 and resnet50 backbones with pretrain… | 32 | 2976 | maintenance |
| Tencent/FaceDetection-DSFD DSFD (Dual Shot Face Detector) is Tencent Youtu's high-accuracy face detection network, released with PyTorch inference code and pretrained… | 56 | 2969 | maintenance |
| CainKernel/CainCamera An open-source Android app and set of libraries demonstrating how to build a beauty camera, image editor, and short-video editor. It implem… | 32 | 2962 | maintenance |