function: image-processing
4273 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| HRNet/HRNet-Image-Classification Official PyTorch implementation and training code for HRNet (High-Resolution Network) image classification models on ImageNet. It provides … | 23 | 1056 | maintenance |
| keras-team/keras-cv KerasCV is a library of modular computer vision components built on Keras 3 that work natively with TensorFlow, JAX, or PyTorch. It provide… | 10 | 1055 | maintenance |
| BerndK/SvgToXaml A hybrid WPF/console tool for viewing SVG files and converting them to XAML resources for .NET applications. It supports interactive browsi… | 23 | 1054 | maintenance |
| caoscott/SReC SReC is the official PyTorch implementation of the paper 'Lossless Image Compression through Super-Resolution', which frames lossless image… | 32 | 1051 | maintenance |
| spicyShrimp/U17 A Swift 5 iOS application that closely replicates the U17 (Youyaoqi) comic reading app, built with UIKit and popular third-party libraries … | 32 | 1050 | maintenance |
| patrickfav/Dali Dali is an Android image blur library offering static blurring, live blurring, and blur animations. It uses RenderScript internally with ca… | 23 | 1050 | maintenance |
| microsoft/SimMIM Official PyTorch implementation of SimMIM, a simple framework for masked image modeling (self-supervised visual pre-training) from Microsof… | 32 | 1048 | maintenance |
| yuval-alaluf/restyle-encoder Official PyTorch implementation of ReStyle, a residual-based StyleGAN encoder that inverts real images into GAN latent codes via iterative … | 32 | 1046 | maintenance |
| 4uiiurz1/pytorch-nested-unet A PyTorch implementation of the UNet++ (Nested U-Net) architecture for image segmentation, based on the paper 'UNet++: A Nested U-Net Archi… | 32 | 1045 | maintenance |
| Silence-GitHub/BBMetalImage A high-performance Swift library for GPU-accelerated image and video processing built on Apple's Metal, inspired by GPUImage. It provides 8… | 32 | 1044 | maintenance |
| pkhungurn/talking-head-anime-3-demo Demo programs for the Talking Head(?) Anime 3 project, which animates an anime character from a single image using machine learning. It inc… | 32 | 1044 | maintenance |
| facebookresearch/FixRes FixRes is a PyTorch implementation of the NeurIPS 2019 paper 'Fixing the train-test resolution discrepancy', providing training and fine-tu… | 10 | 1043 | maintenance |
| pixray/pixray Pixray is a Python library and command-line utility for text-to-image generation, combining CLIP-guided GAN imagery, pixel-art drawers, and… | 23 | 1042 | maintenance |
| asingh33/CNNGestureRecognizer A desktop application that recognizes hand gestures from webcam video using a convolutional neural network built with Keras, TensorFlow/The… | 60 | 1041 | maintenance |
| ithewei/hplayer A multi-screen video player built with Qt, FFmpeg, OpenCV, and OpenGL that plays files, network streams, and capture devices in a configura… | 32 | 1041 | maintenance |
| SysCV/sam-pt SAM-PT extends the Segment Anything Model to zero-shot video segmentation by combining SAM with sparse point-based tracking (PIPS, CoTracke… | 29 | 1041 | maintenance |
| keijiro/Pix2Pix A Unity library that runs pix2pix image-to-image translation neural networks in real time using compute shaders. It includes its own infere… | 23 | 1041 | maintenance |
| stathissideris/ditaa ditaa is a small Java command-line utility that converts diagrams drawn using ASCII art characters (like | / -) into proper bitmap graphics… | 23 | 1041 | maintenance |
| IBM/MAX-Image-Resolution-Enhancer An IBM Model Asset Exchange project that deploys an SRGAN-based image super-resolution model as a web service in a Docker container. It ups… | 42 | 1040 | maintenance |
| draeton/stitches Stitches is an HTML5 sprite sheet generator that combines multiple images into a single sprite sheet with CSS coordinates using a browser-b… | 32 | 1040 | maintenance |
| MaybeShewill-CV/CRNN_Tensorflow A TensorFlow implementation of CRNN (CNN + Bi-LSTM + CTC loss) for scene text recognition, based on the Shi et al. paper. It includes pretr… | 32 | 1039 | maintenance |
| huggingface/pytorch-pretrained-BigGAN A PyTorch reimplementation of DeepMind's BigGAN generator with pretrained weights at 128, 256, and 512 pixel resolutions, plus scripts to c… | 23 | 1039 | maintenance |
| JIA-Lab-research/SNR-Aware-Low-Light-Enhance Official PyTorch implementation of the CVPR 2022 paper 'SNR-aware Low-Light Image Enhancement'. It combines SNR-aware transformers and conv… | 32 | 1037 | maintenance |
| facebook/transform360 Transform360 is a C++ video/image filter that converts 360-degree video between projections, typically from equirectangular to cubemap form… | 69 | 1034 | maintenance |
| xsahil03x/before_after A Flutter package providing a BeforeAfter widget that displays the difference between two images via a draggable slider. It is 100% Dart, s… | 23 | 1034 | maintenance |
| pesser/stable-diffusion The development repository for Stable Diffusion and Latent Diffusion Models, containing research code, training scripts, and pretrained mod… | 32 | 1033 | maintenance |
| antimatter15/whammy Whammy is a small JavaScript library that encodes canvas frames into WebM video files entirely in the browser by reusing WebP-encoded VP8 i… | 32 | 1032 | maintenance |
| xingyizhou/ExtremeNet Official PyTorch implementation of ExtremeNet, a CVPR 2019 bottom-up object detection method that detects four extreme points and one cente… | 32 | 1031 | maintenance |
| ArrowLuo/CLIP4Clip Official PyTorch implementation of the CLIP4Clip paper, a video-text retrieval model that transfers CLIP knowledge to end-to-end video clip… | 23 | 1031 | maintenance |
| kakaobrain/rq-vae-transformer The official PyTorch implementation of 'Autoregressive Image Generation using Residual Quantization' (CVPR 2022), implementing RQ-VAE and R… | 32 | 1030 | maintenance |
| qubvel/ttach TTAch is a Python library for image test time augmentation (TTA) with PyTorch. It wraps existing models to apply augmentations like flips, … | 23 | 1030 | maintenance |
| antwankakki/FabricView FabricView is an Android canvas drawing library inspired by Fabric.js, supporting hand/stylus drawing, text, and images. It lets developers… | 23 | 1029 | maintenance |
| tsulej/GenerateMe A collection of Processing (Java-based) scripts for creating generative glitch art, image distortion, and design effects. It includes dozen… | 32 | 1027 | maintenance |
| yuval-alaluf/hyperstyle Official PyTorch implementation of HyperStyle (CVPR 2022), a hypernetwork that inverts real images into editable regions of StyleGAN's late… | 32 | 1027 | maintenance |
| keijiro/KinoBloom KinoBloom is a bloom/veiling glare image effect for Unity implemented as a shader-based post-processing effect. It offers configurable thre… | 23 | 1027 | maintenance |
| sniklaus/sepconv-slomo A reference PyTorch implementation of Video Frame Interpolation via Adaptive Separable Convolution, which generates intermediate frames bet… | 43 | 1021 | maintenance |
| Gumpest/YOLOv5-Multibackbone-Compression A YOLOv5-based toolbox for swapping in lightweight or high-accuracy backbones (TPH-YOLOv5, GhostNet, ShuffleNetV2, MobileNetV3-Small, Effic… | 32 | 1020 | maintenance |
| neeru1207/AI_Sudoku A Python desktop application with a Tkinter GUI that extracts a Sudoku puzzle from a photo using OpenCV image processing and solves it. Dig… | 32 | 1020 | maintenance |
| NaturalIntelligence/imglab ImgLab is a browser-based image annotation tool for labeling objects and landmark points to train object detectors like dlib. It supports m… | 76 | 1019 | maintenance |
| rmislam/PythonSIFT A pure Python/NumPy implementation of SIFT (Scale-Invariant Feature Transform) that returns OpenCV KeyPoint objects and descriptors, making… | 48 | 1019 | maintenance |
| brain-research/self-attention-gan A TensorFlow implementation of Self-Attention GANs for reproducing results from the paper 'Self-Attention Generative Adversarial Networks' … | 10 | 1019 | maintenance |
| EvgenyKashin/stylegan2-distillation A research implementation of the ECCV 2020 paper 'StyleGAN2 Distillation for Feed-forward Image Manipulation', distilling StyleGAN2 latent-… | 32 | 1018 | maintenance |
| iddan/react-native-canvas A React Native component that provides an HTML5-like Canvas API for drawing 2D graphics in mobile apps, implemented via a WebView bridge. I… | 32 | 1018 | maintenance |
| dkern/jquery.lazy jQuery Lazy is a lightweight, highly configurable lazy-loading plugin for jQuery and Zepto that delays loading of images, background images… | 23 | 1018 | maintenance |
| elevateweb/elevatezoom elevateZoom is a jQuery plugin that adds image zoom functionality to web pages, typically for product detail views. It magnifies a small im… | 32 | 1017 | maintenance |
| trishume/eyeLike eyeLike is an OpenCV-based C++ implementation of Fabian Timm's gradient-based eye center localization algorithm for webcam pupil tracking. … | 32 | 1015 | maintenance |
| torrinworx/Blend_My_NFTs Blend_My_NFTs is a free, open-source Blender add-on that automatically generates thousands of 3D models, images, and animations from user-d… | 23 | 1015 | maintenance |
| FaceTracker ofxFaceTracker is an openFrameworks addon for real-time non-rigid face tracking, based on Jason Saragih's FaceTracker C++ library and OpenC… | 10 | 1014 | maintenance |
| everestpipkin/image-scrubber A browser-based tool for anonymizing photographs taken at protests by stripping Exif metadata and letting users paint over or blur faces an… | 32 | 1012 | maintenance |
| snap-research/NeROIC Official PyTorch implementation of NeROIC, a neural method for capturing 3D object geometry and material from online image collections and … | 32 | 1012 | maintenance |
| afollestad/photo-affix PhotoAffix is an open-source Android app for stitching photos together vertically or horizontally to create side-by-side collage images. It… | 10 | 1011 | maintenance |
| zju3dv/OnePose OnePose is the official PyTorch implementation of the CVPR 2022 paper 'One-Shot Object Pose Estimation without CAD Models'. It estimates th… | 32 | 1010 | maintenance |
| iamvucms/react-native-instagram-clone A React Native mobile application that clones the Instagram mobile app, built with TypeScript, Firebase, Redux, and React Navigation. It in… | 32 | 1009 | maintenance |
| zhanghang1989/PyTorch-Multi-Style-Transfer A PyTorch implementation of MSG-Net and Gatys et al. neural style transfer for applying artistic styles to images in real time. It includes… | 23 | 1009 | maintenance |
| xiaoyufenfei/Efficient-Segmentation-Networks A PyTorch reference implementation collection of lightweight, real-time semantic segmentation models such as ENet, ERFNet, LEDNet, Fast-SCN… | 32 | 1008 | maintenance |
| PeterWang512/CNNDetection A PyTorch research codebase with pretrained models for detecting CNN-generated (GAN/synthetic) images, from the CVPR 2020 paper 'CNN-genera… | 32 | 1005 | maintenance |
| kevinzakka/spatial-transformer-network A TensorFlow implementation of Spatial Transformer Networks, a differentiable module that can be inserted into ConvNet architectures to add… | 32 | 1005 | maintenance |
| CNOliverZhang/PotatofieldImageToolkit Potatofield Image Toolkit is an Electron-based desktop image toolbox for photographers, designers, and other creative professionals. It bun… | 23 | 1005 | maintenance |
| mitallast/diablo-js An isometric minimal-code style game rendered on HTML5 canvas with JavaScript, recreating a level from Diablo 2. It includes tools for extr… | 32 | 1004 | maintenance |
| alex04072000/ObstructionRemoval The official TensorFlow implementation of the CVPR 2020 paper 'Learning to See Through Obstructions', which removes obstructions like windo… | 32 | 1004 | maintenance |
| johannakarras/DreamPose Official PyTorch implementation of DreamPose, a Stable Diffusion-based model that synthesizes animated fashion videos from a single image a… | 30 | 1004 | maintenance |
| emersion/grim grim is a small C utility for taking screenshots on Wayland compositors, supporting full-screen, per-output, and region captures. It integr… | 10 | 1004 | maintenance |
| shaoshengsong/DeepSORT A C++ implementation of multi-object tracking (MOT) combining YOLOv5 object detection with DeepSORT and ByteTrack trackers. It uses ONNX Ru… | 32 | 1003 | maintenance |
| Algebra-FUN/WeReadScan A Python library that uses Selenium headless browsers to scan purchased books from WeRead (WeChat Reading) and convert them into local PDF … | 32 | 1002 | maintenance |
| dawnlabs/alchemy Alchemy is an open-source desktop file converter built with Electron and React that lives in the macOS/Windows menu bar. It lets users drag… | 23 | 1002 | maintenance |
| alibaba/simpleimage SimpleImage is Alibaba's open-source Java image processing library, providing common operations such as scaling, cropping, rotation, format… | 32 | 1001 | maintenance |
| linebender/vello Vello is a GPU compute-centric 2D vector graphics rendering engine written in Rust, built on wgpu. It renders large 2D scenes (shapes, grad… | 95 | 4286 | experimental |
| apple/ml-mgie MGIE (MLLM-Guided Image Editing) is Apple's research implementation of instruction-based image editing guided by multimodal large language … | 26 | 3874 | experimental |
| guoqincode/Open-AnimateAnyone An unofficial PyTorch implementation of Animate Anyone, a diffusion-based method that animates a static character image using pose sequence… | 26 | 2923 | experimental |
| MewPurPur/GodSVG GodSVG is a free, open-source structured SVG editor built with Godot that represents SVG code directly and lets users edit vector graphics … | 82 | 2708 | experimental |
| google/forma Forma is an experimental, thoroughly parallelized vector-graphics renderer written in Rust with both CPU (SIMD + Rayon) and GPU (wgpu/WebGP… | 10 | 2641 | experimental |
| jasonjmcghee/rem rem is an open-source macOS application that locally records everything you view on your screen by taking periodic screenshots and running … | 17 | 2485 | experimental |
| tsoding/olive.c Olive.c is a dependency-free, stb-style single-header 2D graphics library for C that renders pixel by pixel into a memory buffer. It provid… | 52 | 2444 | experimental |
| AIGCDesignGroup/ReplaceAnything ReplaceAnything is a research project from Alibaba's Institute for Intelligent Computing for ultra-high quality content replacement in imag… | 26 | 2426 | experimental |
| lllyasviel/LayerDiffuse LayerDiffuse is a research project that generates transparent images and image layers using diffusion models with latent transparency. It p… | 25 | 2221 | experimental |
| JiauZhang/DragGAN A Python implementation of DragGAN, a research method for interactively manipulating generated images by dragging points on the generative … | 29 | 2128 | experimental |
| QwenLM/Qwen-Image-Layered Qwen-Image-Layered is a diffusion-based model and pipeline that decomposes an input image into multiple independently editable RGBA layers.… | 43 | 2079 | experimental |
| lucidrains/make-a-video-pytorch A PyTorch library implementing Make-A-Video, Meta AI's text-to-video generation approach, built around pseudo-3d (axial) convolutions and s… | 23 | 1986 | experimental |
| ShieldMnt/invisible-watermark A Python library and command line tool for embedding and decoding invisible (blind) image watermarks that do not require the original image… | 23 | 1972 | experimental |
| photonixapp/photonix Photonix is a self-hosted, web-based photo management server built with Django and React. It ingests your photo collection and enables smar… | 65 | 1955 | experimental |
| PiLiDAR/PiLiDAR PiLiDAR is a DIY 360° 3D panorama scanner built on Raspberry Pi that combines a low-cost LDRobot LiDAR (LD06/LD19/STL27L) with a Pi HQ came… | 59 | 1954 | experimental |
| lucidrains/gigagan-pytorch A PyTorch implementation of GigaGAN, Adobe's state-of-the-art generative adversarial network for text-to-image and unconditional image synt… | 21 | 1942 | experimental |
| magic-research/magic-edit MagicEdit is a research implementation of a diffusion-based video editing model from ByteDance that disentangles appearance and motion for … | 10 | 1790 | experimental |
| varunshenoy/opendream Opendream is a web UI for Stable Diffusion that adds layering, non-destructive editing, portable workflow files, and a simple extension sys… | 29 | 1670 | experimental |
| Anything-of-anything/Anything-3D Anything-3D is a Python research project that combines Meta's Segment Anything model with a series of 3D models (3DFuse, Zero 1-to-3, NeRF,… | 30 | 1633 | experimental |
| Ildaron/Laser_control An open-source hardware and software project that uses a camera, deep learning object detection (Darknet/YOLO via OpenCV), and galvanometer… | 66 | 1601 | experimental |
| maplibre/maplibre-rs maplibre-rs is a portable vector map renderer written in Rust that uses WebGPU for cross-platform rendering on web, mobile, and desktop. It… | 76 | 1568 | experimental |
| ali-vilab/composer Official implementation of Composer, a 5-billion-parameter controllable diffusion model for creative image synthesis using composable condi… | 31 | 1557 | experimental |
| KUR-creative/SickZil-Machine SickZil-Machine is a desktop application that automates text removal from manga and comic pages during the scanlation (translation) process… | 23 | 1524 | experimental |
| graphdeco-inria/hierarchical-3d-gaussians Official implementation of the SIGGRAPH 2024 paper 'A Hierarchical 3D Gaussian Representation for Real-Time Rendering of Very Large Dataset… | 35 | 1461 | experimental |
| fizzyedit/fizzy Fizzy is a cross-platform, open-source modular editor written in Zig that starts as an empty shell and loads compiled plugins to provide fu… | 98 | 1457 | experimental |
| edluffy/hologram.nvim A Neovim plugin that displays images inline in the terminal using the Kitty Graphics Protocol. It is written in Lua and C, exposing an exte… | 32 | 1435 | experimental |
| OpnTec/mvisc MVISC (Mobile Visual Classification) is an application that identifies and classifies individual animals from photos using computer vision,… | 32 | 1384 | experimental |
| lizhihao6/Sparc3D Sparc3D is the official implementation of a research framework for high-resolution 3D shape modeling, combining a sparse deformable marchin… | 31 | 1353 | experimental |
| justjake/Gauss Gauss is a native macOS Stable Diffusion app built with SwiftUI and Apple's ml-stable-diffusion CoreML models. It is document-based, storin… | 22 | 1350 | experimental |
| Aliothmoon/MAA-Meow MAA Meow is an Android application that natively runs the MAA (MaaAssistantArknights) automation core on-device, using image recognition to… | 82 | 1349 | experimental |
| google/style-aligned Official research code for 'Style Aligned Image Generation via Shared Attention', implementing style-consistent image generation with diffu… | 10 | 1315 | experimental |
| jlsutherland/doc2text doc2text is a Python library that extracts high-quality text from poorly scanned PDFs by correcting resolution, cropping, and skew before O… | 32 | 1278 | experimental |
| zhouxiyu1997/friendmaker Friend Maker is a desktop application (macOS/Windows) that converts images into pixel grids and controller action scripts, then drives an E… | 79 | 1262 | experimental |
| paradigms-of-intelligence/swissgl SwissGL is a minimalistic JavaScript wrapper around the WebGL2 API that reduces boilerplate for managing GLSL shaders, textures, and frameb… | 63 | 1242 | experimental |