domain: computer-vision
2316 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| yannickl/QRCodeReader.swift QRCodeReader.swift is a Swift library for iOS that provides a simple QR code and machine-readable code scanner built on Apple's AVFoundatio… | 32 | 1337 | maintenance |
| torrvision/crfasrnn Reference implementation of CRF-RNN, an ICCV 2015 semantic image segmentation method that integrates conditional random fields into a neura… | 23 | 1335 | maintenance |
| ahmetozlu/tensorflow_object_counting_api An open-source framework built on TensorFlow and Keras that simplifies developing object counting systems. It supports cumulative counting,… | 23 | 1333 | maintenance |
| google/monster-mash Monster Mash is an open-source sketch-based 3D modeling and animation tool that lets users sketch a 2D character, inflate it into 3D, and a… | 64 | 1330 | maintenance |
| maudzung/Complex-YOLOv4-Pytorch A PyTorch implementation of Complex-YOLO, a YOLOv4-based model for real-time 3D object detection on LiDAR point clouds. It supports distrib… | 32 | 1327 | maintenance |
| google/sg2im A PyTorch research implementation of the CVPR 2018 paper 'Image Generation from Scene Graphs' by Johnson et al. It converts a structured sc… | 10 | 1325 | maintenance |
| go-opencv/go-opencv Go bindings for the OpenCV computer vision library, exposing the OpenCV 1.x C API via CGO and an experimental OpenCV 2.x C++ API (the gocv … | 32 | 1324 | maintenance |
| CharlesShang/DCNv2 A PyTorch implementation of Deformable Convolutional Networks v2 (DCNv2), providing CUDA/CPU operators for deformable convolution as a drop… | 71 | 1323 | maintenance |
| facebookresearch/moco-v3 A PyTorch implementation of MoCo v3, a self-supervised contrastive learning method for ResNet and Vision Transformer (ViT) models. It inclu… | 10 | 1323 | maintenance |
| WisconsinAIVision/yolact_edge YolactEdge is a PyTorch implementation of a real-time instance segmentation model optimized for edge devices like the NVIDIA Jetson AGX Xav… | 32 | 1322 | maintenance |
| aitorzip/PyTorch-CycleGAN A clean, readable PyTorch implementation of CycleGAN for unpaired image-to-image translation using cycle-consistent adversarial networks. I… | 32 | 1319 | maintenance |
| hukkelas/DeepPrivacy DeepPrivacy is a PyTorch-based GAN that automatically anonymizes faces in images and videos by generating realistic synthetic replacements.… | 32 | 1316 | maintenance |
| msracver/Deep-Feature-Flow Official MXNet implementation of Deep Feature Flow (CVPR 2017), an end-to-end framework for video recognition such as object detection and … | 32 | 1315 | maintenance |
| szagoruyko/wide-residual-networks Reference implementation of Wide Residual Networks (WRN), a ResNet variant that trades depth for width to train faster and achieve state-of… | 32 | 1315 | maintenance |
| PRBonn/depth_clustering A fast and robust C++ library for segmenting point clouds from Velodyne LiDAR sensors (16, 32, and 64 beam) into objects using depth cluste… | 23 | 1312 | maintenance |
| LinXueyuanStdio/LaTeX_OCR_PRO A deep-learning application that converts images of math formulas (printed, handwritten, and Chinese-mixed) into LaTeX code using a Seq2Seq… | 32 | 1310 | maintenance |
| d-li14/involution Official PyTorch implementation of the involution neural operator from the CVPR 2021 paper 'Involution: Inverting the Inherence of Convolut… | 32 | 1310 | maintenance |
| tysam-code/hlb-CIFAR10 A single-file PyTorch implementation that trains a neural network to 94% accuracy on CIFAR-10 in under 6.3 seconds on a single A100 GPU, fo… | 22 | 1310 | maintenance |
| NVlabs/prismer Official PyTorch implementation of Prismer, a data- and parameter-efficient vision-language model that ensembles pre-trained task-specific … | 30 | 1309 | maintenance |
| mit-han-lab/data-efficient-gans Official implementation of Differentiable Augmentation (DiffAugment), a NeurIPS 2020 method that improves GAN training data efficiency by a… | 32 | 1308 | maintenance |
| EricGuo5513/momask-codes Official PyTorch implementation of MoMask (CVPR 2024), a masked modeling framework that generates 3D human motion sequences from text promp… | 27 | 1308 | maintenance |
| senlinuc/caffe_ocr An experimental research project built on Caffe implementing CNN+BLSTM+CTC text recognition architectures, with modifications for LSTM, war… | 32 | 1305 | maintenance |
| pqpo/SmartCamera SmartCamera is an Android camera extension library providing a highly customizable real-time scanning module that detects whether an object… | 23 | 1305 | maintenance |
| hengli/camodocal CamOdoCal is a C++ library for automatic intrinsic and extrinsic calibration of a camera rig with multiple generic cameras and odometry. It… | 23 | 1302 | maintenance |
| digantamisra98/Mish Official repository for the Mish activation function, a self-regularized non-monotonic neural activation function published at BMVC 2020. I… | 74 | 1299 | maintenance |
| foolwood/DaSiamRPN PyTorch implementation of DaSiamRPN, an ECCV 2018 distractor-aware Siamese network for visual object tracking, winner of the VOT-18 real-ti… | 32 | 1299 | maintenance |
| ShichenLiu/SoftRas SoftRas is a PyTorch-based differentiable renderer that treats rasterization as a differentiable aggregating process over mesh triangles, e… | 56 | 1298 | maintenance |
| NVlabs/DG-Net DG-Net is a PyTorch implementation of the CVPR 2019 (Oral) paper 'Joint Discriminative and Generative Learning for Person Re-identification… | 32 | 1298 | maintenance |
| bubbliiiing/deeplabv3-plus-pytorch A PyTorch implementation of the DeepLabv3+ semantic segmentation model with MobileNetV2 and Xception backbones. It includes scripts for tra… | 23 | 1292 | maintenance |
| HumanAIGC/OutfitAnyone OutfitAnyone is an Alibaba research project providing ultra-high quality AI virtual try-on, letting users visualize any clothing on any per… | 26 | 5984 | experimental |
| NVlabs/alias-free-gan The official project page and code for Alias-Free GAN, an NVIDIA research method for aliasing-free generative adversarial networks. The imp… | 32 | 1289 | maintenance |
| zeusees/License-Plate-Detector A YOLOv5-based license plate detection model trained on the CCPD dataset and proprietary data, supporting many Chinese plate types. It prov… | 32 | 1288 | maintenance |
| apple/ml-neuman Official reference implementation of NeuMan (ECCV 2022), which reconstructs an animatable human and the background scene from a single vide… | 32 | 1286 | maintenance |
| kwea123/ngp_pl A PyTorch + CUDA implementation of NVIDIA's Instant-NGP focused on NeRF, trained with PyTorch Lightning for fast, high-quality neural radia… | 23 | 1286 | maintenance |
| diwi/PixelFlow PixelFlow is a Processing/Java library for high-performance GPU computing via GLSL shaders. It provides fluid simulation, flow-field partic… | 23 | 1284 | maintenance |
| YuanxunLu/LiveSpeechPortraits A PyTorch implementation of the SIGGRAPH Asia 2021 paper 'Live Speech Portraits', which generates photorealistic personalized talking-head … | 32 | 1283 | maintenance |
| duzexu/ARuler An iOS augmented reality app that measures distances using Apple's ARKit and SceneKit. It uses plane detection and feature point hit-testin… | 63 | 1280 | maintenance |
| OneMoreGres/ScreenTranslator Screen Translator is a desktop utility that captures a selected region of the screen, performs OCR on it, and sends the recognized text to … | 68 | 1279 | maintenance |
| MasterBin-IIAU/UNINEXT UNINEXT is the official PyTorch implementation of the CVPR 2023 paper 'Universal Instance Perception as Object Discovery and Retrieval'. It… | 30 | 1278 | maintenance |
| snap-research/articulated-animation Official research code for the CVPR 2021 paper 'Motion Representations for Articulated Animation' by Snap Research. It animates a static so… | 43 | 1277 | maintenance |
| chonyy/AI-basketball-analysis An AI-powered web app and API that analyzes basketball shots and shooting form from uploaded videos using object detection and OpenPose pos… | 32 | 1277 | maintenance |
| SullyChen/Autopilot-TensorFlow A TensorFlow implementation of Nvidia's end-to-end self-driving steering angle prediction paper (arXiv 1604.07316) with some modifications.… | 32 | 1276 | maintenance |
| j96w/DenseFusion DenseFusion is the official PyTorch implementation of the paper '6D Object Pose Estimation by Iterative Dense Fusion', which estimates the … | 32 | 1276 | maintenance |
| YavorGIvanov/sam.cpp A pure C/C++ implementation of Meta's Segment Anything Model (SAM) for image segmentation inference, built on the ggml tensor library. It c… | 28 | 1275 | maintenance |
| TRI-ML/packnet-sfm Official PyTorch implementation of PackNet and related self-supervised monocular depth estimation methods from Toyota Research Institute's … | 23 | 1274 | maintenance |
| marcbelmont/cnn-watermark-removal A TensorFlow implementation of a fully convolutional neural network that removes transparent watermark overlays from images. It trains on s… | 32 | 1269 | maintenance |
| phonegap/phonegap-plugin-barcodescanner A Cordova/PhoneGap plugin providing cross-platform barcode scanning for hybrid mobile apps. It exposes a JavaScript API (cordova.plugins.ba… | 10 | 1269 | maintenance |
| dsys/match Match is a scalable reverse image search service built on Kubernetes and Elasticsearch, using perceptual hashing to find visually similar i… | 32 | 1265 | maintenance |
| tensorlayer/HyperPose HyperPose is a library for building high-performance custom human pose estimation applications. It combines a C++ inference engine with Ten… | 23 | 1264 | maintenance |
| TimoBolkart/voca VOCA is a speech-driven 3D facial animation framework that synthesizes realistic character face animations from an arbitrary speech signal … | 32 | 1263 | maintenance |
| rohitrango/automatic-watermark-detection A Python implementation of the CVPR 2017 paper 'On The Effectiveness Of Visible Watermarks' that detects and removes visible watermarks fro… | 32 | 1263 | maintenance |
| kennymckormick/pyskl PYSKL is a PyTorch-based toolbox for skeleton-based human action recognition, built on MMAction2. It provides official implementations of P… | 53 | 1261 | maintenance |
| openpifpaf/openpifpaf OpenPifPaf is a PyTorch library implementing Composite Fields for semantic keypoint detection and spatio-temporal association, primarily fo… | 23 | 1261 | maintenance |
| BrandonJoffe/home_surveillance A self-hosted home surveillance application that processes streams from multiple IP cameras, performs motion detection and facial recogniti… | 32 | 1259 | maintenance |
| harvardnlp/im2markup A deep learning system (built on Torch) that converts images of rendered text into presentational markup such as LaTeX or HTML, using a CNN… | 32 | 1259 | maintenance |
| Yujun-Shi/DragDiffusion DragDiffusion is the official CVPR 2024 research code for interactive point-based image editing using pretrained diffusion models. It provi… | 19 | 1259 | maintenance |
| DrSleep/tensorflow-deeplab-resnet A TensorFlow re-implementation of the DeepLab-ResNet model for semantic image segmentation, trained and evaluated on the PASCAL VOC dataset… | 10 | 1258 | maintenance |
| Fantasy-Studio/Paint-by-Example Paint by Example is a PyTorch implementation of exemplar-based image editing with diffusion models, letting users fill masked regions of an… | 32 | 1252 | maintenance |
| GoGoDuck912/Self-Correction-Human-Parsing A deep learning toolkit for human parsing (semantic segmentation of clothing and body parts in images), with pretrained models on LIP, ATR,… | 32 | 1251 | maintenance |
| clovaai/CutMix-PyTorch Official PyTorch implementation of CutMix, an image data augmentation regularizer that cuts and pastes patches between training images whil… | 32 | 1251 | maintenance |
| CQFIO/PhotographicImageSynthesis A TensorFlow implementation of the ICCV 2017 paper 'Photographic Image Synthesis with Cascaded Refinement Networks', which synthesizes phot… | 32 | 1248 | maintenance |
| daijifeng001/R-FCN MATLAB implementation of R-FCN, a region-based fully convolutional object detection framework described in a NIPS 2016 paper. It builds on … | 32 | 1248 | maintenance |
| leftthomas/SRGAN A PyTorch implementation of SRGAN, the CVPR 2017 generative adversarial network for photo-realistic single-image super-resolution. It inclu… | 32 | 1248 | maintenance |
| KMnP/vpt Official PyTorch implementation of Visual Prompt Tuning (VPT), an ECCV 2022 method for parameter-efficient fine-tuning of vision transforme… | 32 | 1244 | maintenance |
| utiasSTARS/pykitti pykitti is a minimal Python library for loading and working with the KITTI autonomous driving dataset, supporting raw and odometry benchmar… | 32 | 1244 | maintenance |
| ranahanocka/point2mesh Point2Mesh is a PyTorch implementation of a SIGGRAPH 2020 technique that reconstructs watertight surface meshes from input point clouds by … | 32 | 1240 | maintenance |
| autonomousvision/giraffe Official PyTorch implementation of GIRAFFE, a CVPR 2021 (oral, best paper award) generative model that represents scenes as compositional g… | 32 | 1238 | maintenance |
| BR-IDL/PaddleViT PaddleViT is a collection of state-of-the-art Vision Transformer and MLP model implementations for PaddlePaddle 2.1+, covering image classi… | 23 | 1238 | maintenance |
| openseg-group/openseg.pytorch Official PyTorch implementations of semantic segmentation models OCNet, OCRNet, and SegFix, achieving state-of-the-art results on benchmark… | 23 | 1237 | maintenance |
| facebookresearch/simsiam A PyTorch implementation of SimSiam, a simple Siamese-based self-supervised representation learning method from Facebook AI Research (CVPR … | 10 | 1237 | maintenance |
| jfzhang95/pytorch-video-recognition A PyTorch library implementing C3D, R3D, and R2Plus1D models for video action recognition, with training scripts for UCF101 and HMDB51 data… | 32 | 1236 | maintenance |
| otaha178/Emotion-recognition A Python application that performs real-time facial emotion recognition from a webcam feed using a convolutional neural network. It display… | 32 | 1236 | maintenance |
| ultralytics/JSON2YOLO A legacy Python toolkit that converts JSON-format annotation datasets (COCO, LabelMe, Labelbox, VoTT, INFOLKS, ATH) into the YOLO format fo… | 67 | 1231 | maintenance |
| qubvel/classification_models A Keras/TensorFlow Keras library providing ImageNet-pretrained image classification model architectures such as ResNet, SE-ResNet, DenseNet… | 23 | 1231 | maintenance |
| ju1ce/April-Tag-VR-FullBody-Tracker A free open-source application that provides full-body tracking in VR using printed AprilTag fiducial markers tracked by a phone or PS Eye … | 23 | 1229 | maintenance |
| huoyijie/AdvancedEAST AdvancedEAST is a deep learning algorithm for detecting text in scene images, built on the EAST architecture with improvements for more acc… | 32 | 1227 | maintenance |
| lukasvst/dm-vio DM-VIO is a C++ implementation of Delayed Marginalization Visual-Inertial Odometry, a direct sparse visual-inertial odometry method publish… | 32 | 1224 | maintenance |
| tylin/coco-caption The official evaluation code for the Microsoft COCO image captioning benchmark, implementing metrics such as BLEU, METEOR, ROUGE-L, CIDEr, … | 32 | 1224 | maintenance |
| xvjiarui/GCNet Official PyTorch implementation of GCNet (Global Context Networks), a paper combining ideas from Non-Local Networks and Squeeze-Excitation … | 32 | 1220 | maintenance |
| GeekAlexis/FastMOT FastMOT is a high-performance multiple object tracking system combining YOLO/SSD detection, Deep SORT with OSNet ReID, and KLT optical flow… | 23 | 1220 | maintenance |
| IBM/MicroscoPy An open-source, motorized, modular microscope built from LEGO bricks, 3D-printed parts, an Arduino, and a Raspberry Pi with an 8MP camera. … | 10 | 1220 | maintenance |
| xingyizhou/CenterNet2 CenterNet2 is a research implementation of probabilistic two-stage object detection built on detectron2, where a class-agnostic one-stage C… | 32 | 1219 | maintenance |
| ucbdrive/few-shot-object-detection FsDet is the official implementation of the ICML 2020 paper 'Frustratingly Simple Few-Shot Object Detection' (TFA), built on detectron2. It… | 23 | 1218 | maintenance |
| The-AI-Summer/self-attention-cv A PyTorch library implementing various self-attention mechanisms and transformer building blocks for computer vision, including multi-head … | 23 | 1214 | maintenance |
| nihui/realsr-ncnn-vulkan A command-line tool implementing the RealSR real-world super-resolution model using the ncnn inference framework with Vulkan GPU accelerati… | 23 | 1213 | maintenance |
| UMass-Embodied-AGI/3D-LLM 3D-LLM is the research code for a large language model that takes 3D representations (objects and scenes) as input, built on BLIP-2/LAVIS. … | 28 | 1212 | maintenance |
| uzh-rpg/rpg_trajectory_evaluation A Python toolbox for quantitatively evaluating visual(-inertial) odometry trajectories, supporting multiple alignment methods and standard … | 23 | 1212 | maintenance |
| NVlabs/VoxFormer Official PyTorch implementation of VoxFormer, a CVPR 2023 highlight paper presenting a sparse voxel transformer for camera-based 3D semanti… | 31 | 1208 | maintenance |
| facebookresearch/ToMe ToMe (Token Merging) is a PyTorch library from Meta AI that speeds up existing Vision Transformers by merging similar tokens inside the net… | 10 | 1207 | maintenance |
| YuliangXiu/ECON ECON is a research tool that reconstructs high-fidelity 3D clothed human avatars from a single color image by combining implicit and explic… | 32 | 1206 | maintenance |
| reiinakano/arbitrary-image-stylization-tfjs A browser-based implementation of arbitrary neural style transfer built with TensorFlow.js, letting users stylize any photo in the style of… | 32 | 1205 | maintenance |
| enazoe/yolo-tensorrt A C++ wrapper around NVIDIA TensorRT for running YOLO object detection models (YOLOv3, YOLOv4, YOLOv5) with support for FP32, FP16, and INT… | 57 | 1202 | maintenance |
| facebookresearch/House3D House3D is a virtual 3D environment of over 45k fully annotated indoor scenes from the SUNCG dataset, built for training embodied AI agents… | 10 | 1200 | maintenance |
| lucidrains/CoCa-pytorch A Pytorch implementation of CoCa (Contrastive Captioners), an image-text foundation model that combines contrastive learning with an encode… | 23 | 1198 | maintenance |
| ethan-li-coding/SemiGlobalMatching A complete, well-commented C++ implementation of the classic Semi-Global Matching (SGM) stereo matching algorithm for computing disparity a… | 32 | 1197 | maintenance |
| rotemtzaban/STIT STIT (Stitch it in Time) is a research implementation of a GAN-based framework for semantic editing of faces in real videos, based on the p… | 32 | 1197 | maintenance |
| facebookresearch/mixup-cifar10 Official PyTorch implementation of mixup, a data augmentation technique that trains neural networks on convex combinations of image pairs a… | 10 | 1197 | maintenance |
| KaihuaTang/Scene-Graph-Benchmark.pytorch A PyTorch codebase for Scene Graph Generation (SGG) built on maskrcnn-benchmark, implementing methods from the CVPR 2020 paper 'Unbiased Sc… | 59 | 1195 | maintenance |
| philipperemy/yolo-9000 A packaging of YOLO9000 (YOLOv2), a real-time object detection model that can detect over 9,000 object classes, built around the Darknet fr… | 32 | 1194 | maintenance |
| Kagami/go-face A Go library that provides face recognition by wrapping dlib's machine learning models, including face detection, landmark prediction, and … | 32 | 1192 | maintenance |