domain: computer-vision
2316 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| DeepVoltaire/AutoAugment An unofficial Python implementation of the AutoAugment data augmentation policies learned for ImageNet, CIFAR-10, and SVHN, built on Pillow… | 32 | 1493 | maintenance |
| ageitgey/show-facebook-computer-vision-tags A simple Chrome extension that overlays the automated computer vision tags Facebook generates for images directly on your Facebook timeline… | 32 | 1488 | maintenance |
| chengdazhi/Deformable-Convolution-V2-PyTorch A PyTorch implementation of Deformable Convolution V2 (DCNv2) custom CUDA operators, ported from the original MXNet implementation. It prov… | 32 | 1484 | maintenance |
| speedinghzl/CCNet Official PyTorch implementation of CCNet, a Criss-Cross Attention network for semantic segmentation published at ICCV 2019 and TPAMI 2020. … | 32 | 1484 | maintenance |
| cvg/pixel-perfect-sfm pixsfm is a Python package with a C++ core that improves Structure-from-Motion and visual localization accuracy by refining keypoints, came… | 23 | 1484 | maintenance |
| AIRLegend/aitrack AITrack is a free 6DoF head tracking application that uses a webcam and neural networks to estimate head position and rotation. It streams … | 23 | 1482 | maintenance |
| chainer/chainercv ChainerCV is a Python library providing tools to train and run neural networks for computer vision tasks on top of the Chainer framework. I… | 10 | 1481 | maintenance |
| Sharpiless/Yolov5-deepsort-inference A Python library combining YOLOv5 object detection with DeepSort multi-object tracking to detect, track, and count vehicles and pedestrians… | 66 | 1479 | maintenance |
| ruotianluo/ImageCaptioning.pytorch A PyTorch research codebase for image captioning, supporting self-critical sequence training, bottom-up features, transformer captioning mo… | 32 | 1476 | maintenance |
| faustomorales/keras-ocr A Python library packaging the CRAFT text detector and a Keras CRNN text recognition model into a high-level OCR pipeline. It supports pret… | 42 | 1473 | maintenance |
| shouzhong/Scanner An Android scanning library that recognizes QR codes, barcodes, ID cards, bank cards, license plates, text, driving licenses, and NSFW imag… | 23 | 1472 | maintenance |
| peteanderson80/bottom-up-attention A bottom-up attention model based on Faster R-CNN with ResNet-101 trained on Visual Genome, producing features for salient image regions. T… | 32 | 1469 | maintenance |
| CSAILVision/LabelMeAnnotationTool The source code for LabelMe, a web-based image annotation tool written in JavaScript that runs on your own Apache server. It lets users dra… | 32 | 1468 | maintenance |
| mmatl/pyrender Pyrender is a pure Python, glTF 2.0-compliant OpenGL renderer for physically-based rendering and visualization of 3D scenes. It includes an… | 34 | 1466 | maintenance |
| lxztju/pytorch_classification A complete PyTorch image classification codebase built on torchvision, covering training, prediction, TTA, model ensembling, knowledge dist… | 32 | 1466 | maintenance |
| sxyu/pixel-nerf Official PyTorch implementation of pixelNeRF, a CVPR 2021 method that predicts neural radiance fields conditioned on one or few input image… | 32 | 1465 | maintenance |
| Sanster/text_renderer A Python tool that generates synthetic text images with configurable visual effects for training deep learning OCR models like CRNN. It sup… | 32 | 1464 | maintenance |
| HRNet/HigherHRNet-Human-Pose-Estimation Official PyTorch implementation of HigherHRNet, a CVPR 2020 bottom-up multi-person human pose estimation model using scale-aware high-resol… | 32 | 1463 | maintenance |
| szagoruyko/attention-transfer PyTorch reference implementation of the ICLR 2017 paper 'Paying More Attention to Attention', which improves convolutional neural networks … | 32 | 1463 | maintenance |
| datitran/face2face-demo A pix2pix demo that learns from facial landmarks and translates them into a trained target face, with a real-time webcam application. It in… | 32 | 1461 | maintenance |
| facebookresearch/MaskFormer MaskFormer is a PyTorch/Detectron2-based implementation of the NeurIPS 2021 paper 'Per-Pixel Classification is Not All You Need for Semanti… | 10 | 1460 | maintenance |
| una-dinosauria/3d-pose-baseline A TensorFlow implementation of a simple yet effective baseline for 3D human pose estimation from 2D keypoints, published at ICCV 2017. It i… | 32 | 1459 | maintenance |
| nerfstudio-project/nerfacc NerfAcc is a PyTorch acceleration toolbox for Neural Radiance Fields (NeRF), focused on efficient sampling in the volumetric rendering pipe… | 23 | 1456 | maintenance |
| wvangansbeke/Unsupervised-Classification PyTorch implementation of SCAN (ECCV 2020), a two-step method for unsupervised image classification that combines self-supervised represent… | 32 | 1455 | maintenance |
| buaacyw/GaussianEditor GaussianEditor is a research application for fast, controllable 3D scene editing built on Gaussian Splatting, released as CVPR 2024 code. I… | 27 | 1455 | maintenance |
| abreheret/PixelAnnotationTool A C++ desktop application for quickly annotating images with pixel-wise labels using OpenCV's watershed algorithm. Users draw markers with … | 23 | 1455 | maintenance |
| xuannianz/EfficientDet A Keras/TensorFlow implementation of the EfficientDet object detection model with pretrained COCO and ImageNet weights. It supports trainin… | 32 | 1454 | maintenance |
| dji-sdk/Tello-Python A collection of Python sample modules for controlling and interacting with the Ryze Tello drone, including command scripting, video streami… | 32 | 1451 | maintenance |
| hysts/pytorch_image_classification A PyTorch library implementing many image classification architectures (ResNet, DenseNet, WRN, PyramidNet, SENet, etc.) and augmentation te… | 10 | 1449 | maintenance |
| mit-han-lab/proxylessnas ProxylessNAS is a neural architecture search (NAS) framework that directly searches CNN architectures on the target task and target hardwar… | 32 | 1447 | maintenance |
| megvii-research/ML-GCN A PyTorch implementation of ML-GCN, the CVPR 2019 paper 'Multi-Label Image Recognition with Graph Convolutional Networks'. It provides trai… | 32 | 1446 | maintenance |
| BachiLi/redner redner is a differentiable Monte Carlo ray tracer that computes exact gradients of rendered images with respect to arbitrary scene paramete… | 32 | 1444 | maintenance |
| toandaominh1997/EfficientDet.Pytorch A PyTorch implementation of EfficientDet, a scalable and efficient object detection model from a 2019 Google Research paper. It supports Ef… | 10 | 1441 | maintenance |
| ZPdesu/Barbershop Official PyTorch implementation of Barbershop (SIGGRAPH Asia 2021), a GAN-inversion-based method for compositing images using segmentation … | 32 | 1439 | maintenance |
| wenhaochai/StableVideo StableVideo is the official ICCV 2023 implementation of text-driven, consistency-aware diffusion-based video editing built on ControlNet an… | 31 | 1439 | maintenance |
| VDIGPKU/M2Det M2Det is a PyTorch implementation of a single-shot object detector based on a Multi-Level Feature Pyramid Network (MLFPN), published at AAA… | 32 | 1438 | maintenance |
| mchong6/JoJoGAN JoJoGAN is the official PyTorch implementation of a one-shot face stylization method that finetunes a pretrained StyleGAN using GAN inversi… | 32 | 1438 | maintenance |
| xuebinqin/BASNet BASNet is the official PyTorch implementation of the CVPR 2019 paper 'BASNet: Boundary-Aware Salient Object Detection', a deep learning mod… | 32 | 1436 | maintenance |
| yangyanli/PointCNN PointCNN is a deep learning framework for feature learning from 3D point clouds, applying convolution on X-transformed points to handle the… | 55 | 1434 | maintenance |
| hollance/CoreMLHelpers A collection of Swift types and helper functions that simplify working with Apple's Core ML framework, such as image-to-CVPixelBuffer conve… | 32 | 1433 | maintenance |
| MSPaintIDE/MSPaintIDE MS Paint IDE is a novelty-but-functional IDE that uses a custom OCR engine to read code written in MS Paint images, then highlights, parses… | 23 | 1433 | maintenance |
| andyzeng/tsdf-fusion-python A lightweight Python script that fuses multiple registered RGB-D images into a projective TSDF voxel volume, from which high-quality 3D sur… | 32 | 1430 | maintenance |
| vlfeat/matconvnet MatConvNet is a MATLAB toolbox implementing convolutional neural networks (CNNs) for computer vision applications. It supports training and… | 32 | 1430 | maintenance |
| qfgaohao/pytorch-ssd A PyTorch implementation of the SSD (Single Shot MultiBox Detector) object detection algorithm with MobileNetV1, MobileNetV2, and VGG backb… | 32 | 1429 | maintenance |
| nianticlabs/simplerecon SimpleRecon is the reference PyTorch implementation of an ECCV 2022 paper for multi-view stereo depth estimation and 3D reconstruction that… | 41 | 1428 | maintenance |
| KaiyangZhou/Dassl.pytorch Dassl is a PyTorch toolbox for research in domain adaptation, domain generalization, and semi-supervised learning. It provides modular comp… | 32 | 1428 | maintenance |
| theAIGuysCode/yolov4-deepsort A Python implementation of multi-object tracking that combines YOLOv4 object detection with the Deep SORT tracking algorithm using TensorFl… | 32 | 1428 | maintenance |
| open-mmlab/mmhuman3d MMHuman3D is an open-source PyTorch-based toolbox and benchmark for 3D human parametric models (e.g., SMPL, SMPL-X) in computer vision and … | 23 | 1426 | maintenance |
| torchgan/torchgan TorchGAN is a PyTorch-based research framework for designing and training Generative Adversarial Networks. It provides modular building blo… | 23 | 1425 | maintenance |
| sfzhang15/RefineDet RefineDet is a C++/Caffe implementation of the CVPR 2018 single-shot object detection model that refines anchors to combine one-stage speed… | 32 | 1424 | maintenance |
| rmokady/CLIP_prefix_caption Official implementation of ClipCap, a CLIP-based image captioning model that maps CLIP image encodings to a GPT-2 prefix to generate captio… | 32 | 1423 | maintenance |
| GOATmessi8/RFBNet PyTorch implementation of RFBNet, a Receptive Field Block Net object detector presented at ECCV 2018. It enhances SSD-style detectors with … | 32 | 1419 | maintenance |
| BangguWu/ECANet Official PyTorch implementation of ECA-Net, an efficient channel attention module for deep convolutional neural networks published at CVPR … | 32 | 1416 | maintenance |
| cdpierse/transformers-interpret A Python library providing model explainability for Hugging Face Transformers models, built on Captum. It explains text classification, que… | 23 | 1416 | maintenance |
| bermanmaxim/LovaszSoftmax Standalone PyTorch and TensorFlow implementations of the Lovász-Softmax and Lovász Hinge loss functions for optimizing the Jaccard index (I… | 32 | 1410 | maintenance |
| chrischoy/3D-R2N2 3D-R2N2 is a PyTorch-based implementation of a recurrent neural network that reconstructs voxelized 3D models of objects from one or multip… | 32 | 1410 | maintenance |
| yu-changqian/TorchSeg A fast, modular PyTorch reference implementation for training and evaluating semantic segmentation models such as FCN, DFN, BiSeNet, PSPNet… | 23 | 1410 | maintenance |
| VerticalResearchGroup/miaow MIAOW is an open source GPU implementation of the AMD Southern Islands ISA written in Verilog. It was developed as a research project at th… | 39 | 1406 | maintenance |
| tonylins/pytorch-mobilenet-v2 A PyTorch implementation of the MobileNetV2 architecture with automatic download of ImageNet pretrained weights. It reproduces the official… | 32 | 1406 | maintenance |
| moeiscool/Shinobi Shinobi CE is a free, open-source CCTV/NVR platform written in Node.js for recording and managing IP security cameras. It supports RTSP/ONV… | 23 | 1406 | maintenance |
| OpenNI/OpenNI OpenNI is a C++ framework and SDK providing a standard interface for depth sensors and 3D vision devices like PrimeSense and Kinect cameras… | 32 | 1394 | maintenance |
| swz30/MPRNet MPRNet is a PyTorch implementation of a multi-stage progressive image restoration network published at CVPR 2021. It provides pretrained mo… | 32 | 1393 | maintenance |
| yinboc/liif LIIF is the official PyTorch implementation of the CVPR 2021 paper 'Learning Continuous Image Representation with Local Implicit Image Func… | 32 | 1388 | maintenance |
| RajSolai/TextSnatcher TextSnatcher is a Linux desktop application for quickly copying text from images using OCR. Built with Vala and GTK3 on top of Tesseract, i… | 23 | 1387 | maintenance |
| doonny/PipeCNN PipeCNN is an OpenCL-based FPGA accelerator for large-scale convolutional neural network inference, written in C with pipelined kernels. It… | 32 | 1386 | maintenance |
| zhubenfu/License-Plate-Detect-Recognition-via-Deep-Neural-Networks-accuracy-up-to-99.9 A C++ application that detects and recognizes Chinese license plates in real time using deep neural networks, claiming up to 99.8% accuracy… | 32 | 1384 | maintenance |
| myhub/tr An offline Chinese text detection and recognition OCR SDK with C++ core code and Python bindings, supporting models like CRNN, CTPN, and Pi… | 50 | 1381 | maintenance |
| PSPNet Reference implementation of the Pyramid Scene Parsing Network (PSPNet), a CVPR 2017 semantic segmentation model that won the ImageNet Scene… | 32 | 1378 | maintenance |
| takuya-takeuchi/FaceRecognitionDotNet A C# port of the popular face_recognition library providing a simple facial recognition API for .NET. It supports face detection, recogniti… | 23 | 1376 | maintenance |
| amazon-science/patchcore-inspection Official implementation of PatchCore, a deep-learning method for industrial image anomaly detection and localization from Roth et al. (2021… | 32 | 1373 | maintenance |
| tleyden/open-ocr OpenOCR is a self-hosted OCR-as-a-Service REST API built in Go on top of Tesseract, containerized with Docker. It uses RabbitMQ for scalabl… | 32 | 1373 | maintenance |
| tensorboy/pytorch_Realtime_Multi-Person_Pose_Estimation A PyTorch implementation of the CVPR'17 Realtime Multi-Person 2D Pose Estimation (OpenPose/rtpose) model. It provides pretrained weights, d… | 32 | 1371 | maintenance |
| msracver/Deep-Image-Analogy Official C++/CUDA implementation of the SIGGRAPH 2017 'Visual Attribute Transfer through Deep Image Analogy' technique from Microsoft Resea… | 32 | 1370 | maintenance |
| nv-tlabs/lift-splat-shoot PyTorch implementation of Lift-Splat-Shoot (ECCV 2020), an end-to-end model that converts images from arbitrary multi-camera rigs into a bi… | 32 | 1369 | maintenance |
| open-mmlab/mmfashion MMFashion is an open-source PyTorch-based toolbox for visual fashion analysis from the OpenMMLab project. It provides modular implementatio… | 32 | 1369 | maintenance |
| akamaster/pytorch_resnet_cifar10 A PyTorch implementation of ResNet architectures (ResNet20 through ResNet1202) for CIFAR10/CIFAR100 that faithfully matches the original pa… | 32 | 1368 | maintenance |
| Cysu/open-reid Open-ReID is a lightweight Python library for person re-identification built on PyTorch, providing a uniform dataset interface, models, and… | 32 | 1368 | maintenance |
| ethz-asl/okvis OKVIS is a C++ implementation of keyframe-based visual-inertial SLAM/odometry using nonlinear optimization, from ETH Zurich research. It pr… | 32 | 1366 | maintenance |
| wzzheng/TPVFormer TPVFormer is a CVPR 2023 research implementation of a tri-perspective view transformer for vision-based 3D semantic occupancy prediction, s… | 32 | 1364 | maintenance |
| ali-vilab/ACE_plus ACE++ is a Python library and model release from Alibaba's Tongyi Lab for instruction-based image creation and editing via context-aware co… | 28 | 1364 | maintenance |
| HKUST-Aerial-Robotics/VINS-Mobile VINS-Mobile is a real-time monocular visual-inertial state estimator that runs on iOS devices, providing high-accuracy visual-inertial odom… | 32 | 1363 | maintenance |
| taoxugit/AttnGAN A PyTorch implementation of AttnGAN, a fine-grained text-to-image generative adversarial network with multi-stage attention mechanisms from… | 32 | 1363 | maintenance |
| Image-Py/imagepy ImagePy is an open-source, ImageJ-like interactive image processing application and plugin framework written in Python, built on wxPython, … | 23 | 1363 | maintenance |
| sail-sg/poolformer PoolFormer is a PyTorch implementation of the CVPR 2022 paper 'MetaFormer Is Actually What You Need for Vision', which replaces attention i… | 23 | 1362 | maintenance |
| mchehab/zbar ZBar is an open-source C library and suite of tools for reading barcodes from video streams, image files, and sensors, supporting formats l… | 55 | 1361 | maintenance |
| MoyGcc/vid2avatar Vid2Avatar is the official PyTorch implementation of a CVPR 2023 method that reconstructs detailed 3D human avatars from monocular in-the-w… | 56 | 1360 | maintenance |
| atulapra/Emotion-detection A Python application that performs real-time facial emotion detection from a webcam feed using a CNN trained on the FER-2013 dataset. It us… | 32 | 1359 | maintenance |
| mayuelala/FollowYourPose Official PyTorch implementation of Follow-Your-Pose (AAAI 2024), a pose-guided text-to-video generation model that tunes a text-to-image mo… | 21 | 1358 | maintenance |
| atduskgreg/opencv-processing A Processing library wrapping OpenCV's official Java bindings to provide beginner-friendly computer vision functions within the Processing … | 23 | 1356 | maintenance |
| dlunion/DBFace DBFace is a real-time, single-stage face detection model implemented in Python, offering small model sizes with high accuracy on the WiderF… | 32 | 1355 | maintenance |
| orobix/retina-unet A Python implementation of a U-Net convolutional neural network for segmenting blood vessels in retina fundus images. It performs binary pi… | 32 | 1354 | maintenance |
| ShuLiu1993/PANet A PyTorch re-implementation of PANet (Path Aggregation Network), the CVPR 2018 paper that won 1st place in the COCO 2017 Instance Segmentat… | 32 | 1347 | maintenance |
| microsoft/X-Decoder Official PyTorch implementation of X-Decoder, a generalized decoding model from CVPR 2023 that unifies pixel-level segmentation, image-leve… | 22 | 1345 | maintenance |
| Star-Clouds/CenterFace CenterFace is a lightweight (7.3MB) anchor-free face detection and alignment model that detects faces and predicts facial landmarks using a… | 32 | 1344 | maintenance |
| dorarad/gansformer GANformer is a research implementation of a generative adversarial transformer that uses a bipartite attention structure for efficient high… | 23 | 1344 | maintenance |
| panrafal/depthy Depthy is a web application that extracts depth maps from Google Camera Lens Blur photos and displays them with a 3D parallax effect. It ca… | 44 | 1343 | maintenance |
| PeizeSun/SparseR-CNN Sparse R-CNN is a PyTorch implementation (built on Detectron2) of the CVPR 2021 / PAMI 2023 paper 'End-to-End Object Detection with Learnab… | 23 | 1343 | maintenance |
| HobbitLong/CMC Official PyTorch implementation of Contrastive Multiview Coding (CMC), a self-supervised visual representation learning method that contras… | 32 | 1340 | maintenance |
| kuaikuaikim/dface DFace is an open-source Python library implementing face detection and recognition with PyTorch, based on the MTCNN cascaded convolutional … | 23 | 1339 | maintenance |
| timojl/clipseg CLIPSeg is a Python implementation of the CVPR 2022 paper 'Image Segmentation Using Text and Image Prompts', enabling zero-shot segmentatio… | 32 | 1338 | maintenance |