function: computer-vision
1555 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| QingyongHu/RandLA-Net Official TensorFlow implementation of RandLA-Net, a neural architecture for efficient semantic segmentation of large-scale 3D point clouds,… | 32 | 1560 | maintenance |
| V2AI/Det3D Det3D is a PyTorch-based toolbox for 3D object detection from point clouds, offering implementations of models like PointPillars, SECOND, a… | 32 | 1560 | maintenance |
| IDEA-Research/MaskDINO Official PyTorch implementation of Mask DINO, a unified transformer-based framework for object detection and segmentation, built on detectr… | 23 | 1557 | maintenance |
| microsoft/SoM A research toolbox from Microsoft implementing Set-of-Mark (SoM) visual prompting, which overlays numbered spatial marks on images to impro… | 18 | 1557 | maintenance |
| vt-vl-lab/FGVC FGVC is a PyTorch implementation of the ECCV 2020 paper 'Flow-edge Guided Video Completion'. It completes missing regions in videos by extr… | 32 | 1551 | maintenance |
| DevashishPrasad/CascadeTabNet CascadeTabNet is a PyTorch/mmdetection implementation of a CVPR 2020 paper for end-to-end table detection and structure recognition from im… | 32 | 1549 | maintenance |
| JiaRenChang/PSMNet PSMNet is the official PyTorch implementation of the CVPR 2018 paper 'Pyramid Stereo Matching Network' for stereo depth estimation. It uses… | 32 | 1548 | maintenance |
| Javacr/PyQt5-YOLOv5 A desktop GUI application built with PyQt5 that wraps YOLOv5 (v6.1) object detection models. It supports running detection on images, video… | 32 | 1547 | maintenance |
| YoYo000/MVSNet MVSNet is a deep learning architecture for depth map inference from unstructured multi-view images, and R-MVSNet is its recurrent extension… | 32 | 1546 | maintenance |
| allgood/OpenNoteScanner OpenNoteScanner is an open-source Android app for scanning handwritten notes and printed documents using a mobile device camera. It uses Op… | 31 | 1539 | maintenance |
| mzucker/page_dewarp A Python command-line tool that dewarps photos of curled or warped book pages using a cubic sheet optimization model, flattening them into … | 32 | 1526 | maintenance |
| yu4u/age-gender-estimation A Keras/TensorFlow implementation of a convolutional neural network that estimates age and gender from face images, trained on the IMDB-WIK… | 23 | 1520 | maintenance |
| zhan-xu/RigNet RigNet is a PyTorch implementation of the SIGGRAPH 2020 paper on neural rigging for articulated 3D characters. It takes a character mesh as… | 32 | 1519 | maintenance |
| marsauto/europilot Europilot is a Python toolkit that bridges Euro Truck Simulator 2 with deep-learning frameworks, enabling end-to-end self-driving research … | 32 | 1511 | maintenance |
| compphoto/BoostingMonocularDepth A Python research implementation for boosting monocular depth estimation to high resolution using a double-estimation merging operator, sup… | 32 | 1507 | maintenance |
| tinyvision/SOLIDER SOLIDER is a semantic-controllable self-supervised learning framework that learns general human representations from massive unlabeled huma… | 32 | 1504 | maintenance |
| AITTSMD/MTCNN-Tensorflow A TensorFlow reproduction of MTCNN (Multi-task Cascaded Convolutional Networks) for joint face detection and facial landmark alignment. It … | 32 | 1502 | maintenance |
| joschuck/matrix-webcam A Python CLI tool that renders your webcam video feed as a Matrix-style falling-character animation in the terminal. The output can be pipe… | 32 | 1490 | maintenance |
| ageitgey/show-facebook-computer-vision-tags A simple Chrome extension that overlays the automated computer vision tags Facebook generates for images directly on your Facebook timeline… | 32 | 1488 | maintenance |
| chengdazhi/Deformable-Convolution-V2-PyTorch A PyTorch implementation of Deformable Convolution V2 (DCNv2) custom CUDA operators, ported from the original MXNet implementation. It prov… | 32 | 1484 | maintenance |
| speedinghzl/CCNet Official PyTorch implementation of CCNet, a Criss-Cross Attention network for semantic segmentation published at ICCV 2019 and TPAMI 2020. … | 32 | 1484 | maintenance |
| cvg/pixel-perfect-sfm pixsfm is a Python package with a C++ core that improves Structure-from-Motion and visual localization accuracy by refining keypoints, came… | 23 | 1484 | maintenance |
| AIRLegend/aitrack AITrack is a free 6DoF head tracking application that uses a webcam and neural networks to estimate head position and rotation. It streams … | 23 | 1482 | maintenance |
| chainer/chainercv ChainerCV is a Python library providing tools to train and run neural networks for computer vision tasks on top of the Chainer framework. I… | 10 | 1481 | maintenance |
| Sharpiless/Yolov5-deepsort-inference A Python library combining YOLOv5 object detection with DeepSort multi-object tracking to detect, track, and count vehicles and pedestrians… | 66 | 1479 | maintenance |
| faustomorales/keras-ocr A Python library packaging the CRAFT text detector and a Keras CRNN text recognition model into a high-level OCR pipeline. It supports pret… | 42 | 1473 | maintenance |
| shouzhong/Scanner An Android scanning library that recognizes QR codes, barcodes, ID cards, bank cards, license plates, text, driving licenses, and NSFW imag… | 23 | 1472 | maintenance |
| peteanderson80/bottom-up-attention A bottom-up attention model based on Faster R-CNN with ResNet-101 trained on Visual Genome, producing features for salient image regions. T… | 32 | 1469 | maintenance |
| CSAILVision/LabelMeAnnotationTool The source code for LabelMe, a web-based image annotation tool written in JavaScript that runs on your own Apache server. It lets users dra… | 32 | 1468 | maintenance |
| HRNet/HigherHRNet-Human-Pose-Estimation Official PyTorch implementation of HigherHRNet, a CVPR 2020 bottom-up multi-person human pose estimation model using scale-aware high-resol… | 32 | 1463 | maintenance |
| una-dinosauria/3d-pose-baseline A TensorFlow implementation of a simple yet effective baseline for 3D human pose estimation from 2D keypoints, published at ICCV 2017. It i… | 32 | 1459 | maintenance |
| wvangansbeke/Unsupervised-Classification PyTorch implementation of SCAN (ECCV 2020), a two-step method for unsupervised image classification that combines self-supervised represent… | 32 | 1455 | maintenance |
| abreheret/PixelAnnotationTool A C++ desktop application for quickly annotating images with pixel-wise labels using OpenCV's watershed algorithm. Users draw markers with … | 23 | 1455 | maintenance |
| xuannianz/EfficientDet A Keras/TensorFlow implementation of the EfficientDet object detection model with pretrained COCO and ImageNet weights. It supports trainin… | 32 | 1454 | maintenance |
| dji-sdk/Tello-Python A collection of Python sample modules for controlling and interacting with the Ryze Tello drone, including command scripting, video streami… | 32 | 1451 | maintenance |
| toandaominh1997/EfficientDet.Pytorch A PyTorch implementation of EfficientDet, a scalable and efficient object detection model from a 2019 Google Research paper. It supports Ef… | 10 | 1441 | maintenance |
| VDIGPKU/M2Det M2Det is a PyTorch implementation of a single-shot object detector based on a Multi-Level Feature Pyramid Network (MLFPN), published at AAA… | 32 | 1438 | maintenance |
| yangyanli/PointCNN PointCNN is a deep learning framework for feature learning from 3D point clouds, applying convolution on X-transformed points to handle the… | 55 | 1434 | maintenance |
| andyzeng/tsdf-fusion-python A lightweight Python script that fuses multiple registered RGB-D images into a projective TSDF voxel volume, from which high-quality 3D sur… | 32 | 1430 | maintenance |
| qfgaohao/pytorch-ssd A PyTorch implementation of the SSD (Single Shot MultiBox Detector) object detection algorithm with MobileNetV1, MobileNetV2, and VGG backb… | 32 | 1429 | maintenance |
| nianticlabs/simplerecon SimpleRecon is the reference PyTorch implementation of an ECCV 2022 paper for multi-view stereo depth estimation and 3D reconstruction that… | 41 | 1428 | maintenance |
| theAIGuysCode/yolov4-deepsort A Python implementation of multi-object tracking that combines YOLOv4 object detection with the Deep SORT tracking algorithm using TensorFl… | 32 | 1428 | maintenance |
| open-mmlab/mmhuman3d MMHuman3D is an open-source PyTorch-based toolbox and benchmark for 3D human parametric models (e.g., SMPL, SMPL-X) in computer vision and … | 23 | 1426 | maintenance |
| sfzhang15/RefineDet RefineDet is a C++/Caffe implementation of the CVPR 2018 single-shot object detection model that refines anchors to combine one-stage speed… | 32 | 1424 | maintenance |
| GOATmessi8/RFBNet PyTorch implementation of RFBNet, a Receptive Field Block Net object detector presented at ECCV 2018. It enhances SSD-style detectors with … | 32 | 1419 | maintenance |
| chrischoy/3D-R2N2 3D-R2N2 is a PyTorch-based implementation of a recurrent neural network that reconstructs voxelized 3D models of objects from one or multip… | 32 | 1410 | maintenance |
| imaginary-cloud/CameraManager CameraManager is a Swift library that wraps AVFoundation to simplify building custom camera views in iOS apps. It handles camera configurat… | 32 | 1396 | maintenance |
| OpenNI/OpenNI OpenNI is a C++ framework and SDK providing a standard interface for depth sensors and 3D vision devices like PrimeSense and Kinect cameras… | 32 | 1394 | maintenance |
| zhubenfu/License-Plate-Detect-Recognition-via-Deep-Neural-Networks-accuracy-up-to-99.9 A C++ application that detects and recognizes Chinese license plates in real time using deep neural networks, claiming up to 99.8% accuracy… | 32 | 1384 | maintenance |
| myhub/tr An offline Chinese text detection and recognition OCR SDK with C++ core code and Python bindings, supporting models like CRNN, CTPN, and Pi… | 50 | 1381 | maintenance |
| PSPNet Reference implementation of the Pyramid Scene Parsing Network (PSPNet), a CVPR 2017 semantic segmentation model that won the ImageNet Scene… | 32 | 1378 | maintenance |
| takuya-takeuchi/FaceRecognitionDotNet A C# port of the popular face_recognition library providing a simple facial recognition API for .NET. It supports face detection, recogniti… | 23 | 1376 | maintenance |
| amazon-science/patchcore-inspection Official implementation of PatchCore, a deep-learning method for industrial image anomaly detection and localization from Roth et al. (2021… | 32 | 1373 | maintenance |
| tensorboy/pytorch_Realtime_Multi-Person_Pose_Estimation A PyTorch implementation of the CVPR'17 Realtime Multi-Person 2D Pose Estimation (OpenPose/rtpose) model. It provides pretrained weights, d… | 32 | 1371 | maintenance |
| msracver/Deep-Image-Analogy Official C++/CUDA implementation of the SIGGRAPH 2017 'Visual Attribute Transfer through Deep Image Analogy' technique from Microsoft Resea… | 32 | 1370 | maintenance |
| nv-tlabs/lift-splat-shoot PyTorch implementation of Lift-Splat-Shoot (ECCV 2020), an end-to-end model that converts images from arbitrary multi-camera rigs into a bi… | 32 | 1369 | maintenance |
| ethz-asl/okvis OKVIS is a C++ implementation of keyframe-based visual-inertial SLAM/odometry using nonlinear optimization, from ETH Zurich research. It pr… | 32 | 1366 | maintenance |
| wzzheng/TPVFormer TPVFormer is a CVPR 2023 research implementation of a tri-perspective view transformer for vision-based 3D semantic occupancy prediction, s… | 32 | 1364 | maintenance |
| HKUST-Aerial-Robotics/VINS-Mobile VINS-Mobile is a real-time monocular visual-inertial state estimator that runs on iOS devices, providing high-accuracy visual-inertial odom… | 32 | 1363 | maintenance |
| mchehab/zbar ZBar is an open-source C library and suite of tools for reading barcodes from video streams, image files, and sensors, supporting formats l… | 55 | 1361 | maintenance |
| MoyGcc/vid2avatar Vid2Avatar is the official PyTorch implementation of a CVPR 2023 method that reconstructs detailed 3D human avatars from monocular in-the-w… | 56 | 1360 | maintenance |
| atulapra/Emotion-detection A Python application that performs real-time facial emotion detection from a webcam feed using a CNN trained on the FER-2013 dataset. It us… | 32 | 1359 | maintenance |
| atduskgreg/opencv-processing A Processing library wrapping OpenCV's official Java bindings to provide beginner-friendly computer vision functions within the Processing … | 23 | 1356 | maintenance |
| dlunion/DBFace DBFace is a real-time, single-stage face detection model implemented in Python, offering small model sizes with high accuracy on the WiderF… | 32 | 1355 | maintenance |
| bumble-tech/private-detector Bumble's Private Detector is a pretrained TensorFlow image classifier based on EfficientNet-v2 that detects lewd images. The repo provides … | 23 | 1352 | maintenance |
| ShuLiu1993/PANet A PyTorch re-implementation of PANet (Path Aggregation Network), the CVPR 2018 paper that won 1st place in the COCO 2017 Instance Segmentat… | 32 | 1347 | maintenance |
| microsoft/X-Decoder Official PyTorch implementation of X-Decoder, a generalized decoding model from CVPR 2023 that unifies pixel-level segmentation, image-leve… | 22 | 1345 | maintenance |
| Star-Clouds/CenterFace CenterFace is a lightweight (7.3MB) anchor-free face detection and alignment model that detects faces and predicts facial landmarks using a… | 32 | 1344 | maintenance |
| panrafal/depthy Depthy is a web application that extracts depth maps from Google Camera Lens Blur photos and displays them with a 3D parallax effect. It ca… | 44 | 1343 | maintenance |
| PeizeSun/SparseR-CNN Sparse R-CNN is a PyTorch implementation (built on Detectron2) of the CVPR 2021 / PAMI 2023 paper 'End-to-End Object Detection with Learnab… | 23 | 1343 | maintenance |
| kuaikuaikim/dface DFace is an open-source Python library implementing face detection and recognition with PyTorch, based on the MTCNN cascaded convolutional … | 23 | 1339 | maintenance |
| timojl/clipseg CLIPSeg is a Python implementation of the CVPR 2022 paper 'Image Segmentation Using Text and Image Prompts', enabling zero-shot segmentatio… | 32 | 1338 | maintenance |
| yannickl/QRCodeReader.swift QRCodeReader.swift is a Swift library for iOS that provides a simple QR code and machine-readable code scanner built on Apple's AVFoundatio… | 32 | 1337 | maintenance |
| torrvision/crfasrnn Reference implementation of CRF-RNN, an ICCV 2015 semantic image segmentation method that integrates conditional random fields into a neura… | 23 | 1335 | maintenance |
| ahmetozlu/tensorflow_object_counting_api An open-source framework built on TensorFlow and Keras that simplifies developing object counting systems. It supports cumulative counting,… | 23 | 1333 | maintenance |
| maudzung/Complex-YOLOv4-Pytorch A PyTorch implementation of Complex-YOLO, a YOLOv4-based model for real-time 3D object detection on LiDAR point clouds. It supports distrib… | 32 | 1327 | maintenance |
| go-opencv/go-opencv Go bindings for the OpenCV computer vision library, exposing the OpenCV 1.x C API via CGO and an experimental OpenCV 2.x C++ API (the gocv … | 32 | 1324 | maintenance |
| WisconsinAIVision/yolact_edge YolactEdge is a PyTorch implementation of a real-time instance segmentation model optimized for edge devices like the NVIDIA Jetson AGX Xav… | 32 | 1322 | maintenance |
| msracver/Deep-Feature-Flow Official MXNet implementation of Deep Feature Flow (CVPR 2017), an end-to-end framework for video recognition such as object detection and … | 32 | 1315 | maintenance |
| PRBonn/depth_clustering A fast and robust C++ library for segmenting point clouds from Velodyne LiDAR sensors (16, 32, and 64 beam) into objects using depth cluste… | 23 | 1312 | maintenance |
| pqpo/SmartCamera SmartCamera is an Android camera extension library providing a highly customizable real-time scanning module that detects whether an object… | 23 | 1305 | maintenance |
| ToanTech/py-apple-quadruped-robot Py-Apple Dog (Pineapple Dog) is a low-cost, open-source hardware and software project for building a DIY quadruped robot dog. It bundles fo… | 32 | 1302 | maintenance |
| hengli/camodocal CamOdoCal is a C++ library for automatic intrinsic and extrinsic calibration of a camera rig with multiple generic cameras and odometry. It… | 23 | 1302 | maintenance |
| foolwood/DaSiamRPN PyTorch implementation of DaSiamRPN, an ECCV 2018 distractor-aware Siamese network for visual object tracking, winner of the VOT-18 real-ti… | 32 | 1299 | maintenance |
| NVlabs/DG-Net DG-Net is a PyTorch implementation of the CVPR 2019 (Oral) paper 'Joint Discriminative and Generative Learning for Person Re-identification… | 32 | 1298 | maintenance |
| zeusees/License-Plate-Detector A YOLOv5-based license plate detection model trained on the CCPD dataset and proprietary data, supporting many Chinese plate types. It prov… | 32 | 1288 | maintenance |
| apple/ml-neuman Official reference implementation of NeuMan (ECCV 2022), which reconstructs an animatable human and the background scene from a single vide… | 32 | 1286 | maintenance |
| duzexu/ARuler An iOS augmented reality app that measures distances using Apple's ARKit and SceneKit. It uses plane detection and feature point hit-testin… | 63 | 1280 | maintenance |
| MasterBin-IIAU/UNINEXT UNINEXT is the official PyTorch implementation of the CVPR 2023 paper 'Universal Instance Perception as Object Discovery and Retrieval'. It… | 30 | 1278 | maintenance |
| chonyy/AI-basketball-analysis An AI-powered web app and API that analyzes basketball shots and shooting form from uploaded videos using object detection and OpenPose pos… | 32 | 1277 | maintenance |
| SullyChen/Autopilot-TensorFlow A TensorFlow implementation of Nvidia's end-to-end self-driving steering angle prediction paper (arXiv 1604.07316) with some modifications.… | 32 | 1276 | maintenance |
| j96w/DenseFusion DenseFusion is the official PyTorch implementation of the paper '6D Object Pose Estimation by Iterative Dense Fusion', which estimates the … | 32 | 1276 | maintenance |
| YavorGIvanov/sam.cpp A pure C/C++ implementation of Meta's Segment Anything Model (SAM) for image segmentation inference, built on the ggml tensor library. It c… | 28 | 1275 | maintenance |
| TRI-ML/packnet-sfm Official PyTorch implementation of PackNet and related self-supervised monocular depth estimation methods from Toyota Research Institute's … | 23 | 1274 | maintenance |
| phonegap/phonegap-plugin-barcodescanner A Cordova/PhoneGap plugin providing cross-platform barcode scanning for hybrid mobile apps. It exposes a JavaScript API (cordova.plugins.ba… | 10 | 1269 | maintenance |
| tensorlayer/HyperPose HyperPose is a library for building high-performance custom human pose estimation applications. It combines a C++ inference engine with Ten… | 23 | 1264 | maintenance |
| rohitrango/automatic-watermark-detection A Python implementation of the CVPR 2017 paper 'On The Effectiveness Of Visible Watermarks' that detects and removes visible watermarks fro… | 32 | 1263 | maintenance |
| kennymckormick/pyskl PYSKL is a PyTorch-based toolbox for skeleton-based human action recognition, built on MMAction2. It provides official implementations of P… | 53 | 1261 | maintenance |
| openpifpaf/openpifpaf OpenPifPaf is a PyTorch library implementing Composite Fields for semantic keypoint detection and spatio-temporal association, primarily fo… | 23 | 1261 | maintenance |
| BrandonJoffe/home_surveillance A self-hosted home surveillance application that processes streams from multiple IP cameras, performs motion detection and facial recogniti… | 32 | 1259 | maintenance |