domain: computer-vision
2316 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| NVlabs/PWC-Net Official NVIDIA implementation of PWC-Net, a CNN for optical flow estimation using pyramid, warping, and cost volume, released with CVPR 20… | 32 | 1736 | maintenance |
| zijundeng/pytorch-semantic-segmentation A PyTorch library providing implementations of popular semantic segmentation models such as FCN, U-Net, SegNet, PSPNet, GCN, and DUC/HDC. I… | 32 | 1735 | maintenance |
| experiencor/keras-yolo2 A Keras/TensorFlow implementation of the YOLOv2 real-time object detection model with support for training on custom datasets. It offers mu… | 23 | 1733 | maintenance |
| ranahanocka/MeshCNN MeshCNN is a PyTorch library implementing a convolutional neural network that operates directly on 3D triangular meshes, with mesh-specific… | 32 | 1731 | maintenance |
| NVIDIA/Cosmos-Tokenizer NVIDIA Cosmos Tokenizer is a suite of neural tokenizers for images and videos that convert visual data into continuous latents or discrete … | 10 | 1731 | maintenance |
| natethegreate/hent-AI A Python application that automatically detects censor bars and mosaic blurs in illustrated adult content using deep learning (Mask R-CNN) … | 23 | 1730 | maintenance |
| cszn/DnCNN DnCNN is the official implementation of the TIP 2017 paper 'Beyond a Gaussian Denoiser: Residual Learning of Deep CNN for Image Denoising',… | 32 | 1727 | maintenance |
| bubbliiiing/unet-pytorch A PyTorch implementation of the U-Net semantic segmentation model with training, prediction, and mIoU evaluation scripts. It supports multi… | 23 | 1725 | maintenance |
| Lightning-Universe/lightning-flash Lightning Flash is a high-level PyTorch library built on PyTorch Lightning that provides ready-made 'recipes' for over 15 AI tasks across 7… | 10 | 1722 | maintenance |
| svip-lab/impersonator A PyTorch implementation of Liquid Gating GAN (ICCV 2019) that performs human motion imitation, appearance transfer, and novel view synthes… | 32 | 1717 | maintenance |
| dunbar12138/pix2pix3D pix2pix3D is the official PyTorch implementation of a CVPR 2023 paper on 3D-aware conditional image synthesis. It generates 3D objects (neu… | 31 | 1716 | maintenance |
| cvlab-columbia/viper ViperGPT is a research codebase that composes vision-and-language models with code generated by large language models (GPT-3.5/GPT-4) to pe… | 30 | 1716 | maintenance |
| sipeed/MaixPy-v1 MaixPy-v1 is a MicroPython port for the Kendryte K210 RISC-V AI chip, letting users program Sipeed Maix boards in Python. It provides APIs … | 23 | 1712 | maintenance |
| aiff22/DPED DPED is a Python/TensorFlow implementation of a deep convolutional network approach that translates ordinary smartphone photos into DSLR-qu… | 49 | 1710 | maintenance |
| Fyusion/LLFF LLFF (Local Light Field Fusion) is a TensorFlow implementation of the SIGGRAPH 2019 paper for novel view synthesis from sparse input images… | 32 | 1703 | maintenance |
| microsoft/i-Code Microsoft's i-Code is a collection of research models and frameworks for integrative, composable multimodal AI spanning vision, language, a… | 32 | 1703 | maintenance |
| hanleyweng/CoreML-in-ARKit A Swift demo/template iOS app that runs CoreML object detection (Inception V3) on live ARKit camera frames and renders 3D text labels above… | 32 | 1697 | maintenance |
| VITA-Group/TransGAN Official PyTorch implementation of TransGAN, a NeurIPS 2021 paper that builds a GAN whose generator and discriminator are both pure transfo… | 32 | 1695 | maintenance |
| LoSealL/VideoSuperResolution A Python library (pip-installable as VSR) collecting reimplementation of state-of-the-art single-image and video super-resolution neural ne… | 23 | 1687 | maintenance |
| koide3/fast_gicp A C++ library of fast GICP-based point cloud registration algorithms, including multi-threaded GICP, voxelized GICP (VGICP), and CUDA-accel… | 40 | 1686 | maintenance |
| adobe/antialiased-cnns A PyTorch library providing antialiased CNN models and a BlurPool layer from the ICML 2019 paper 'Making Convolutional Networks Shift-Invar… | 23 | 1683 | maintenance |
| open-mmlab/mmrazor MMRazor is OpenMMLab's model compression toolbox and benchmark built on PyTorch. It provides implementations of neural architecture search,… | 23 | 1682 | maintenance |
| CuriousAI/mean-teacher Reference implementations (TensorFlow and PyTorch) of the Mean Teacher semi-supervised learning method from the NIPS 2017 paper by Tarvaine… | 32 | 1678 | maintenance |
| Lam1360/YOLOv3-model-pruning A PyTorch implementation of YOLOv3 channel pruning (network slimming) applied to hand detection on the Oxford Hand dataset. It provides spa… | 32 | 1676 | maintenance |
| argusswift/YOLOv4-pytorch A PyTorch re-implementation of YOLOv4 object detection with variants including attentive YOLOv4 (SEnet, CBAM, CoordAttention) and MobileNet… | 23 | 1676 | maintenance |
| YuliangXiu/ICON ICON is a PyTorch research implementation of a CVPR 2022 method that reconstructs detailed, animatable 3D clothed human avatars from 2D ima… | 23 | 1675 | maintenance |
| jamriska/ebsynth Ebsynth is a fast example-based image synthesis tool that performs style transfer, guided texture synthesis, inpainting, and super-resoluti… | 32 | 1674 | maintenance |
| charlesq34/frustum-pointnets Official TensorFlow code release for the CVPR 2018 paper 'Frustum PointNets for 3D Object Detection from RGB-D Data' by Stanford and Nuro r… | 32 | 1668 | maintenance |
| akanazawa/hmr HMR (Human Mesh Recovery) is a TensorFlow implementation of the CVPR 2018 paper 'End-to-end Recovery of Human Shape and Pose', which regres… | 32 | 1666 | maintenance |
| natanielruiz/deep-head-pose Hopenet is a PyTorch deep learning model for fine-grained head pose estimation from images and video, without requiring facial keypoints. I… | 32 | 1666 | maintenance |
| autonomousvision/occupancy_networks Official PyTorch implementation of the CVPR 2019 paper 'Occupancy Networks: Learning 3D Reconstruction in Function Space'. It learns contin… | 32 | 1663 | maintenance |
| rwightman/efficientdet-pytorch A PyTorch implementation of EfficientDet object detection, faithful to the original Google TensorFlow implementation with ported pretrained… | 23 | 1654 | maintenance |
| HonglinChu/SiamTrackers A PyTorch collection of Siamese-based visual object tracking models including SiamFC, SiamRPN++, SiamMask, Ocean, LightTrack, and the light… | 32 | 1653 | maintenance |
| vlfeat/vlfeat VLFeat is an open-source C library of popular computer vision algorithms specializing in image understanding and local feature extraction a… | 32 | 1648 | maintenance |
| invictus717/MetaTransformer Meta-Transformer is a research framework for unified multimodal learning that maps inputs from 12 modalities (text, images, point clouds, a… | 19 | 1647 | maintenance |
| facebookresearch/consistent_depth A research library from Facebook AI Research implementing Consistent Video Depth Estimation (SIGGRAPH 2020). It reconstructs dense, flicker… | 10 | 1634 | maintenance |
| raulmur/ORB_SLAM ORB-SLAM is a real-time monocular SLAM system written in C++ that computes camera trajectories and sparse 3D reconstructions from a single … | 32 | 1632 | maintenance |
| DT42/BerryNet BerryNet is a deep learning gateway that turns edge devices like Raspberry Pi into intelligent, offline AI hubs for analyzing camera images… | 23 | 1610 | maintenance |
| experiencor/keras-yolo3 A Keras/TensorFlow implementation of YOLOv3 for object detection, supporting detection with pretrained weights, custom model training with … | 32 | 1608 | maintenance |
| ialhashim/DenseDepth Official Keras/TensorFlow implementation (with experimental PyTorch and TF2 code) of the DenseDepth paper for high-quality monocular depth … | 32 | 1606 | maintenance |
| wy1iu/sphereface Official implementation of SphereFace (Deep Hypersphere Embedding for Face Recognition, CVPR 2017), built on Caffe with a full face recogni… | 32 | 1606 | maintenance |
| chandrikadeb7/Face-Mask-Detection A face mask detection system built with OpenCV and TensorFlow/Keras that uses deep learning (SSD MobileNetV2) to detect whether people are … | 23 | 1606 | maintenance |
| kakaobrain/fast-autoaugment Official PyTorch implementation of Fast AutoAugment (NeurIPS 2019), which learns image augmentation policies via density-matching search. I… | 32 | 1605 | maintenance |
| uzh-rpg/rpg_svo_pro_open SVO Pro is a C++ implementation of Semi-direct Visual Odometry from the Robotics and Perception Group at University of Zurich, supporting m… | 32 | 1601 | maintenance |
| PeterWang512/FALdetector FALdetector is the official PyTorch implementation of the ICCV 2019 paper 'Detecting Photoshopped Faces by Scripting Photoshop'. It provide… | 32 | 1600 | maintenance |
| jiupinjia/stylized-neural-painting Official PyTorch implementation of the CVPR 2021 paper 'Stylized Neural Painting', which translates photos into vectorized painting artwork… | 32 | 1600 | maintenance |
| silvanmelchior/RPi_Cam_Web_Interface A PHP-based web interface for controlling the Raspberry Pi Camera module. It provides live streaming, motion detection, time-lapse capture,… | 34 | 1598 | maintenance |
| cvg/nice-slam NICE-SLAM is a dense RGB-D SLAM system that combines neural implicit decoders with hierarchical grid-based scene representations, published… | 32 | 1597 | maintenance |
| niessner/BundleFusion BundleFusion is a real-time, globally consistent 3D reconstruction system for RGB-D input, published at SIGGRAPH 2017. It estimates globall… | 32 | 1588 | maintenance |
| NVlabs/FUNIT FUNIT is NVIDIA's PyTorch implementation of a few-shot unsupervised image-to-image translation model (ICCV 2019) that can translate images … | 32 | 1586 | maintenance |
| lufficc/SSD A high-quality, fast, modular reference implementation of the SSD (Single Shot MultiBox Detector) object detection model in PyTorch. It sup… | 23 | 1585 | maintenance |
| melodyguan/enas Authors' TensorFlow implementation of Efficient Neural Architecture Search (ENAS), which discovers neural network architectures via paramet… | 32 | 1578 | maintenance |
| snavely/bundler_sfm Bundler is a structure-from-motion (SfM) system that takes unordered image collections with features and matches and produces a 3D reconstr… | 32 | 1578 | maintenance |
| Temporal Segment Networks (TSN) Official code and pretrained models for Temporal Segment Networks (TSN), a deep learning framework for video action recognition published a… | 32 | 1577 | maintenance |
| xinntao/EDVR EDVR is the winning solution of the NTIRE19 video restoration challenges, built on enhanced deformable convolutional networks. The repo is … | 32 | 1577 | maintenance |
| microsoft/Azure-Kinect-Sensor-SDK A cross-platform (Linux and Windows) user-mode SDK in C++ for reading data from the Azure Kinect depth and RGB camera device. It provides a… | 10 | 1575 | maintenance |
| Ewenwan/ORB_SLAM2_SSD_Semantic A C++ research project extending ORB_SLAM2 with dynamic object detection and semantic mapping. It combines SSD-based object detection (Mobi… | 32 | 1574 | maintenance |
| vturrisi/solo-learn solo-learn is a Python library of state-of-the-art self-supervised methods for unsupervised visual representation learning, built on PyTorc… | 65 | 1573 | maintenance |
| cmdbug/YOLOv5_NCNN A mobile demo application that deploys the ncnn inference framework on Android and iOS, running a variety of computer vision models includi… | 32 | 1572 | maintenance |
| kumar-shridhar/PyTorch-BayesianCNN A PyTorch library implementing Bayesian convolutional neural networks with variational inference via Bayes by Backprop. It provides drop-in… | 32 | 1570 | maintenance |
| HumanAIGC/EMO EMO (Emote Portrait Alive) is a research codebase from Alibaba's Institute for Intelligent Computing that generates expressive talking port… | 25 | 7594 | experimental |
| sniklaus/3d-ken-burns A PyTorch reference implementation of the 3D Ken Burns Effect from a Single Image paper, which animates a still photo with a virtual camera… | 70 | 1569 | maintenance |
| AlexHex7/Non-local_pytorch A PyTorch implementation of the Non-local Neural Block from the paper 'Non-local Neural Networks', providing multiple variants (concatenati… | 32 | 1564 | maintenance |
| thu-ml/prolificdreamer Official PyTorch implementation of ProlificDreamer, a NeurIPS 2023 method for high-fidelity text-to-3D generation using Variational Score D… | 29 | 1564 | maintenance |
| NVlabs/noise2noise Official TensorFlow implementation of the Noise2Noise ICML 2018 paper, which trains image restoration networks using only corrupted (noisy)… | 32 | 1563 | maintenance |
| facebookresearch/DeepSDF DeepSDF is Facebook Research's official implementation of the CVPR 2019 paper on learning continuous signed distance functions for 3D shape… | 10 | 1563 | maintenance |
| genforce/interfacegan InterFaceGAN is a Python research library that interprets the latent space of pretrained GANs (PGGAN, StyleGAN) to find semantic subspaces … | 32 | 1561 | maintenance |
| msracver/FCIS FCIS is the official MXNet implementation of the CVPR 2017 paper 'Fully Convolutional Instance-aware Semantic Segmentation', which won firs… | 32 | 1561 | maintenance |
| QingyongHu/RandLA-Net Official TensorFlow implementation of RandLA-Net, a neural architecture for efficient semantic segmentation of large-scale 3D point clouds,… | 32 | 1560 | maintenance |
| V2AI/Det3D Det3D is a PyTorch-based toolbox for 3D object detection from point clouds, offering implementations of models like PointPillars, SECOND, a… | 32 | 1560 | maintenance |
| IDEA-Research/MaskDINO Official PyTorch implementation of Mask DINO, a unified transformer-based framework for object detection and segmentation, built on detectr… | 23 | 1557 | maintenance |
| microsoft/SoM A research toolbox from Microsoft implementing Set-of-Mark (SoM) visual prompting, which overlays numbered spatial marks on images to impro… | 18 | 1557 | maintenance |
| vt-vl-lab/FGVC FGVC is a PyTorch implementation of the ECCV 2020 paper 'Flow-edge Guided Video Completion'. It completes missing regions in videos by extr… | 32 | 1551 | maintenance |
| pytorch/QNNPACK QNNPACK is a mobile-optimized C library of high-performance kernels for 8-bit quantized neural network operators such as convolution, pooli… | 10 | 1551 | maintenance |
| DevashishPrasad/CascadeTabNet CascadeTabNet is a PyTorch/mmdetection implementation of a CVPR 2020 paper for end-to-end table detection and structure recognition from im… | 32 | 1549 | maintenance |
| JiaRenChang/PSMNet PSMNet is the official PyTorch implementation of the CVPR 2018 paper 'Pyramid Stereo Matching Network' for stereo depth estimation. It uses… | 32 | 1548 | maintenance |
| Javacr/PyQt5-YOLOv5 A desktop GUI application built with PyQt5 that wraps YOLOv5 (v6.1) object detection models. It supports running detection on images, video… | 32 | 1547 | maintenance |
| YoYo000/MVSNet MVSNet is a deep learning architecture for depth map inference from unstructured multi-view images, and R-MVSNet is its recurrent extension… | 32 | 1546 | maintenance |
| JingyunLiang/VRT VRT is the official PyTorch implementation of the paper 'VRT: A Video Restoration Transformer', a transformer-based model for video restora… | 23 | 1546 | maintenance |
| google-research/big_transfer Official repository for the Big Transfer (BiT) paper, providing ResNet models pre-trained on ImageNet and ImageNet-21k for transfer learnin… | 10 | 1541 | maintenance |
| dandelin/ViLT Official PyTorch code for the ICML 2021 paper ViLT, a vision-and-language transformer that performs multimodal pre-training without convolu… | 23 | 1538 | maintenance |
| gpleiss/efficient_densenet_pytorch A memory-efficient PyTorch implementation of DenseNets that uses gradient checkpointing to reduce feature map memory consumption from quadr… | 32 | 1535 | maintenance |
| lucidrains/lambda-networks A PyTorch implementation of Lambda Networks, featuring the λ layer that models long-range interactions by transforming contexts into linear… | 23 | 1528 | maintenance |
| mzucker/page_dewarp A Python command-line tool that dewarps photos of curled or warped book pages using a cubic sheet optimization model, flattening them into … | 32 | 1526 | maintenance |
| yu4u/age-gender-estimation A Keras/TensorFlow implementation of a convolutional neural network that estimates age and gender from face images, trained on the IMDB-WIK… | 23 | 1520 | maintenance |
| zhan-xu/RigNet RigNet is a PyTorch implementation of the SIGGRAPH 2020 paper on neural rigging for articulated 3D characters. It takes a character mesh as… | 32 | 1519 | maintenance |
| tanluren/yolov3-channel-and-layer-pruning A Python toolkit built on ultralytics/yolov3 that implements channel pruning, layer pruning, and knowledge distillation for YOLOv3/v4 (incl… | 32 | 1515 | maintenance |
| junyanz/BicycleGAN A PyTorch implementation of BicycleGAN, a model for multimodal image-to-image translation that generates diverse outputs from a single inpu… | 32 | 1514 | maintenance |
| megvii-model/ShuffleNet-Series A collection of ShuffleNet-series efficient convolutional neural network models (V1, V2, V2+, Large, ExLarge) plus NAS-derived backbones li… | 32 | 1513 | maintenance |
| Eric-mingjie/rethinking-network-pruning A PyTorch research codebase reproducing the ICLR 2019 paper 'Rethinking the Value of Network Pruning', which shows pruned models trained fr… | 32 | 1512 | maintenance |
| open-mmlab/Multimodal-GPT Multimodal-GPT is an open-source project for training a multimodal chatbot that combines vision and language instructions, built on OpenFla… | 30 | 1512 | maintenance |
| krasserm/super-resolution A TensorFlow 2.x implementation of EDSR, WDSR, and SRGAN models for single image super-resolution, with a high-level training API and DIV2K… | 32 | 1511 | maintenance |
| compphoto/BoostingMonocularDepth A Python research implementation for boosting monocular depth estimation to high resolution using a double-estimation merging operator, sup… | 32 | 1507 | maintenance |
| tinyvision/SOLIDER SOLIDER is a semantic-controllable self-supervised learning framework that learns general human representations from massive unlabeled huma… | 32 | 1504 | maintenance |
| yulunzhang/RCAN PyTorch implementation of RCAN, a very deep residual channel attention network for single image super-resolution from an ECCV 2018 paper. I… | 39 | 1503 | maintenance |
| AITTSMD/MTCNN-Tensorflow A TensorFlow reproduction of MTCNN (Multi-task Cascaded Convolutional Networks) for joint face detection and facial landmark alignment. It … | 32 | 1502 | maintenance |
| cruxopen/openISP An open-source Python implementation of an Image Signal Processor (ISP) pipeline that converts RAW sensor images to RGB/YUV. It simulates h… | 32 | 1502 | maintenance |
| vacancy/Synchronized-BatchNorm-PyTorch A PyTorch library implementing Synchronized Batch Normalization, which computes batch statistics across all GPUs during multi-device traini… | 32 | 1502 | maintenance |
| luuuyi/CBAM.PyTorch A non-official PyTorch re-implementation of the CBAM (Convolutional Block Attention Module) paper from ECCV 2018. It provides channel and s… | 32 | 1500 | maintenance |
| filipradenovic/cnnimageretrieval-pytorch A PyTorch toolbox for training and evaluating convolutional neural networks for image retrieval, implementing the authors' TPAMI 2018 and E… | 23 | 1495 | maintenance |