domain: deep-learning
2771 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| trevin-creator/autoresearch-mlx An Apple Silicon (MLX) port of Karpathy's autoresearch that runs autonomous AI research loops on a Mac without PyTorch or CUDA. A coding ag… | 55 | 1813 | active |
| yerfor/GeneFacePlusPlus GeneFace++ is the official PyTorch implementation of a NeRF-based system for generalized and stable real-time 3D talking face generation. I… | 26 | 1809 | active |
| apple/ml-4m 4M is a framework from Apple and EPFL for training any-to-any multimodal foundation models using masked modeling over discrete tokens acros… | 35 | 1808 | active |
| zju3dv/4K4D 4K4D is a research implementation of a 4D point cloud representation for real-time dynamic view synthesis at up to 4K resolution, built on … | 27 | 1807 | active |
| Robbyant/lingbot-va LingBot-VA is an autoregressive diffusion framework that unifies video world modeling and robot policy learning in a single interleaved vid… | 56 | 1806 | active |
| triple-mu/YOLOv8-TensorRT A library for running YOLOv8 inference accelerated with NVIDIA TensorRT, supporting detection, segmentation, pose estimation, oriented boun… | 75 | 1804 | active |
| HuangJunJie2017/BEVDet BEVDet is a Python research codebase implementing the BEVDet series of bird's-eye-view (BEV) 3D object detection models for autonomous driv… | 23 | 1801 | active |
| Xilinx/Vitis-AI AMD Vitis AI is an integrated development environment and stack for accelerating AI inference on AMD/Xilinx adaptable platforms, including … | 71 | 1800 | active |
| OpenImagingLab/FlashVSR FlashVSR is a one-step diffusion-based streaming video super-resolution framework that runs at ~17 FPS for 768x1408 video on a single A100 … | 61 | 1799 | active |
| acids-ircam/RAVE RAVE is the official PyTorch implementation of a realtime audio variational autoencoder for fast, high-quality neural audio synthesis. It s… | 55 | 1790 | active |
| lehaifeng/T-GCN A collection of research source code implementing Temporal Graph Convolutional Networks (T-GCN) and related variants for urban traffic flow… | 50 | 1781 | active |
| huchenlei/ComfyUI-layerdiffuse A ComfyUI custom node plugin implementing LayerDiffuse, enabling generation of transparent images with RGBA output and foreground/backgroun… | 29 | 1779 | active |
| thu-ml/RoboticsDiffusionTransformer RDT-1B is a 1B-parameter diffusion foundation model for robot bimanual manipulation, pre-trained on 1M+ multi-robot episodes to predict rob… | 50 | 1778 | active |
| Totoro97/NeuS Official PyTorch implementation of NeuS, a neural implicit surface reconstruction method that learns SDF-based surfaces via volume renderin… | 32 | 1777 | stable |
| CalculatedContent/WeightWatcher WeightWatcher is an open-source Python diagnostic tool for analyzing pre-trained deep neural networks without access to training or test da… | 59 | 1772 | active |
| xinntao/ESRGAN ESRGAN (Enhanced SRGAN) is a PyTorch-based image super-resolution model that won the PIRM 2018 Challenge on Perceptual Super-Resolution. Th… | 32 | 6568 | maintenance |
| SamsungSAILMontreal/TinyRecursiveModels Official codebase for the Tiny Recursive Model (TRM) paper, a recursive reasoning approach where a tiny 7M-parameter neural network iterati… | 10 | 6566 | maintenance |
| VAST-AI-Research/TripoSG TripoSG is an open-source image-to-3D generation foundation model that produces high-fidelity 3D meshes from single images using large-scal… | 27 | 1755 | active |
| webonnx/wonnx Wonnx is a GPU-accelerated ONNX inference runtime written entirely in Rust, built on wgpu and usable natively or in the browser via WebGPU … | 10 | 1754 | active |
| microsoft/mup The `mup` Python package implements Maximal Update Parametrization (μP) for PyTorch models, enabling optimal hyperparameters to remain stab… | 23 | 1753 | stable |
| facebookresearch/metaseq Metaseq is a PyTorch codebase from Meta AI for training and working with large-scale Open Pre-trained Transformers (OPT), forked from fairs… | 10 | 6548 | maintenance |
| google-research/text-to-text-transfer-transformer The official T5 library from Google Research, implementing the text-to-text transfer transformer for NLP tasks like summarization, question… | 64 | 6544 | maintenance |
| octo-models/octo Octo is an open-source generalist robot policy: a transformer-based diffusion policy pretrained on 800k robot trajectories from the Open X-… | 17 | 1751 | active |
| Stability-AI/StableCascade Official codebase for Stable Cascade, a text-to-image generation model built on the Würstchen architecture with a highly compressed latent … | 26 | 6540 | maintenance |
| character-ai/Ovi Ovi is a video-plus-audio generation model from Character AI that simultaneously generates synchronized video and audio from text or text+i… | 40 | 1748 | active |
| codertimo/BERT-pytorch A PyTorch implementation of Google AI's 2018 BERT model with simple, readable code. It provides CLI tools for building vocabulary and pre-t… | 23 | 6527 | maintenance |
| CompVis/taming-transformers The official implementation of 'Taming Transformers for High-Resolution Image Synthesis' (CVPR 2021), combining a convolutional VQGAN codeb… | 32 | 6521 | maintenance |
| zhouhaoyi/Informer2020 The official PyTorch implementation of Informer, an efficient Transformer architecture for long sequence time-series forecasting that won t… | 44 | 6516 | maintenance |
| rusty1s/pytorch_scatter A PyTorch extension library providing highly optimized scatter and segment (sparse update) operations with sum, mean, min, and max reductio… | 61 | 1745 | stable |
| TencentARC/BrushNet BrushNet is the official PyTorch implementation of an ECCV 2024 plug-and-play image inpainting model that embeds pixel-level masked image f… | 25 | 1745 | active |
| deepseek-ai/TileKernels TileKernels is a Python library of optimized GPU kernels for LLM operations, written in the TileLang DSL. It provides kernels for MoE routi… | 49 | 1743 | active |
| openai/consistency_models Official PyTorch implementation of Consistency Models, a generative image model family from OpenAI supporting consistency distillation, con… | 10 | 6486 | maintenance |
| tile-ai/TileRT TileRT is a tile-based runtime for ultra-low-latency LLM inference, achieving hundreds to over 1000 tokens/s decode speeds on frontier mode… | 77 | 1738 | active |
| Xiaojiu-z/EasyControl EasyControl is the official implementation of an ICCV 2025 paper adding efficient and flexible conditional control to Diffusion Transformer… | 35 | 1737 | active |
| SHI-Labs/OneFormer OneFormer is a universal image segmentation framework (CVPR 2023) that unifies semantic, instance, and panoptic segmentation in a single tr… | 32 | 1736 | stable |
| google/automl Google Brain's AutoML repository containing implementations of AutoML models and libraries such as EfficientNet, EfficientNetV2, and Effici… | 10 | 6474 | maintenance |
| facebookresearch/multimodal TorchMultimodal is a PyTorch library from Meta for training state-of-the-art multimodal multi-task models at scale, covering both content u… | 77 | 1732 | active |
| cleverhans-lab/cleverhans CleverHans is a Python library for benchmarking machine learning systems' vulnerability to adversarial examples, providing reference implem… | 23 | 6450 | maintenance |
| NVIDIA/FasterTransformer NVIDIA's highly optimized C++/CUDA library for fast inference of Transformer-based models such as BERT, GPT, and encoder-decoder models, wi… | 23 | 6447 | maintenance |
| mahmoodlab/CLAM CLAM is an open-source Python toolkit for data-efficient, weakly supervised classification of whole-slide images (WSIs) in computational pa… | 39 | 1728 | active |
| MultimediaTechLab/YOLO Official MIT-licensed implementation of the YOLOv9, YOLOv7, and YOLO-RD real-time object detection models, including pre-trained weights, t… | 56 | 1723 | active |
| facebookresearch/ConvNeXt Official PyTorch implementation of ConvNeXt, a pure convolutional neural network architecture from the CVPR 2022 paper 'A ConvNet for the 2… | 10 | 6416 | maintenance |
| meta-pytorch/tnt TNT (torchtnt) is a lightweight library from Meta providing tools and utilities for PyTorch training workflows. It offers building blocks f… | 76 | 1721 | active |
| cocodataset/cocoapi Official API for the COCO (Common Objects in Context) dataset, providing Matlab, Python, and Lua interfaces to load, parse, and visualize C… | 32 | 6384 | maintenance |
| kingoflolz/mesh-transformer-jax A JAX/Haiku library implementing model-parallel training and inference of transformer models using xmap/pjit operators, similar to Megatron… | 32 | 6380 | maintenance |
| AnswerDotAI/ModernBERT ModernBERT is the research repository for a modernized BERT-family bidirectional encoder trained on 2 trillion tokens with an 8192-token co… | 56 | 1713 | active |
| feizc/FluxMusic FluxMusic is the official PyTorch implementation of a rectified flow Transformer model for text-to-music generation, from the paper 'Flux t… | 23 | 1711 | active |
| omerbt/TokenFlow TokenFlow is the official PyTorch implementation of an ICLR 2024 paper for text-driven, temporally consistent video editing using a pre-tra… | 30 | 1708 | stable |
| lsdefine/simple_GRPO A minimal (~200 lines) Python implementation of GRPO reinforcement learning for training LLMs to develop r1-like reasoning, built on PyTorc… | 44 | 1702 | active |
| iMoonLab/yolov13 Official PyTorch implementation of YOLOv13, a real-time object detection model family (Nano to X-Large) featuring Hypergraph-based Adaptive… | 32 | 1702 | active |
| PaddlePaddle/PaddleVideo PaddleVideo is a video understanding toolkit built on PaddlePaddle, offering state-of-the-art models for action recognition, temporal actio… | 26 | 1702 | active |
| AvaLovelace1/BrickGPT BrickGPT is the official implementation of an ICCV 2025 Best Paper approach that generates physically stable, buildable toy brick (LEGO) mo… | 57 | 1701 | active |
| sihyun-yu/REPA REPA is the official PyTorch implementation of the ICLR 2025 paper 'Representation Alignment for Generation', a regularization technique th… | 27 | 1700 | active |
| jiaweizzhao/GaLore GaLore is a PyTorch library implementing Gradient Low-Rank Projection for memory-efficient full-parameter training of large language models… | 25 | 1700 | active |
| BindsNET/bindsnet BindsNET is a Python package for simulating spiking neural networks (SNNs) built on PyTorch tensor functionality, running on CPUs or GPUs. … | 84 | 1695 | active |
| tensorpack/tensorpack Tensorpack is a high-level neural network training interface built on graph-mode TensorFlow, focused on training speed and flexibility for … | 23 | 6286 | maintenance |
| lucidrains/alphafold3-pytorch A PyTorch implementation of Google DeepMind's Alphafold 3 model for protein and molecular structure prediction. It is a research-oriented r… | 68 | 1691 | active |
| facebookresearch/coconut Official PyTorch implementation of Coconut, a method for training large language models to reason in a continuous latent space instead of e… | 61 | 1689 | active |
| elixir-nx/axon Axon is a neural network library for Elixir built on top of Nx numerical definitions. It provides a functional API, a high-level model crea… | 79 | 1688 | active |
| intelligent-machine-learning/dlrover DLRover is an automatic distributed deep learning system that manages training of large AI models on Kubernetes and Ray clusters. It provid… | 79 | 1680 | active |
| EnzymeAD/Enzyme Enzyme is a high-performance automatic differentiation plugin for LLVM and MLIR that computes derivatives and gradients of arbitrary existi… | 95 | 1679 | active |
| takuseno/d3rlpy d3rlpy is a Python library for offline and online deep reinforcement learning built on PyTorch, offering state-of-the-art algorithms throug… | 52 | 1679 | active |
| franciszzj/Leffa Leffa is a diffusion-based framework for controllable person image generation, supporting virtual try-on and pose transfer via a regulariza… | 40 | 1672 | active |
| google/vizier Open Source Vizier is a Python-based service and library for black-box and hyperparameter optimization, based on Google's internal Vizier t… | 79 | 1671 | active |
| shubham-goel/4D-Humans 4DHumans is a Python research codebase implementing HMR 2.0, a transformer-based model for 3D human mesh recovery from single images, plus … | 59 | 1671 | active |
| jd-opensource/JoyAI-Video-Edit JoyAI-Video-Edit is a real-time, instruction-guided video editing system that applies natural-language edits to live or uploaded video stre… | 57 | 1671 | active |
| ZrrSkywalker/Personalize-SAM PerSAM is the official implementation of 'Personalize Segment Anything Model with One Shot', which customizes the Segment Anything Model (S… | 29 | 1671 | active |
| zihangdai/xlnet XLNet is a generalized autoregressive pretraining method for language representation learning, built on Transformer-XL. This repo provides … | 32 | 6188 | maintenance |
| tkarras/progressive_growing_of_gans Official TensorFlow implementation of the ICLR 2018 NVIDIA paper 'Progressive Growing of GANs', which trains generators and discriminators … | 32 | 6179 | maintenance |
| Gen-Verse/MMaDA MMaDA is an open-source family of multimodal large diffusion language models that unify textual reasoning, multimodal understanding, and te… | 49 | 1668 | active |
| JiuhaiChen/BLIP3o Official implementation of the BLIP3o-Series, a unified autoregressive-plus-diffusion model for text-to-image generation and editing. It co… | 44 | 1666 | active |
| elixir-nx/bumblebee Bumblebee is an Elixir library providing pre-trained neural network models built on Axon, with integration for downloading models from Hugg… | 89 | 1662 | active |
| mit-han-lab/torchquantum TorchQuantum is a PyTorch-based framework for quantum computing simulation, supporting statevector and pulse-level simulation on GPUs with … | 64 | 1662 | active |
| thunil/TecoGAN TecoGAN is the official source code for a temporally coherent GAN for video super-resolution, published at SIGGRAPH/ACM TOG. It includes in… | 32 | 6141 | maintenance |
| thtrieu/darkflow Darkflow is a Python library that translates Darknet's YOLO neural network definitions to TensorFlow, enabling real-time object detection a… | 32 | 6139 | maintenance |
| chineseocr A Python OCR toolkit that combines YOLO3-based text detection with CRNN/Dense recognition for Chinese and English text in natural scene ima… | 32 | 6123 | maintenance |
| mhamilton723/FeatUp FeatUp is a model-agnostic framework that upsamples the spatial resolution of deep neural network features by 16-32x without changing their… | 16 | 1654 | active |
| taki0112/UGATIT Official TensorFlow implementation of U-GAT-IT, an unsupervised image-to-image translation model using attention modules and adaptive layer… | 32 | 6116 | maintenance |
| maziarraissi/PINNs The original reference implementation of Physics-Informed Neural Networks (PINNs), which train neural networks to solve and discover nonlin… | 62 | 6111 | maintenance |
| lightly-ai/lightly-train LightlyTrain is a Python framework for training computer vision models, covering pretraining of vision foundation models (DINOv2/v3) on unl… | 85 | 1652 | active |
| sunlabuiuc/PyHealth PyHealth is an open-source Python toolkit for clinical deep learning, unifying healthcare datasets (EHRs, physiological signals, medical im… | 91 | 1650 | active |
| XueZeyue/DanceGRPO Official implementation of DanceGRPO, a framework applying Group Relative Policy Optimization (GRPO) to fine-tune visual generation models … | 40 | 1648 | active |
| JIA-Lab-research/ControlNeXt ControlNeXt is the official implementation of a controllable generation method for images and videos, built on Stable Diffusion XL, Stable … | 24 | 1646 | active |
| tenstorrent/tt-metal TT-Metal is Tenstorrent's open-source software stack containing TT-NN, a Python and C++ neural network operator library, and TT-Metalium, a… | 97 | 1640 | active |
| CoinCheung/BiSeNet A PyTorch implementation of the BiSeNet V1 and V2 real-time semantic segmentation models, with pretrained weights for Cityscapes, COCO-Stuf… | 57 | 1638 | active |
| yenchenlin/nerf-pytorch A PyTorch implementation of NeRF (Neural Radiance Fields) that reproduces the original paper's results for synthesizing novel views of comp… | 32 | 6044 | maintenance |
| luanfujun/deep-painterly-harmonization Research code implementing the 'Deep Painterly Harmonization' algorithm, which seamlessly blends a pasted object into a painting's style us… | 32 | 6042 | maintenance |
| Picsart-AI-Research/StreamingT2V StreamingT2V is a research codebase (CVPR 2025) implementing an autoregressive technique that turns short text-to-video diffusion models li… | 31 | 1630 | active |
| yoshitomo-matsubara/torchdistill torchdistill is a modular, configuration-driven PyTorch framework for knowledge distillation and general deep learning experiments, requiri… | 86 | 1629 | active |
| Fantasy-AMAP/fantasy-talking FantasyTalking is a research codebase and model for generating realistic talking portrait videos from a single image and an audio clip, bui… | 48 | 1628 | active |
| InterDigitalInc/CompressAI CompressAI is a PyTorch library and evaluation platform for end-to-end learned data compression research. It provides custom layers, entrop… | 73 | 1627 | active |
| ZhangGe6/onnx-modifier onnx-modifier is a web-based visual editor for ONNX neural network models, built on Netron for graph visualization and Flask as the backend… | 72 | 1627 | active |
| ZiqiaoPeng/SyncTalk SyncTalk is the official PyTorch implementation of a CVPR 2024 paper that synthesizes speech-driven, synchronized talking head videos using… | 46 | 1626 | active |
| gomlx/gomlx GoMLX is an accelerated machine learning and math framework for Go, comparable to PyTorch/JAX/TensorFlow. It offers differentiable operator… | 95 | 1621 | active |
| jeffffffli/HybrIK HybrIK is the official PyTorch implementation of a hybrid analytical-neural inverse kinematics method for 3D human pose and shape estimatio… | 23 | 1618 | stable |
| PaddlePaddle/PaddleSlim PaddleSlim is an open-source library built on PaddlePaddle for deep learning model compression and architecture search. It provides low-bit… | 50 | 1612 | active |
| gorgonia/gorgonia Gorgonia is a Go library for machine learning that lets you define and evaluate mathematical equations over multidimensional arrays using a… | 23 | 5929 | maintenance |
| chainer/chainer Chainer is a Python-based deep learning framework that pioneered the define-by-run approach with dynamic computational graphs and automatic… | 23 | 5924 | maintenance |
| dmlc/gluon-cv GluonCV is a deep learning toolkit providing state-of-the-art computer vision model implementations with 170+ pre-trained models. It suppor… | 23 | 5916 | maintenance |
| OpenGVLab/LLaMA-Adapter Official implementation of LLaMA-Adapter and its V2 successor, a parameter-efficient fine-tuning method that adapts LLaMA models to follow … | 21 | 5914 | maintenance |