domain: deep-learning
2771 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| NVIDIAGameWorks/kaolin Kaolin is NVIDIA's PyTorch library of GPU-optimized modules for 3D deep learning research, covering meshes, point clouds, and 3D Gaussian s… | 68 | 5161 | active |
| aigc-apps/sd-webui-EasyPhoto EasyPhoto is a Stable Diffusion WebUI plugin for generating AI portraits by training a personal 'digital doppelganger' from 5-20 user photo… | 28 | 5155 | active |
| open-mmlab/mmaction2 MMAction2 is OpenMMLab's PyTorch-based toolbox and benchmark for video understanding, covering action recognition, temporal action localiza… | 55 | 5142 | active |
| ai-dawang/PlugNPlay-Modules A curated collection of plug-and-play deep learning modules (convolutions, attention mechanisms, downsampling, and feature fusion blocks) i… | 38 | 5105 | active |
| facebookresearch/AugLy AugLy is a Python data augmentation library from Meta AI supporting audio, image, text, and video with over 100 augmentations. It focuses o… | 67 | 5089 | stable |
| AILab-CVC/VideoCrafter VideoCrafter is an open-source video generation and editing toolbox built on video diffusion models, offering Text-to-Video and Image-to-Vi… | 48 | 5073 | active |
| Deci-AI/super-gradients SuperGradients is an open-source PyTorch-based training library for building, training, and fine-tuning state-of-the-art computer vision mo… | 54 | 5052 | active |
| deepseek-ai/DeepSeek-V2 DeepSeek-V2 is a strong, economical Mixture-of-Experts language model released with open weights, inference code, and evaluation tooling. T… | 24 | 5039 | active |
| lightvector/KataGo KataGo is an open-source Go (baduk) engine trained via AlphaZero-like self-play, one of the strongest Go bots available. It runs as a GTP e… | 99 | 5036 | active |
| NVIDIA/nccl NVIDIA's Collective Communication Library (NCCL) is a C++ library providing topology-aware, high-bandwidth inter-GPU communication primitiv… | 98 | 5027 | stable |
| Plachtaa/VITS-fast-fine-tuning A Python pipeline for fast fine-tuning of VITS text-to-speech models, enabling speaker adaptation in under an hour from short audio, long a… | 10 | 5012 | active |
| pytorch/executorch ExecuTorch is PyTorch's framework for exporting and running AI models on-device across mobile, embedded, and edge hardware, with a tiny (~5… | 95 | 4953 | active |
| microsoft/muzic Muzic is a Microsoft Research project providing deep learning models for music understanding and generation, including MusicBERT, SongMASS,… | 66 | 4952 | active |
| KaiyangZhou/deep-person-reid Torchreid is a PyTorch library for deep-learning person re-identification, supporting both image and video reid with end-to-end training an… | 50 | 4900 | stable |
| PaddlePaddle/VisualDL VisualDL is a deep learning visualization toolkit for PaddlePaddle that provides charts for tracking training metrics, visualizing model st… | 24 | 4884 | stable |
| deepjavalibrary/djl Deep Java Library (DJL) is an engine-agnostic, high-level deep learning framework for Java that supports backends like PyTorch, TensorFlow,… | 80 | 4842 | active |
| ace-step/ACE-Step ACE-Step is an open-source foundation model for music generation that combines diffusion-based generation with a deep compression autoencod… | 49 | 4788 | active |
| pytorch/ignite PyTorch-Ignite is a high-level library for training and evaluating neural networks in PyTorch flexibly and transparently. It provides an En… | 89 | 4778 | stable |
| facebookresearch/lingua Meta Lingua is a minimal, fast LLM training and inference library built on easy-to-modify PyTorch components for research purposes. It supp… | 36 | 4766 | active |
| open-mmlab/mmocr MMOCR is OpenMMLab's PyTorch-based toolbox for text detection, recognition, and key information extraction. It provides a model zoo of OCR … | 23 | 4752 | active |
| FluxML/Flux.jl Flux.jl is a machine learning library written entirely in Julia, providing lightweight abstractions over Julia's native GPU support and aut… | 98 | 4739 | stable |
| OpenDriveLab/UniAD UniAD is a unified end-to-end autonomous driving framework that hierarchically casts perception, prediction, and planning tasks under a pla… | 44 | 4737 | active |
| cvg/LightGlue LightGlue is a deep neural network library that matches sparse local features across image pairs with high accuracy and fast inference. It … | 50 | 4728 | stable |
| facebookresearch/flow_matching A PyTorch library for implementing flow matching algorithms with continuous, discrete, and Riemannian flow matching implementations. It acc… | 48 | 4705 | active |
| mindspore-ai/mindspore MindSpore is an open-source deep learning framework for training and inference across mobile, edge, and cloud scenarios. It provides automa… | 32 | 4700 | active |
| PKU-Alignment/align-anything Align-Anything is a modular Python framework for aligning any-to-any (all-modality) large models with human intentions and values using fee… | 48 | 4667 | active |
| Blealtan/efficient-kan An efficient pure-PyTorch implementation of Kolmogorov-Arnold Networks (KAN) that reformulates the computation as matrix multiplications ov… | 24 | 4655 | active |
| Tencent/TNN TNN is a high-performance, lightweight deep learning inference framework developed by Tencent Youtu Lab, supporting mobile, desktop, and se… | 32 | 4648 | active |
| OpenNMT/CTranslate2 CTranslate2 is a C++ and Python library for fast, memory-efficient inference of Transformer models on CPU and GPU. It uses quantization, la… | 96 | 4644 | stable |
| WhisperSpeech/WhisperSpeech WhisperSpeech is an open-source text-to-speech system built by inverting OpenAI's Whisper model, aiming to be 'Stable Diffusion for speech'… | 56 | 4639 | active |
| Rikorose/DeepFilterNet DeepFilterNet is a low-complexity speech enhancement framework that performs real-time noise suppression on full-band 48kHz audio using dee… | 23 | 4632 | active |
| NVlabs/neuralangelo Official PyTorch implementation of Neuralangelo, a CVPR 2023 method for high-fidelity neural surface reconstruction from multi-view images.… | 29 | 4615 | active |
| Kwai-Kolors/Kolors Kolors is a large-scale latent diffusion model for photorealistic text-to-image synthesis, trained with bilingual (Chinese and English) tex… | 23 | 4615 | active |
| deepseek-ai/Engram Official implementation of Engram, a conditional memory module from DeepSeek that modernizes N-gram embeddings for O(1) lookup as a new spa… | 43 | 4614 | active |
| sensity-ai/dot dot (Deepfake Offensive Toolkit) is a Python tool that generates real-time, controllable deepfakes from a webcam feed and injects them into… | 23 | 4586 | active |
| RecBole RecBole is a unified, comprehensive and efficient recommendation library built on Python and PyTorch for reproducing and developing recomme… | 26 | 4541 | stable |
| Tencent-Hunyuan/HunyuanVideo-1.5 HunyuanVideo-1.5 is Tencent's lightweight 8.3B-parameter video generation model supporting text-to-video and image-to-video synthesis. The … | 50 | 4534 | active |
| OAID/Tengine Tengine is a lightweight, high-performance, modular deep learning inference engine developed by OPEN AI LAB for embedded and edge devices. … | 27 | 4531 | active |
| NVlabs/tiny-cuda-nn A small, self-contained C++/CUDA framework for training and querying neural networks, featuring a lightning-fast fully fused MLP and a vers… | 58 | 4528 | active |
| facebookresearch/vjepa2 Official PyTorch codebase and pretrained models for V-JEPA 2, a self-supervised video encoder trained on internet-scale video, plus V-JEPA … | 52 | 4527 | active |
| real-stanford/diffusion_policy Official PyTorch implementation of Diffusion Policy, a visuomotor robot policy learning method that represents robot behavior as a conditio… | 30 | 4491 | stable |
| Mobile ALOHA Mobile ALOHA is an open-source system for low-cost whole-body teleoperation and data collection with a bimanual mobile manipulation robot, … | 29 | 4465 | active |
| layumi/Person_reID_baseline_pytorch A small, friendly PyTorch baseline implementation for person and vehicle re-identification (ReID). It reproduces strong top-conference resu… | 65 | 4446 | stable |
| modelscope/ClearerVoice-Studio ClearerVoice-Studio is an open-source, AI-powered speech processing toolkit from ModelScope/Alibaba offering state-of-the-art pretrained mo… | 38 | 4445 | active |
| mosaicml/llm-foundry LLM Foundry is a PyTorch-based codebase for training, finetuning, evaluating, and deploying large language models from 125M to 70B+ paramet… | 65 | 4441 | active |
| xlite-dev/lite.ai.toolkit A lightweight C++ toolkit providing unified APIs for 100+ pre-trained AI models across inference backends like ONNX Runtime, MNN, TensorRT,… | 74 | 4427 | active |
| tensorflow/probability TensorFlow Probability is a Python library for probabilistic reasoning and statistical analysis built on TensorFlow, with a JAX substrate. … | 66 | 4427 | active |
| huawei-noah/Efficient-AI-Backbones A collection of efficient neural network backbone architectures (GhostNet, TNT, ViG, WaveMLP, TinyNet, etc.) from Huawei Noah's Ark Lab, wi… | 28 | 4418 | active |
| zai-org/GLM-4.5 GLM-4.5 (and successors GLM-4.6/4.7) is a family of open-weight Mixture-of-Experts foundation models from Z.ai focused on agentic tasks, re… | 47 | 4416 | active |
| ROCm/hip HIP is a C++ runtime API and kernel programming language for AMD GPUs that mirrors the NVIDIA CUDA programming interface. It enables develo… | 95 | 4392 | active |
| iperov/DeepFaceLab DeepFaceLab is the leading open-source Windows application for creating deepfakes, allowing users to swap, de-age, or replace faces and hea… | 10 | 19292 | maintenance |
| onnxsim/onnxsim ONNX Simplifier (onnxsim) is a tool that simplifies ONNX models by inferring the whole computation graph and replacing redundant operators … | 98 | 4390 | active |
| lululxvi/deepxde DeepXDE is a Python library for scientific machine learning and physics-informed learning, built on top of PyTorch, TensorFlow, JAX, and Pa… | 81 | 4386 | stable |
| bowang-lab/MedSAM MedSAM is a fine-tuned Segment Anything Model (SAM) foundation model for universal medical image segmentation, trained on over 1.5 million … | 29 | 4379 | active |
| AI4Finance-Foundation/ElegantRL ElegantRL is a lightweight, modular deep reinforcement learning library built on PyTorch that implements core model-free RL algorithms (PPO… | 53 | 4355 | active |
| Tencent-Hunyuan/HunyuanDiT Hunyuan-DiT is Tencent's open-source diffusion transformer model for text-to-image generation with fine-grained Chinese language understand… | 49 | 4291 | active |
| richzhang/PerceptualSimilarity A PyTorch library implementing the LPIPS (Learned Perceptual Image Patch Similarity) metric, which measures perceptual distance between ima… | 23 | 4269 | stable |
| Nixtla/neuralforecast NeuralForecast is a Python library offering a large collection of state-of-the-art neural forecasting models (NBEATS, NHITS, TFT, PatchTST,… | 98 | 4257 | active |
| ThilinaRajapakse/simpletransformers Simple Transformers is a Python library built on Hugging Face Transformers that lets users train, fine-tune, and evaluate Transformer model… | 61 | 4254 | active |
| ModelTC/LightLLM LightLLM is a Python-based LLM inference and serving framework designed for lightweight deployment, easy scalability, and high throughput. … | 84 | 4243 | active |
| SwanHubX/SwanLab SwanLab is an open-source AI training tracking and visualization platform with a Python SDK, CLI, and cloud or self-hosted dashboard. It in… | 89 | 4175 | active |
| facebookresearch/vggt-omega VGGT-Omega is a research library from Oxford VGG and Meta AI providing pretrained transformer models for 3D vision tasks such as camera pos… | 58 | 4165 | active |
| torchgeo/torchgeo TorchGeo is a PyTorch domain library, similar to torchvision, providing datasets, samplers, transforms, and pre-trained models specific to … | 95 | 4159 | active |
| ArcInstitute/evo2 Evo 2 is a DNA foundation language model (1B-40B parameters) that models genomes at single-nucleotide resolution with up to 1 million base … | 62 | 4157 | active |
| openai/transformer-debugger Transformer Debugger (TDB) is an OpenAI tool for investigating specific behaviors of small language models, combining automated interpretab… | 60 | 4124 | active |
| IDEA-CCNL/Fengshenbang-LM Fengshenbang-LM is an open-source suite of Chinese large language models and a PyTorch training framework from IDEA Research's CCNL lab, ai… | 71 | 4123 | active |
| opengeos/segment-geospatial SamGeo (segment-geospatial) is a Python package that applies Meta AI's Segment Anything Model (SAM, SAM2, SAM3, HQ-SAM) to geospatial data … | 97 | 4122 | active |
| facebookresearch/jepa Official PyTorch implementation of V-JEPA, a self-supervised method for learning visual representations from video using a joint-embedding … | 29 | 4105 | active |
| GuyTevet/motion-diffusion-model Official PyTorch implementation of the Human Motion Diffusion Model (MDM) paper, generating 3D human motion sequences from text prompts usi… | 52 | 4092 | active |
| princeton-vl/RAFT Official PyTorch implementation of RAFT (Recurrent All Pairs Field Transforms for Optical Flow), an ECCV 2020 model for estimating dense op… | 49 | 4091 | stable |
| PaddlePaddle/PaddleRec PaddleRec is a large-scale recommendation algorithm library built on PaddlePaddle, containing classic and state-of-the-art recommendation m… | 29 | 4085 | active |
| hao-ai-lab/FastVideo FastVideo is a unified Python framework for post-training and real-time inference of video diffusion models, covering data preprocessing, f… | 80 | 4076 | active |
| QwenLM/Qwen2.5-Omni Qwen2.5-Omni is an end-to-end multimodal model from Alibaba's Qwen team that understands text, images, audio, and video, and generates stre… | 31 | 4074 | active |
| facebookresearch/dlrm A PyTorch-based implementation of the Deep Learning Recommendation Model (DLRM) from Facebook Research, which processes dense and sparse fe… | 60 | 4065 | stable |
| FedML-AI/FedML FedML (TensorOpera) is a unified Python library for large-scale distributed training, model serving, and federated learning across GPU clou… | 45 | 4062 | active |
| thinking-machines-lab/tinker-cookbook Tinker Cookbook is a Python library of realistic examples and abstractions for post-training (fine-tuning) language models via the Tinker A… | 85 | 4058 | active |
| uxlfoundation/oneDNN oneDNN is an open-source cross-platform performance library providing optimized building blocks (primitives) for deep learning applications… | 99 | 4042 | stable |
| tensorflow/tensor2tensor Tensor2Tensor (T2T) is a Python library of deep learning models and datasets built on TensorFlow, developed by the Google Brain team to mak… | 10 | 17464 | maintenance |
| lucidrains/vector-quantize-pytorch A PyTorch library implementing vector and scalar quantization, including Residual VQ and techniques like DiVeQ codebook updates. It origina… | 83 | 3996 | active |
| ML-GSAI/LLaDA Official PyTorch implementation of LLaDA, a family of large language diffusion models (8B base/instruct, MoE, and iLLaDA variants) with pre… | 62 | 3943 | active |
| Nunchaku Nunchaku is a high-performance inference engine for 4-bit quantized diffusion models (and LLMs) based on the SVDQuant technique from an ICL… | 65 | 3937 | active |
| ali-vilab/VACE VACE is the official implementation of an all-in-one video creation and editing model from Tongyi Lab, built on Wan2.1 diffusion models. It… | 41 | 3934 | active |
| thuml/Transfer-Learning-Library TLlib is a PyTorch-based open-source library for transfer learning, covering domain adaptation, task adaptation (finetuning), and domain ge… | 23 | 3931 | active |
| Lightning-AI/LitServe LitServe is a Python framework for building custom AI inference servers with full control over batching, routing, streaming, and scaling lo… | 92 | 3930 | active |
| stanford-futuredata/ColBERT ColBERT is a fast and accurate neural retrieval model that encodes passages and queries into token-level embedding matrices and scores them… | 45 | 3924 | active |
| Tencent-Hunyuan/Hunyuan3D-2.1 Tencent's open-source 3D asset generation model that creates high-fidelity 3D meshes with production-ready PBR materials from single images… | 40 | 3917 | active |
| iree-org/iree IREE is an MLIR-based end-to-end machine learning compiler and runtime that lowers models from frameworks like PyTorch, TensorFlow, JAX, an… | 87 | 3901 | active |
| hustvl/Vim Vision Mamba (Vim) is a PyTorch implementation of a generic vision backbone built on bidirectional Mamba state space models, published at I… | 29 | 3899 | active |
| hustvl/4DGaussians An official PyTorch implementation of 4D Gaussian Splatting (4D-GS) for real-time rendering of dynamic scenes, published at CVPR 2024. It c… | 27 | 3895 | active |
| Plachtaa/seed-vc Seed-VC is a Python tool and model for zero-shot voice conversion, real-time voice conversion, and singing voice conversion, cloning a voic… | 10 | 3888 | active |
| safetensors/safetensors Safetensors is a simple, secure file format and library for storing and distributing tensors, designed as a fast zero-copy alternative to p… | 91 | 3876 | stable |
| hkust-nlp/simpleRL-reason A research codebase from HKUST-NLP implementing a simple reinforcement learning recipe (rule-based rewards on GSM8K/Math data) to train LLM… | 47 | 3874 | active |
| mseitzer/pytorch-fid A PyTorch port of the official TensorFlow implementation of the Fréchet Inception Distance (FID), a metric for measuring similarity between… | 23 | 3851 | stable |
| open-mmlab/mmpretrain MMPretrain is OpenMMLab's PyTorch-based toolbox and benchmark for image classification model pre-training, covering supervised, self-superv… | 23 | 3850 | active |
| Stability-AI/stable-audio-tools Stability AI's training and inference toolkit for conditional audio generation models, including Stable Audio Open. It supports training cu… | 73 | 3849 | active |
| Avatarify Avatarify is an open-source application that drives photorealistic avatars in real time for video-conferencing apps like Zoom and Skype, ba… | 23 | 16515 | maintenance |
| neuraloperator/neuraloperator A PyTorch library for learning neural operators, which map between function spaces rather than finite-dimensional vectors. It provides the … | 82 | 3836 | active |
| TransformerLensOrg/TransformerLens TransformerLens is a Python library for mechanistic interpretability of GPT-style transformer language models. It lets researchers load tho… | 99 | 3825 | active |
| google-research/scenic Scenic is a JAX-based library from Google Research focused on attention-based models for computer vision, providing shared lightweight libr… | 76 | 3821 | active |
| ddbourgin/numpy-ml numpy-ml is a collection of machine learning models and algorithms implemented exclusively in NumPy and the Python standard library, coveri… | 32 | 16330 | maintenance |