function: deep-learning
2653 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| PaddlePaddle/PaddleClas PaddleClas is a Python library and toolkit for image classification, recognition, and retrieval built on the PaddlePaddle deep learning fra… | 66 | 5838 | active |
| pjreddie/darknet Darknet is an open-source neural network framework written in C and CUDA, best known as the original home of the YOLO real-time object dete… | 32 | 26492 | maintenance |
| shibing624/MedicalGPT MedicalGPT is a Python training framework for building medical-domain large language models using the full ChatGPT-style training pipeline:… | 83 | 5748 | active |
| DeepLabCut/DeepLabCut DeepLabCut is an open-source Python toolbox for markerless 2D and 3D pose estimation of user-defined body parts using deep neural networks … | 89 | 5745 | stable |
| google/gemma_pytorch The official PyTorch implementation of Google's Gemma family of open large language models, including text-only and multimodal variants. It… | 10 | 5719 | active |
| mlpack/mlpack mlpack is a fast, header-only C++ machine learning library built on Armadillo, ensmallen, and cereal, offering a wide range of algorithms f… | 87 | 5703 | stable |
| open-mmlab/OpenPCDet OpenPCDet is a PyTorch-based open-source toolbox for LiDAR-based 3D object detection. It provides official implementations of models like P… | 53 | 5692 | active |
| apple/ml-depth-pro Depth Pro is Apple's reference implementation of a foundation model for zero-shot metric monocular depth estimation, producing sharp high-r… | 30 | 5683 | active |
| huggingface/alignment-handbook A collection of robust training recipes and scripts from Hugging Face for aligning large language models with human and AI preferences, cov… | 66 | 5671 | active |
| OpenSenseNova/SenseNova-U1 SenseNova-U is a series of open-weight unified multimodal models (e.g., SenseNova-U1.5-8B-MoT) built on the NEO-unify architecture that com… | 59 | 5668 | active |
| pytorch/torchtitan torchtitan is a PyTorch-native platform for large-scale training of generative AI models, offering a clean-room implementation of PyTorch's… | 79 | 5667 | active |
| Fanghua-Yu/SUPIR SUPIR is a Python-based photo-realistic image restoration system built on SDXL diffusion priors and LLaVA captioning, presented at CVPR 202… | 36 | 5649 | active |
| Vision-CAIR/MiniGPT-4 Official code for MiniGPT-4 and MiniGPT-v2, vision-language models that align a frozen visual encoder with a frozen LLM (Vicuna) via a sing… | 30 | 25627 | maintenance |
| mosaicml/composer Composer is an open-source PyTorch-based deep learning training library by MosaicML (now Databricks) for training neural networks faster an… | 65 | 5495 | active |
| LaurentMazare/tch-rs tch-rs is a Rust crate providing thin bindings to the C++ API of PyTorch (libtorch), staying close to the original API. It enables tensor o… | 67 | 5479 | active |
| tensorflow/rust TensorFlow Rust provides idiomatic Rust language bindings for TensorFlow via its C API. It lets Rust programs build and run TensorFlow comp… | 10 | 5476 | active |
| karpathy/minGPT A minimal, clean PyTorch re-implementation of OpenAI's GPT covering both training and inference in roughly 300 lines of code. It is designe… | 32 | 24840 | maintenance |
| xxlong0/Wonder3D Wonder3D is a cross-domain diffusion model that reconstructs high-fidelity textured 3D meshes from a single image in 2-3 minutes. It genera… | 32 | 5425 | active |
| isl-org/MiDaS MiDaS is a Python library with pretrained models for robust monocular depth estimation from a single image, based on the TPAMI 2022 paper a… | 10 | 5420 | stable |
| deepseek-ai/DeepSeek-VL2 DeepSeek-VL2 is a series of Mixture-of-Experts vision-language models (Tiny, Small, and 4.5B activated parameters) with inference code and … | 25 | 5374 | active |
| microsoft/SynapseML SynapseML (formerly MMLSpark) is an open-source machine learning library built on Apache Spark that provides simple, composable, distribute… | 88 | 5240 | active |
| awslabs/gluonts GluonTS is a Python library for probabilistic time series modeling, focused on deep learning based forecasting models built on PyTorch. It … | 88 | 5227 | stable |
| wenet-e2e/wenet WeNet is a production-first, end-to-end automatic speech recognition (ASR) toolkit built on PyTorch with transformer/conformer models. It p… | 62 | 5227 | active |
| InternLM/xtuner XTuner is an open-source LLM training engine from InternLM designed for fine-tuning ultra-large-scale Mixture-of-Experts (MoE) models, with… | 67 | 5183 | active |
| NVIDIAGameWorks/kaolin Kaolin is NVIDIA's PyTorch library of GPU-optimized modules for 3D deep learning research, covering meshes, point clouds, and 3D Gaussian s… | 68 | 5161 | active |
| yisol/IDM-VTON Official implementation of IDM-VTON, an ECCV 2024 paper that improves diffusion models for high-fidelity virtual try-on, swapping garments … | 30 | 5156 | active |
| open-mmlab/mmaction2 MMAction2 is OpenMMLab's PyTorch-based toolbox and benchmark for video understanding, covering action recognition, temporal action localiza… | 55 | 5142 | active |
| ai-dawang/PlugNPlay-Modules A curated collection of plug-and-play deep learning modules (convolutions, attention mechanisms, downsampling, and feature fusion blocks) i… | 38 | 5105 | active |
| facebookresearch/co-tracker CoTracker is a transformer-based model from Meta AI and Oxford VGG that jointly tracks any point (pixel) across a video, handling occlusion… | 60 | 5080 | active |
| AILab-CVC/VideoCrafter VideoCrafter is an open-source video generation and editing toolbox built on video diffusion models, offering Text-to-Video and Image-to-Vi… | 48 | 5073 | active |
| Deci-AI/super-gradients SuperGradients is an open-source PyTorch-based training library for building, training, and fine-tuning state-of-the-art computer vision mo… | 54 | 5052 | active |
| deepseek-ai/DeepSeek-V2 DeepSeek-V2 is a strong, economical Mixture-of-Experts language model released with open weights, inference code, and evaluation tooling. T… | 24 | 5039 | active |
| lightvector/KataGo KataGo is an open-source Go (baduk) engine trained via AlphaZero-like self-play, one of the strongest Go bots available. It runs as a GTP e… | 99 | 5036 | active |
| sktime/pytorch-forecasting A PyTorch-based Python library for time series forecasting with state-of-the-art deep learning architectures. It provides a high-level API … | 92 | 4977 | active |
| pytorch/executorch ExecuTorch is PyTorch's framework for exporting and running AI models on-device across mobile, embedded, and edge hardware, with a tiny (~5… | 95 | 4953 | active |
| microsoft/muzic Muzic is a Microsoft Research project providing deep learning models for music understanding and generation, including MusicBERT, SongMASS,… | 66 | 4952 | active |
| KaiyangZhou/deep-person-reid Torchreid is a PyTorch library for deep-learning person re-identification, supporting both image and video reid with end-to-end training an… | 50 | 4900 | stable |
| google-deepmind/alphageometry Google DeepMind's implementation of AlphaGeometry and DDAR, AI systems that solve Olympiad-level geometry theorem proving problems. It comb… | 55 | 4879 | active |
| tyxsspa/AnyText AnyText is the official implementation of a diffusion-based model for multilingual visual text generation and editing in images, accepted a… | 32 | 4874 | active |
| deepjavalibrary/djl Deep Java Library (DJL) is an engine-agnostic, high-level deep learning framework for Java that supports backends like PyTorch, TensorFlow,… | 80 | 4842 | active |
| ace-step/ACE-Step ACE-Step is an open-source foundation model for music generation that combines diffusion-based generation with a deep compression autoencod… | 49 | 4788 | active |
| pytorch/ignite PyTorch-Ignite is a high-level library for training and evaluating neural networks in PyTorch flexibly and transparently. It provides an En… | 89 | 4778 | stable |
| open-mmlab/mmocr MMOCR is OpenMMLab's PyTorch-based toolbox for text detection, recognition, and key information extraction. It provides a model zoo of OCR … | 23 | 4752 | active |
| FluxML/Flux.jl Flux.jl is a machine learning library written entirely in Julia, providing lightweight abstractions over Julia's native GPU support and aut… | 98 | 4739 | stable |
| OpenDriveLab/UniAD UniAD is a unified end-to-end autonomous driving framework that hierarchically casts perception, prediction, and planning tasks under a pla… | 44 | 4737 | active |
| cvg/LightGlue LightGlue is a deep neural network library that matches sparse local features across image pairs with high accuracy and fast inference. It … | 50 | 4728 | stable |
| facebookresearch/flow_matching A PyTorch library for implementing flow matching algorithms with continuous, discrete, and Riemannian flow matching implementations. It acc… | 48 | 4705 | active |
| mindspore-ai/mindspore MindSpore is an open-source deep learning framework for training and inference across mobile, edge, and cloud scenarios. It provides automa… | 32 | 4700 | active |
| PKU-Alignment/align-anything Align-Anything is a modular Python framework for aligning any-to-any (all-modality) large models with human intentions and values using fee… | 48 | 4667 | active |
| Blealtan/efficient-kan An efficient pure-PyTorch implementation of Kolmogorov-Arnold Networks (KAN) that reformulates the computation as matrix multiplications ov… | 24 | 4655 | active |
| Tencent/TNN TNN is a high-performance, lightweight deep learning inference framework developed by Tencent Youtu Lab, supporting mobile, desktop, and se… | 32 | 4648 | active |
| OpenNMT/CTranslate2 CTranslate2 is a C++ and Python library for fast, memory-efficient inference of Transformer models on CPU and GPU. It uses quantization, la… | 96 | 4644 | stable |
| WhisperSpeech/WhisperSpeech WhisperSpeech is an open-source text-to-speech system built by inverting OpenAI's Whisper model, aiming to be 'Stable Diffusion for speech'… | 56 | 4639 | active |
| Rikorose/DeepFilterNet DeepFilterNet is a low-complexity speech enhancement framework that performs real-time noise suppression on full-band 48kHz audio using dee… | 23 | 4632 | active |
| Kwai-Kolors/Kolors Kolors is a large-scale latent diffusion model for photorealistic text-to-image synthesis, trained with bilingual (Chinese and English) tex… | 23 | 4615 | active |
| sensity-ai/dot dot (Deepfake Offensive Toolkit) is a Python tool that generates real-time, controllable deepfakes from a webcam feed and injects them into… | 23 | 4586 | active |
| joanrod/star-vector StarVector is a foundation model that generates scalable vector graphics (SVG) code from images and text by treating vectorization as a cod… | 49 | 4560 | active |
| RecBole RecBole is a unified, comprehensive and efficient recommendation library built on Python and PyTorch for reproducing and developing recomme… | 26 | 4541 | stable |
| Tencent-Hunyuan/HunyuanVideo-1.5 HunyuanVideo-1.5 is Tencent's lightweight 8.3B-parameter video generation model supporting text-to-video and image-to-video synthesis. The … | 50 | 4534 | active |
| OAID/Tengine Tengine is a lightweight, high-performance, modular deep learning inference engine developed by OPEN AI LAB for embedded and edge devices. … | 27 | 4531 | active |
| NVlabs/tiny-cuda-nn A small, self-contained C++/CUDA framework for training and querying neural networks, featuring a lightning-fast fully fused MLP and a vers… | 58 | 4528 | active |
| facebookresearch/vjepa2 Official PyTorch codebase and pretrained models for V-JEPA 2, a self-supervised video encoder trained on internet-scale video, plus V-JEPA … | 52 | 4527 | active |
| TencentARC/InstantMesh InstantMesh is a feed-forward framework for generating 3D meshes from a single image using sparse-view large reconstruction models (LRM/Ins… | 25 | 4509 | active |
| real-stanford/diffusion_policy Official PyTorch implementation of Diffusion Policy, a visuomotor robot policy learning method that represents robot behavior as a conditio… | 30 | 4491 | stable |
| zjunlp/DeepKE DeepKE is an open-source PyTorch-based knowledge extraction toolkit for knowledge graph construction, supporting named entity recognition, … | 64 | 4472 | active |
| layumi/Person_reID_baseline_pytorch A small, friendly PyTorch baseline implementation for person and vehicle re-identification (ReID). It reproduces strong top-conference resu… | 65 | 4446 | stable |
| mosaicml/llm-foundry LLM Foundry is a PyTorch-based codebase for training, finetuning, evaluating, and deploying large language models from 125M to 70B+ paramet… | 65 | 4441 | active |
| xlite-dev/lite.ai.toolkit A lightweight C++ toolkit providing unified APIs for 100+ pre-trained AI models across inference backends like ONNX Runtime, MNN, TensorRT,… | 74 | 4427 | active |
| huawei-noah/Efficient-AI-Backbones A collection of efficient neural network backbone architectures (GhostNet, TNT, ViG, WaveMLP, TinyNet, etc.) from Huawei Noah's Ark Lab, wi… | 28 | 4418 | active |
| iperov/DeepFaceLab DeepFaceLab is the leading open-source Windows application for creating deepfakes, allowing users to swap, de-age, or replace faces and hea… | 10 | 19292 | maintenance |
| lululxvi/deepxde DeepXDE is a Python library for scientific machine learning and physics-informed learning, built on top of PyTorch, TensorFlow, JAX, and Pa… | 81 | 4386 | stable |
| dreamgaussian/dreamgaussian DreamGaussian is the official PyTorch implementation of an ICLR 2024 Oral paper for efficient 3D content creation using generative Gaussian… | 18 | 4352 | active |
| lucas-maes/le-wm LeWorldModel (LeWM) is the official PyTorch codebase for a JEPA-based world model that trains stably end-to-end from raw pixels using only … | 52 | 4344 | active |
| VectorSpaceLab/OmniGen OmniGen is a unified diffusion-based image generation model that produces and edits images from multi-modal prompts without auxiliary modul… | 47 | 4340 | active |
| ourownstory/neural_prophet NeuralProphet is a Python library for interpretable time series forecasting built on PyTorch, combining neural networks with traditional ti… | 23 | 4295 | active |
| Tencent-Hunyuan/HunyuanDiT Hunyuan-DiT is Tencent's open-source diffusion transformer model for text-to-image generation with fine-grained Chinese language understand… | 49 | 4291 | active |
| richzhang/PerceptualSimilarity A PyTorch library implementing the LPIPS (Learned Perceptual Image Patch Similarity) metric, which measures perceptual distance between ima… | 23 | 4269 | stable |
| Nixtla/neuralforecast NeuralForecast is a Python library offering a large collection of state-of-the-art neural forecasting models (NBEATS, NHITS, TFT, PatchTST,… | 98 | 4257 | active |
| ThilinaRajapakse/simpletransformers Simple Transformers is a Python library built on Hugging Face Transformers that lets users train, fine-tune, and evaluate Transformer model… | 61 | 4254 | active |
| ali-vilab/AnyDoor AnyDoor is the official implementation of a diffusion-based model that teleports target objects into new scenes at user-specified locations… | 28 | 4238 | active |
| jwohlwend/boltz Boltz is a family of open-source biomolecular interaction models (Boltz-1 and Boltz-2) that predict complex structures and binding affiniti… | 63 | 4177 | active |
| lllyasviel/style2paints Style2Paints is an AI-driven tool that colorizes lineart sketches, optionally guided by human hints, style reference images, and lighting. … | 32 | 18179 | maintenance |
| facebookresearch/vggt-omega VGGT-Omega is a research library from Oxford VGG and Meta AI providing pretrained transformer models for 3D vision tasks such as camera pos… | 58 | 4165 | active |
| torchgeo/torchgeo TorchGeo is a PyTorch domain library, similar to torchvision, providing datasets, samplers, transforms, and pre-trained models specific to … | 95 | 4159 | active |
| ArcInstitute/evo2 Evo 2 is a DNA foundation language model (1B-40B parameters) that models genomes at single-nucleotide resolution with up to 1 million base … | 62 | 4157 | active |
| VectorSpaceLab/OmniGen2 OmniGen2 is an open-source unified multimodal generation model supporting text-to-image generation, instruction-guided image editing, and i… | 52 | 4112 | active |
| huggingface/distil-whisper Distil-Whisper is a distilled version of OpenAI's Whisper model for English speech recognition, offering 6x faster inference, 49% fewer par… | 27 | 4112 | active |
| facebookresearch/jepa Official PyTorch implementation of V-JEPA, a self-supervised method for learning visual representations from video using a joint-embedding … | 29 | 4105 | active |
| GuyTevet/motion-diffusion-model Official PyTorch implementation of the Human Motion Diffusion Model (MDM) paper, generating 3D human motion sequences from text prompts usi… | 52 | 4092 | active |
| princeton-vl/RAFT Official PyTorch implementation of RAFT (Recurrent All Pairs Field Transforms for Optical Flow), an ECCV 2020 model for estimating dense op… | 49 | 4091 | stable |
| PaddlePaddle/PaddleRec PaddleRec is a large-scale recommendation algorithm library built on PaddlePaddle, containing classic and state-of-the-art recommendation m… | 29 | 4085 | active |
| hao-ai-lab/FastVideo FastVideo is a unified Python framework for post-training and real-time inference of video diffusion models, covering data preprocessing, f… | 80 | 4076 | active |
| lllyasviel/Paints-UNDO Paints-UNDO is a family of deep learning models that take an image as input and generate the step-by-step drawing sequence (sketching, inki… | 40 | 4067 | active |
| facebookresearch/dlrm A PyTorch-based implementation of the Deep Learning Recommendation Model (DLRM) from Facebook Research, which processes dense and sparse fe… | 60 | 4065 | stable |
| open-spaced-repetition/fsrs4anki FSRS4Anki is a modern spaced-repetition scheduler for Anki based on the Free Spaced Repetition Scheduler (FSRS) algorithm. It replaces Anki… | 80 | 4053 | active |
| limix-ldm-ai/LimiX LimiX is the first large structured-data foundation model (LDM), a transformer-based model for tabular data that handles classification, re… | 60 | 4045 | active |
| uxlfoundation/oneDNN oneDNN is an open-source cross-platform performance library providing optimized building blocks (primitives) for deep learning applications… | 99 | 4042 | stable |
| tensorflow/tensor2tensor Tensor2Tensor (T2T) is a Python library of deep learning models and datasets built on TensorFlow, developed by the Google Brain team to mak… | 10 | 17464 | maintenance |
| Nixtla/nixtla Nixtla's TimeGPT is a production-ready pre-trained foundation model for time series forecasting and anomaly detection, accessed via a Pytho… | 97 | 3996 | active |
| lucidrains/vector-quantize-pytorch A PyTorch library implementing vector and scalar quantization, including Residual VQ and techniques like DiVeQ codebook updates. It origina… | 83 | 3996 | active |