function: deep-learning
2653 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| andabi/deep-voice-conversion A TensorFlow implementation of deep neural networks for voice conversion (voice style transfer) that converts a source speaker's voice into… | 32 | 3938 | maintenance |
| chengtan9907/OpenSTL OpenSTL is a comprehensive benchmark and modular framework for spatio-temporal predictive learning, covering video prediction methods acros… | 54 | 1137 | active |
| facebookresearch/watermark-anything Official PyTorch implementation and pretrained models for the paper 'Watermark Anything with Localized Messages', which embeds multiple loc… | 10 | 1137 | active |
| xzf-thu/Mega-ASR Mega-ASR is a foundation automatic speech recognition model trained on 2.6M samples spanning 7 atomic acoustic conditions and 54 compound r… | 59 | 1136 | active |
| horseee/LLM-Pruner LLM-Pruner is a PyTorch library implementing structural pruning of large language models based on gradient information, as published at Neu… | 29 | 1136 | active |
| Dao-AILab/quack QuACK is a collection of high-performance GPU kernels (RMSNorm, LayerNorm, softmax, cross-entropy, GEMM with epilogues) written in NVIDIA's… | 85 | 1135 | active |
| Kiteretsu77/APISR APISR is a deep-learning based super-resolution tool that restores and enhances low-quality, low-resolution anime images and videos using t… | 37 | 1135 | active |
| rhymes-ai/Allegro Allegro is an open-source text-to-video generation model that produces high-quality 720p videos up to 6 seconds at 15 FPS from text prompts… | 24 | 1135 | active |
| Janspiry/Image-Super-Resolution-via-Iterative-Refinement An unofficial PyTorch implementation of SR3 (Image Super-Resolution via Iterative Refinement), a diffusion-based model for image super-reso… | 32 | 3923 | maintenance |
| zai-org/SCAIL-2 Official implementation of SCAIL-2, an open-source model for end-to-end controlled character animation that drives character videos from re… | 58 | 1132 | active |
| FlagOpen/RoboBrain2.5 RoboBrain 2.5 is an open-source embodied AI foundation model from BAAI that combines multimodal large language model capabilities with 3D s… | 50 | 1132 | active |
| sooftware/conformer An unofficial PyTorch implementation of the Conformer architecture (convolution-augmented Transformer) from the INTERSPEECH 2020 paper, tar… | 72 | 1131 | active |
| noahcao/OC_SORT OC-SORT is a pure motion-model-based multi-object tracker for video, improving on SORT by fixing Kalman filter limitations to handle occlus… | 67 | 1131 | stable |
| ikostrikov/pytorch-a2c-ppo-acktr-gail A PyTorch implementation of several deep reinforcement learning algorithms: A2C, PPO, ACKTR, and GAIL (imitation learning). It works with O… | 32 | 3903 | maintenance |
| Dobiasd/frugally-deep frugally-deep is a lightweight header-only C++ library for running inference (forward passes) on Keras/TensorFlow models without linking ag… | 85 | 1128 | active |
| BeingBeyond/Being-H Being-H is a family of human-centric embodied foundation models, including VLA models (Being-H0.5, Being-H0) and latent world-action models… | 63 | 1126 | active |
| OpenGVLab/VideoMamba VideoMamba is a state space model (Mamba-based) architecture for efficient video understanding, released with code and pretrained models fr… | 25 | 1125 | active |
| geopavlakos/hamer HaMeR (Hand Mesh Recovery) is a transformer-based model that reconstructs 3D hand meshes from single monocular images using the MANO parame… | 56 | 1124 | active |
| OpenDriveLab/UniVLA UniVLA is an open-source framework for training cross-embodiment vision-language-action (VLA) robot policies using task-centric latent acti… | 43 | 1124 | active |
| THUDM/SwissArmyTransformer SwissArmyTransformer (sat) is a PyTorch library for developing custom Transformer model variants where models like BERT, GPT, T5, GLM, and … | 23 | 1121 | active |
| FlagAI-Open/FlagAI FlagAI is a Python toolkit for training, fine-tuning, and deploying large-scale AI models across NLP, CV, and vision-language tasks. It int… | 64 | 3869 | maintenance |
| PaddlePaddle/PaddleHelix PaddleHelix is a bio-computing platform built on PaddlePaddle featuring large-scale representation learning and multi-task deep learning fo… | 57 | 1119 | active |
| alibaba-damo-academy/RynnVLA-002 RynnVLA-002 is a unified autoregressive Vision-Language-Action and world model that generates robot actions from text and image observation… | 43 | 1119 | active |
| caiyuanhao1998/MST A Python toolbox for spectral compressive imaging reconstruction that implements over 15 algorithms including MST, CST, DAUHST, BiSCI, HDNe… | 53 | 1118 | active |
| yangxue0827/RotationDetection AlphaRotate is a TensorFlow-based benchmark and toolbox for rotated (oriented) object detection, implementing detectors such as R2CNN, Reti… | 23 | 1118 | active |
| Lasagne/Lasagne Lasagne is a lightweight Python library for building and training neural networks on top of Theano. It supports feed-forward, convolutional… | 23 | 3857 | maintenance |
| HITsz-TMG/Uni-MoE Uni-MoE is a family of open-source Mixture-of-Experts (MoE) based omnimodal large language models that understand and generate across text,… | 68 | 1116 | active |
| NTMC-Community/MatchZoo MatchZoo is a Python toolkit for designing, comparing, and sharing deep text matching models. It provides a unified data pipeline, pre-buil… | 23 | 3849 | maintenance |
| yahoo/TensorFlowOnSpark TensorFlowOnSpark is a Python library that lets existing TensorFlow programs run distributed training and inference on Apache Spark and Had… | 23 | 3845 | maintenance |
| openai/improved-diffusion The official codebase for OpenAI's Improved Denoising Diffusion Probabilistic Models paper, providing a Python package for training and sam… | 32 | 3844 | maintenance |
| NVIDIA/earth2studio Earth2Studio is a Python deep-learning framework from NVIDIA for building, exploring, and deploying AI-driven weather and climate workflows… | 88 | 1112 | active |
| baofff/U-ViT U-ViT is the official PyTorch implementation of a ViT-based backbone architecture for diffusion models from the CVPR 2023 paper 'All are Wo… | 32 | 1110 | stable |
| unitreerobotics/unifolm-world-model-action UnifoLM-WMA-0 is Unitree's open-source world-model-action framework for general-purpose robot learning across multiple robotic embodiments.… | 50 | 1109 | active |
| princeton-vl/DPVO DPVO is a deep learning-based visual odometry and SLAM system that estimates camera trajectories from video or image sequences using patch-… | 32 | 1108 | active |
| THU-MIG/RepViT Official PyTorch implementation of RepViT, a family of lightweight CNNs designed by integrating efficient ViT architectural designs into Mo… | 19 | 1108 | stable |
| open-mmlab/PowerPaint PowerPaint is a versatile image inpainting model (ECCV 2024) built on diffusion models that handles text-guided object insertion, object re… | 70 | 1107 | active |
| OpenMOSS/MOVA MOVA is an open-source foundation model and toolkit for joint video-audio generation, synthesizing synchronized video and audio in a single… | 60 | 1106 | active |
| szymanowiczs/splatter-image Official PyTorch implementation of 'Splatter Image: Ultra-Fast Single-View 3D Reconstruction' (CVPR 2024), which uses an image-to-image net… | 26 | 1106 | active |
| google-research/multinerf Google Research's official code release for three NeRF papers: Mip-NeRF 360, Ref-NeRF, and RawNeRF, written in JAX. It trains neural radian… | 10 | 3808 | maintenance |
| TencentARC/T2I-Adapter Official implementation of T2I-Adapter, lightweight adapter models that add controllable conditioning (sketch, canny, lineart, depth, pose)… | 31 | 3801 | maintenance |
| PharMolix/OpenBioMed OpenBioMed is an open-source toolkit and agent platform for biomedicine and life science, offering multimodal models like BioMedGPT-Mol, 45… | 72 | 1103 | active |
| yangheng95/PyABSA PyABSA is a PyTorch-based library providing state-of-the-art models for aspect-based sentiment analysis, including aspect term extraction, … | 67 | 1102 | active |
| yzslab/gaussian-splatting-lightning A PyTorch Lightning implementation of 3D Gaussian Splatting with many derived algorithms (Mip-Splatting, LightGaussian, 2DGS, deformable Ga… | 58 | 1101 | active |
| zai-org/CogView4 CogView4, CogView3-Plus and CogView3 are open-source text-to-image generation models from Zhipu AI, with CogView4 being a 6B-parameter DiT-… | 28 | 1101 | active |
| naturomics/CapsNet-Tensorflow A TensorFlow implementation of CapsNet (Capsule Networks) based on Geoffrey Hinton's paper 'Dynamic Routing Between Capsules'. It supports … | 32 | 3786 | maintenance |
| yongliang-wu/DFT DFT (Dynamic Fine-Tuning) is the official implementation of an ICLR 2026 paper that improves Supervised Fine-Tuning of LLMs by dynamically … | 61 | 1100 | active |
| lucidrains/stylegan2-pytorch A simple PyTorch implementation of StyleGAN2, a state-of-the-art generative adversarial network, trainable entirely from the command line w… | 23 | 3783 | maintenance |
| yandex-research/tabm TabM is a PyTorch-based deep learning model for tabular data that efficiently imitates an ensemble of MLPs through parameter-efficient ense… | 36 | 1099 | active |
| LAMDA-CL/PyCIL PyCIL is a PyTorch-based Python toolbox for class-incremental learning, implementing the largest collection of CIL methods for reproducible… | 52 | 1098 | active |
| WangLibo1995/GeoSeg GeoSeg is an open-source PyTorch-based semantic segmentation toolbox focused on Vision Transformers for remote sensing imagery, featuring t… | 32 | 1096 | active |
| Eyeline-Labs/Go-with-the-Flow Official implementation of the CVPR 2025 Oral paper 'Go-with-the-Flow', which controls motion in video diffusion models by replacing i.i.d.… | 41 | 1093 | active |
| lizhe00/AnimatableGaussians Official PyTorch implementation of the CVPR 2024 paper 'Animatable Gaussians', which learns pose-dependent Gaussian maps for high-fidelity … | 27 | 1093 | active |
| lucidrains/tab-transformer-pytorch A PyTorch implementation of the TabTransformer architecture, an attention-based neural network for tabular data, also including the FT Tran… | 70 | 1091 | stable |
| yerfor/Real3DPortrait Official PyTorch implementation of Real3D-Portrait, an ICLR 2024 Spotlight paper for one-shot realistic 3D talking portrait synthesis. It g… | 26 | 1091 | active |
| Toni-SM/skrl skrl is an open-source modular Reinforcement Learning library written in Python, implemented in PyTorch, JAX, and NVIDIA Warp. It supports … | 81 | 1089 | active |
| mymusise/ChatGLM-Tuning A Python toolkit for fine-tuning the ChatGLM-6B large language model using LoRA (Low-Rank Adaptation) with the Alpaca dataset. It provides … | 10 | 3740 | maintenance |
| rhymes-ai/Aria Aria is an open multimodal native Mixture-of-Experts (MoE) model with 25.3B total parameters (3.9B activated per token) and a 64K multimoda… | 23 | 1087 | active |
| ndif-team/nnsight nnsight is a Python library for interpreting and intervening on the internals of deep learning models, built on PyTorch. It lets researcher… | 88 | 1085 | active |
| microsoft/TimeCraft TimeCraft is a diffusion model-based framework for generating high-quality synthetic time series data across domains. It uses learned seman… | 64 | 1085 | active |
| DSE-MSU/DeepRobust DeepRobust is a PyTorch library for adversarial robustness research, providing implementations of attack and defense methods for both image… | 45 | 1085 | active |
| mlc-ai/web-stable-diffusion A project that compiles and runs Stable Diffusion text-to-image models entirely inside web browsers using WebGPU and WebAssembly, with no s… | 30 | 3721 | maintenance |
| Alpha-VLLM/Lumina-mGPT-2.0 Lumina-mGPT 2.0 is a stand-alone decoder-only autoregressive model trained from scratch that unifies a broad range of image generation task… | 42 | 1084 | active |
| GraphSAGE Reference implementation of the GraphSAGE algorithm for inductive representation learning on large graphs using stochastic graph convolutio… | 32 | 3719 | maintenance |
| bytedance/byteps BytePS is a high-performance parameter server framework for distributed deep neural network training, supporting TensorFlow, Keras, PyTorch… | 10 | 3717 | maintenance |
| kimiyoung/transformer-xl Official implementation of Transformer-XL, an attention-based language model architecture that extends context beyond a fixed length via se… | 32 | 3714 | maintenance |
| agi-brain/xuance XuanCe is an open-source Python library of deep reinforcement learning (DRL) and multi-agent reinforcement learning (MARL) algorithm implem… | 95 | 1082 | active |
| NVlabs/Fast-dLLM NVIDIA's official implementation of Fast-dLLM, a family of training-free and fine-tuning-based acceleration techniques for diffusion-based … | 57 | 1082 | active |
| charlesq34/pointnet2 Official TensorFlow implementation of PointNet++, a deep neural network that learns hierarchical features on 3D point clouds using metric-s… | 32 | 3700 | maintenance |
| openai/glide-text2im Official codebase for GLIDE, a diffusion-based text-conditional image synthesis model from OpenAI. It provides pretrained models and notebo… | 10 | 3684 | maintenance |
| facebookresearch/hiera Hiera is the official PyTorch implementation of a hierarchical vision transformer from Meta AI (ICML 2023 Oral). It achieves state-of-the-a… | 20 | 1074 | active |
| cleardusk/3DDFA A PyTorch implementation of the TPAMI 2017 paper 'Face Alignment in Full Pose Range: A 3D Total Solution' (3DDFA). It fits a 3D Morphable M… | 23 | 3677 | maintenance |
| zju3dv/InfiniDepth InfiniDepth is a CVPR 2026 research library for monocular depth estimation that represents depth as neural implicit fields, allowing depth … | 53 | 1073 | active |
| open-gigaai/giga-train GigaTrain is an efficient and scalable Python training framework for large AI models, supporting distributed multi-GPU/multi-node execution… | 62 | 1072 | active |
| AILab-CVC/UniRepLKNet UniRepLKNet is a large-kernel ConvNet architecture (CVPR 2024, TPAMI 2025) that provides universal perception across image, audio, video, p… | 43 | 1072 | stable |
| NVIDIA-Merlin/HugeCTR HugeCTR is a GPU-accelerated deep learning framework from NVIDIA designed for training and inference of large recommender models, especiall… | 77 | 1071 | active |
| memoavatar/memo MEMO is an open-weight diffusion model for generating expressive, identity-consistent talking videos from a single reference image and an a… | 40 | 1070 | active |
| juntang-zhuang/Adabelief-Optimizer AdaBelief is a deep learning optimizer that adapts step sizes based on the 'belief' in observed gradients, combining Adam's fast convergenc… | 32 | 1070 | stable |
| yeates/PromptFix PromptFix is a PyTorch implementation of a diffusion-model-based image restoration model that follows natural language instructions to fix … | 24 | 1070 | active |
| RL-VIG/LibFewShot LibFewShot is a comprehensive PyTorch library for few-shot learning, implementing many fine-tuning, meta-learning, and metric-learning meth… | 54 | 1069 | active |
| princeton-nlp/SimCSE SimCSE is a Python library and research codebase implementing simple contrastive learning for sentence embeddings, with pre-trained unsuper… | 23 | 3654 | maintenance |
| guochengqian/PointNeXt PointNeXt is the official PyTorch implementation of the NeurIPS'22 paper that improves PointNet++ via better training and model scaling str… | 93 | 1067 | stable |
| hujie-frank/SENet Official Caffe/CUDA implementation of Squeeze-and-Excitation Networks (SENet), channel-attention building blocks for convolutional neural n… | 32 | 3646 | maintenance |
| gangweix/pixel-perfect-depth Pixel-Perfect Depth is a monocular depth estimation model based on pixel-space diffusion transformers that produces flying-pixel-free depth… | 49 | 1064 | active |
| lucidrains/mlp-mixer-pytorch A PyTorch implementation of Google AI's MLP-Mixer, an all-MLP architecture for image classification that uses neither convolutions nor atte… | 48 | 1064 | active |
| abertsch72/unlimiformer Unlimiformer is a method and official implementation for augmenting pretrained encoder-decoder transformers with retrieval-based attention,… | 30 | 1062 | stable |
| vijishmadhavan/ArtLine ArtLine is a deep learning project that converts portrait photos into line art portraits, with a ControlNet-based variant that adjusts styl… | 32 | 3630 | maintenance |
| NVIDIA/DreamDojo NVIDIA's official PyTorch codebase for DreamDojo, a generalist robot world model pretrained on 44k hours of human egocentric video and post… | 48 | 1059 | active |
| clovaai/stargan-v2 The official PyTorch implementation of StarGAN v2, a CVPR 2020 paper on diverse image-to-image translation across multiple domains using a … | 32 | 3617 | maintenance |
| mgsalem/Tensorflow-Project-Template A Python project template that provides a recommended folder structure and object-oriented skeleton (base model, base trainer, data loader,… | 32 | 3616 | maintenance |
| YunYang1994/tensorflow-yolov3 A TensorFlow 1.x implementation of the YOLOv3 real-time object detector, reproducing the 'YOLOv3: An Incremental Improvement' paper. It sup… | 23 | 3614 | maintenance |
| open-gigaai/giga-models GigaModels is an open-source Python framework providing pipelines for training, inference, deployment, and compression of multi-modal, gene… | 62 | 1057 | active |
| X-LANCE/SLAM-LLM SLAM-LLM is a deep learning toolkit for training custom multimodal large language models focused on speech, language, audio, and music proc… | 55 | 1056 | active |
| apache/singa Apache SINGA is a distributed deep learning platform for training neural networks across multiple devices and machines. It provides a C++ c… | 64 | 3606 | maintenance |
| yoyo-nb/Thin-Plate-Spline-Motion-Model The official PyTorch implementation of the CVPR 2022 paper 'Thin-Plate Spline Motion Model for Image Animation'. It animates a source image… | 32 | 3604 | maintenance |
| BAAI-DCAI/Bunny Bunny is a family of lightweight multimodal vision-language models that combine plug-and-play vision encoders (EVA-CLIP, SigLIP) with langu… | 26 | 1053 | active |
| Tencent-Hunyuan/HunyuanVideo-Foley HunyuanVideo-Foley is a multimodal diffusion model from Tencent Hunyuan that generates high-fidelity Foley sound effects synchronized with … | 37 | 1052 | active |
| showlab/MotionDirector MotionDirector is a research library for customizing text-to-video diffusion models to generate videos with desired motions from a small se… | 27 | 1050 | active |
| drprojects/superpoint_transformer Official PyTorch implementation of Superpoint Transformer (ICCV'23), SuperCluster (3DV'24), and EZ-SP (ICRA'26) for efficient semantic and … | 64 | 1049 | active |
| arcee-ai/DistillKit DistillKit is an open-source Python toolkit for knowledge distillation of large language models, supporting both online and offline distill… | 60 | 1047 | active |
| facebookresearch/pytorchvideo PyTorchVideo is a deep learning library from Facebook Research focused on video understanding research, built on PyTorch. It provides reusa… | 59 | 3566 | maintenance |