function: deep-learning
2653 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| nv-tlabs/PiD PiD is a plug-and-play pixel diffusion decoder from NVIDIA that replaces VAE/RAE decoders, decoding latent representations directly into hi… | 56 | 1045 | active |
| liuyuan-pal/SyncDreamer SyncDreamer is a synchronized multiview diffusion model that generates multiview-consistent images from a single-view image, released with … | 50 | 1045 | active |
| k2-fsa/ZipVoice ZipVoice is a series of fast, high-quality zero-shot text-to-speech models based on flow matching, with a compact 123M-parameter Zipformer-… | 43 | 1045 | active |
| 3DTopia/3DTopia-XL 3DTopia-XL is a 3D diffusion transformer model that generates high-quality 3D assets with PBR materials from a single image or text prompt … | 37 | 1045 | active |
| Shawn1993/cnn-text-classification-pytorch A PyTorch implementation of Kim's CNN architecture for sentence classification, reproducing results from the paper 'Convolutional Neural Ne… | 65 | 1043 | active |
| aigc3d/LAM LAM is a PyTorch implementation of a Large Avatar Model that reconstructs an animatable 3D Gaussian head from a single image in one forward… | 58 | 1043 | active |
| ZhengYinan-AIR/Diffusion-Planner Official PyTorch implementation of Diffusion Planner, an ICLR 2025 Oral paper using a DiT-based diffusion model for closed-loop autonomous … | 53 | 1043 | active |
| zai-org/SCAIL SCAIL is the official inference implementation of a 14B diffusion transformer model that generates studio-grade character animation videos … | 52 | 1043 | active |
| jiachenzhu/DyT Official PyTorch implementation of DynamicTanh (DyT), a learnable element-wise tanh operation that replaces normalization layers in Transfo… | 26 | 1043 | active |
| Xiaoqi-Zhao-DLUT/MSNet-M2SNet Official PyTorch implementations of MSNet and M2SNet, multi-scale subtraction networks for medical image segmentation such as polyp, lung i… | 74 | 1042 | active |
| HannesStark/boltzgen BoltzGen is an open-source all-atom generative diffusion model for designing protein and peptide binders against arbitrary biomolecular tar… | 71 | 1042 | active |
| zju3dv/EfficientLoFTR Efficient LoFTR is a PyTorch implementation of a semi-dense local feature matching model that matches keypoints between image pairs with sp… | 40 | 1042 | active |
| PetarV-/GAT Reference implementation of Graph Attention Networks (GAT), the ICLR 2018 model by Veličković et al., written in TensorFlow 1.x with a mini… | 32 | 3548 | maintenance |
| foolwood/SiamMask Official PyTorch implementation of SiamMask, a deep learning framework for fast online visual object tracking and video object segmentation… | 35 | 3547 | maintenance |
| tonbistudio/turboquant-pytorch A from-scratch PyTorch implementation of Google's TurboQuant (ICLR 2026) algorithm for compressing LLM key-value caches, including an impro… | 50 | 1040 | active |
| tensorflow/minigo Minigo is an open-source, minimalist implementation of the AlphaGo Zero algorithm for the game of Go, built on TensorFlow. It provides a re… | 10 | 3542 | maintenance |
| TorchEnsemble-Community/Ensemble-Pytorch A unified ensemble learning framework for PyTorch that implements strategies like voting, bagging, and gradient boosting to improve deep le… | 23 | 1037 | active |
| pytorch/tensordict TensorDict is a PyTorch library providing a batched, nested dict-like tensor container where operations like slicing, stacking, and arithme… | 99 | 1036 | active |
| ustcllm/RecFM RecFM is a collection of tools and frameworks from USTCLLM for developing foundation models tailored to recommendation systems. It bundles … | 40 | 1033 | active |
| podgorskiy/ALAE Official PyTorch implementation of Adversarial Latent Autoencoders (ALAE/StyleALAE), a CVPR 2020 paper combining autoencoders with GAN trai… | 32 | 3511 | maintenance |
| awentzonline/image-analogies A Python library implementing neural image analogies using VGG16 feature maps with PatchMatch-based matching and blending, based on the 'Im… | 23 | 3502 | maintenance |
| rl-tools/rl-tools RLtools is a pure C++ header-only, dependency-free deep reinforcement learning library supporting algorithms like SAC, TD3, and PPO. It com… | 65 | 1028 | active |
| EchoMimic EchoMimic is a series of open-source models (V1-V3) from Ant Group for audio-driven human animation, generating lifelike talking-head, port… | 50 | 1027 | active |
| soCzech/TransNetV2 TransNet V2 is a deep neural network for shot boundary detection in videos, achieving state-of-the-art results on benchmarks like ClipShots… | 32 | 1027 | stable |
| JackAILab/ConsistentID ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using … | 52 | 1026 | active |
| DingXiaoH/RepVGG RepVGG is a PyTorch implementation of the VGG-style ConvNet architecture from the CVPR 2021 paper, achieving over 84% top-1 ImageNet accura… | 32 | 3478 | maintenance |
| aim-uofa/AdelaiDet AdelaiDet is an open-source Python toolbox built on Detectron2 that implements multiple instance-level detection and recognition algorithms… | 32 | 3478 | maintenance |
| open-mmlab/mmyolo MMYOLO is the OpenMMLab toolbox and benchmark for the YOLO series of object detection models, implemented on PyTorch. It provides unified i… | 23 | 3468 | maintenance |
| TencentARC/SEED-Voken SEED-Voken is a collection of visual tokenizers (Open-MAGVIT2 and IBQ) that convert images and videos into discrete tokens for autoregressi… | 48 | 1021 | active |
| JiahuiYu/generative_inpainting An open-source implementation of DeepFill v1/v2 generative image inpainting models, featuring Contextual Attention (CVPR 2018) and Gated Co… | 32 | 3466 | maintenance |
| tensorlayer/SRGAN Reference implementation of SRGAN, a generative adversarial network for photo-realistic single image super-resolution, built on TensorLayer… | 23 | 3466 | maintenance |
| fla-org/native-sparse-attention Efficient Triton kernel implementations of Native Sparse Attention (NSA), a hardware-aligned, natively trainable sparse attention mechanism… | 49 | 1020 | active |
| richzhang/colorization A Python library implementing automatic colorization of grayscale photos using deep neural networks from the ECCV 2016 'Colorful Image Colo… | 32 | 3461 | maintenance |
| zai-org/GLM-Image GLM-Image is an open-source image generation model combining a 9B autoregressive generator with a 7B diffusion decoder, excelling at text r… | 48 | 1019 | active |
| bowang-lab/U-Mamba U-Mamba is a hybrid CNN-state-space-model (Mamba) network for biomedical image segmentation, built on top of the nnU-Net framework. It comb… | 26 | 1019 | active |
| google/prompt-to-prompt Google's official implementation of the Prompt-to-Prompt paper, which enables text-driven image editing in Latent Diffusion and Stable Diff… | 10 | 3456 | maintenance |
| facebookresearch/PyTorch-BigGraph PyTorch-BigGraph is a distributed system for learning embeddings of very large graph-structured data, scaling to billions of entities and t… | 10 | 3454 | maintenance |
| tensorflow/adanet AdaNet is a lightweight TensorFlow-based AutoML framework that automatically learns high-quality neural network architectures and ensembles… | 10 | 3454 | maintenance |
| HobbitLong/SupContrast A PyTorch reference implementation of the Supervised Contrastive Learning paper (SupCon loss) that also supports SimCLR when labels are omi… | 32 | 3449 | maintenance |
| ShangtongZhang/DeepRL A modularized PyTorch implementation of popular deep reinforcement learning algorithms including DQN variants, PPO, DDPG, TD3, A2C, and Opt… | 23 | 3448 | maintenance |
| Alpha-VLLM/Lumina-DiMOO Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outpu… | 55 | 1015 | active |
| MeshAnything MeshAnything is an autoregressive transformer model that generates artist-created 3D meshes (up to 1600 faces in V2) aligned with a given s… | 31 | 1015 | active |
| fallenshock/FlowEdit Official PyTorch implementation of FlowEdit, an ICCV 2025 method for inversion-free, text-based editing of real images using pre-trained fl… | 66 | 1014 | active |
| clab/dynet DyNet is a C++ neural network toolkit with Python bindings, designed for efficient CPU/GPU training of networks with dynamic per-instance s… | 23 | 3436 | maintenance |
| EvolvingLMMs-Lab/Otter Otter is a multi-modal vision-language model built on OpenFlamingo, instruction-tuned on the MIMIC-IT dataset with image and video understa… | 21 | 3436 | maintenance |
| Soul-AILab/SoulX-FlashHead SoulX-FlashHead is a 1.3B-parameter framework for high-fidelity, infinite-length, real-time streaming talking-head portrait video generatio… | 53 | 1011 | active |
| microsoft/aurora Aurora is Microsoft's implementation of a deep learning foundation model for Earth system forecasting, predicting atmospheric variables lik… | 89 | 1010 | active |
| TensorSpeech/TensorFlowASR TensorFlowASR is a Python library implementing automatic speech recognition architectures such as DeepSpeech2, Jasper, RNN Transducer, Cont… | 66 | 1010 | active |
| thuml/Large-Time-Series-Model Official code, datasets, and checkpoints for Timer and Sundial, generative pre-trained Transformer foundation models for general time serie… | 58 | 1010 | active |
| wang-rui/phishguard-scaffold PhishGuard is a Python research framework that jointly performs phishing detection and dissemination control on social media using LLaMA-ba… | 47 | 1009 | active |
| TRI-ML/prismatic-vlms Prismatic VLMs is a PyTorch-based codebase for training visually-conditioned language models (VLMs) with flexible vision backbones like CLI… | 25 | 1009 | active |
| waifu2x (nunif) waifu2x is an image super-resolution and noise-reduction tool for anime-style art (and photos) using deep convolutional neural networks, or… | 78 | 3418 | maintenance |
| shaoanlu/faceswap-GAN A Jupyter Notebook-based implementation of face swapping using a denoising autoencoder architecture enhanced with adversarial losses, VGGFa… | 32 | 3416 | maintenance |
| graspnet/graspnet-baseline The official baseline deep learning model for the GraspNet-1Billion benchmark, detecting dense 6-DoF grasp poses from point clouds of clutt… | 35 | 1008 | stable |
| MeiGen-AI/PosterCraft PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight… | 48 | 1007 | active |
| meetps/pytorch-semseg A PyTorch library implementing popular semantic segmentation architectures such as FCN, U-Net, SegNet, PSPNet, ICNet, FRRN, and LinkNet, wi… | 23 | 3402 | maintenance |
| RIFE RIFE is a deep learning model for real-time video frame interpolation, estimating intermediate flow between frames to generate smooth slow-… | 77 | 1004 | active |
| huawei-noah/noah-research A collection of research code subprojects released by Huawei Noah's Ark Lab, each in its own directory. It is not an official Huawei produc… | 76 | 1004 | active |
| TinyLLaVA/TinyLLaVA_Factory TinyLLaVA Factory is an open-source modular PyTorch/HuggingFace codebase for training small-scale large multimodal models (LMMs) that combi… | 68 | 1004 | active |
| SciSharp/TensorFlow.NET TensorFlow.NET provides .NET Standard bindings for Google's TensorFlow, aiming to implement the complete TensorFlow API in C# and F#. It in… | 24 | 3398 | maintenance |
| ZJUI-AI4H/Hulu-Med Hulu-Med is a family of open-source transparent generalist medical vision-language models ranging from 4B to 235B parameters, covering text… | 62 | 1001 | active |
| pjlab-sys4nlp/llama-moe LLaMA-MoE is a Python toolkit and model series for building Mixture-of-Experts (MoE) language models from dense LLaMA models via expert con… | 19 | 1001 | active |
| catalyst-team/catalyst Catalyst is a high-level PyTorch framework for deep learning research and development, focused on reproducibility, rapid experimentation, a… | 64 | 3382 | maintenance |
| eladrich/pixel2style2pixel Official PyTorch implementation of pixel2style2pixel (pSp), a StyleGAN encoder from CVPR 2021 that maps real images directly into the W+ la… | 32 | 3350 | maintenance |
| shelhamer/fcn.berkeleyvision.org Reference implementation of Fully Convolutional Networks (FCN) for semantic segmentation from the CVPR 2015 / PAMI 2016 papers, built on Ca… | 32 | 3350 | maintenance |
| tianzhi0549/FCOS Official PyTorch implementation of FCOS, a fully convolutional one-stage, anchor-free object detector published at ICCV 2019. It provides t… | 32 | 3345 | maintenance |
| tamarott/SinGAN Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat… | 32 | 3344 | maintenance |
| NVlabs/eg3d Official PyTorch implementation of EG3D, an efficient geometry-aware 3D generative adversarial network from NVIDIA Research (CVPR 2022). It… | 32 | 3338 | maintenance |
| HRNet/HRNet-Semantic-Segmentation Official PyTorch implementation of HRNet (High-Resolution Network) and the Segmentation Transformer (OCR) approach for semantic segmentatio… | 32 | 3331 | maintenance |
| run-youngjoo/SC-FEGAN SC-FEGAN is a GUI application that uses a generative adversarial network (SN-patchGAN discriminator with a U-Net generator) to edit face im… | 32 | 3329 | maintenance |
| pytorch-yolo-v3 A minimal PyTorch implementation of the YOLO v3 object detection algorithm, supporting detection on images and video with configurable reso… | 32 | 3312 | maintenance |
| PixArt-alpha/PixArt-alpha PixArt-α is a Transformer-based text-to-image diffusion model with PyTorch model definitions, pre-trained weights, and inference/training c… | 27 | 3304 | maintenance |
| open-mmlab/mmselfsup OpenMMLab's PyTorch-based toolbox and benchmark for self-supervised and unsupervised visual representation learning. It provides implementa… | 23 | 3302 | maintenance |
| facebookresearch/vissl VISSL is Facebook AI Research's extensible, modular and scalable PyTorch library for state-of-the-art self-supervised learning with images.… | 10 | 3293 | maintenance |
| NVIDIA/flownet2-pytorch A PyTorch implementation of FlowNet 2.0 for deep-learning-based optical flow estimation, released by NVIDIA. It provides multiple network a… | 66 | 3289 | maintenance |
| google-deepmind/dm-haiku Haiku is a simple neural network library for JAX from Google DeepMind, providing an object-oriented module abstraction (hk.Module) and a fu… | 89 | 3275 | maintenance |
| zhanghang1989/ResNeSt ResNeSt is a PyTorch implementation of the Split-Attention Network, a ResNet variant that applies channel-wise attention across network bra… | 23 | 3261 | maintenance |
| martinarjovsky/WassersteinGAN Reference PyTorch implementation of the Wasserstein GAN paper, providing training scripts for DCGAN and MLP architectures on datasets like … | 32 | 3243 | maintenance |
| google/model_search Model Search is a Google AutoML framework that implements neural architecture search algorithms at scale to find optimal DNN architectures … | 10 | 3238 | maintenance |
| tinyvision/DAMO-YOLO DAMO-YOLO is a fast and accurate object detection framework built on PyTorch, featuring NAS-searched backbones, RepGFPN, a lightweight Zero… | 32 | 3183 | maintenance |
| yusugomori/DeepLearning A multi-language (Python, C, C++, Java, Scala, Go) educational implementation of classic deep learning algorithms such as Deep Belief Nets,… | 32 | 3177 | maintenance |
| adambielski/siamese-triplet A PyTorch library implementing siamese and triplet networks for learning image embeddings, with online pair/triplet mining strategies. It p… | 32 | 3174 | maintenance |
| jettify/pytorch-optimizer torch-optimizer is a Python library providing a collection of additional optimization algorithms for PyTorch, drop-in compatible with the t… | 23 | 3171 | maintenance |
| farizrahman4u/seq2seq A sequence-to-sequence learning add-on library for Keras, providing modular encoder-decoder layers and ready-made Seq2Seq models. It suppor… | 32 | 3169 | maintenance |
| migueldeicaza/TensorFlowSharp TensorFlowSharp provides .NET bindings to the TensorFlow C API, exposing a strongly-typed low-level API for C# and F#. It is designed mainl… | 10 | 3147 | maintenance |
| FudanNLP/fastNLP fastNLP is a lightweight, modularized and extensible NLP framework in Python that reduces engineering boilerplate such as data processing l… | 23 | 3141 | maintenance |
| DAMO-NLP-SG/Video-LLaMA Video-LLaMA is an instruction-tuned audio-visual language model that extends LLaMA with video and audio understanding via cross-modal pretr… | 29 | 3139 | maintenance |
| microsoft/torchscale A PyTorch library from Microsoft implementing foundation Transformer architectures such as DeepNet, Magneto, RetNet, LongNet, BitNet, and X… | 32 | 3138 | maintenance |
| open-mmlab/mmdeploy MMDeploy is the OpenMMLab model deployment framework that converts PyTorch-based OpenMMLab models (mmdetection, mmsegmentation, etc.) into … | 23 | 3137 | maintenance |
| google-deepmind/trfl TRFL is a Python library built on TensorFlow that provides building-block loss operations (e.g., Q-learning, TD learning, distributional RL… | 32 | 3131 | maintenance |
| open-mmlab/mmskeleton MMSkeleton is an OpenMMLAB toolbox for skeleton-based human understanding, built on PyTorch. It supports 2D pose estimation, skeleton-based… | 23 | 3127 | maintenance |
| diegoantognini/pyGAT A PyTorch implementation of the Graph Attention Network (GAT) model from Veličković et al. (2017), including a sparse-matrix variant. It re… | 32 | 3123 | maintenance |
| cysmith/neural-style-tf A TensorFlow implementation of neural style transfer based on Gatys et al.'s convolutional neural network approach, with support for video … | 32 | 3104 | maintenance |
| tusen-ai/simpledet SimpleDet is a Python framework built on MXNet for object detection and instance recognition. It provides state-of-the-art detection models… | 32 | 3085 | maintenance |
| qwopqwop200/GPTQ-for-LLaMa A Python library that applies GPTQ 4-bit weight quantization to LLaMA large language models, drastically reducing memory usage and checkpoi… | 30 | 3073 | maintenance |
| Tramac/awesome-semantic-segmentation-pytorch A PyTorch library providing concise, modifiable reference implementations of many semantic segmentation models such as FCN, PSPNet, DeepLab… | 32 | 3069 | maintenance |
| stellargraph/stellargraph StellarGraph is a Python library for machine learning on graphs and networks, offering state-of-the-art graph neural network algorithms suc… | 23 | 3060 | maintenance |
| argman/EAST A TensorFlow re-implementation of the EAST (Efficient and Accurate Scene Text Detector) deep learning model for detecting text in natural s… | 32 | 3059 | maintenance |
| rinongal/textual_inversion Official implementation of the Textual Inversion paper, which learns new word embeddings in a frozen text-to-image (Latent Diffusion) model… | 32 | 3055 | maintenance |
| avinashpaliwal/Super-SloMo A PyTorch implementation of the Super SloMo paper for high-quality video frame interpolation, generating multiple intermediate frames to co… | 10 | 3025 | maintenance |