domain: deep-learning
2771 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| PaddlePaddle/PaddleGAN PaddleGAN is a Python library providing high-performance implementations of classic and state-of-the-art Generative Adversarial Networks bu… | 23 | 8048 | maintenance |
| ContinualAI/avalanche Avalanche is an end-to-end continual learning library built on PyTorch, developed by ContinualAI. It provides modules for benchmarks, train… | 27 | 2088 | active |
| vitoplantamura/OnnxStream A lightweight C++ inference library for ONNX models that streams weights to run large models in very little memory, accelerated by XNNPACK.… | 59 | 2086 | active |
| ali-vilab/In-Context-LoRA Official repository for In-Context LoRA (IC-LoRA), a framework for adapting Diffusion Transformers to diverse visual generation tasks via L… | 22 | 2083 | active |
| alex-damian/pulse PULSE is a Python research implementation of a CVPR 2020 paper that upscales low-resolution face photos by searching the latent space of a … | 32 | 8023 | maintenance |
| PKU-YuanGroup/Helios Helios is a 14B autoregressive diffusion model for real-time, minute-scale video generation supporting text-to-video, image-to-video, and v… | 59 | 2076 | active |
| facebookresearch/ConvNeXt-V2 Official PyTorch implementation of ConvNeXt V2, a family of pure convolutional neural network models co-designed with a fully convolutional… | 10 | 2069 | stable |
| brightmart/text_classification A collection of deep learning baseline models for text classification in NLP, implemented in TensorFlow. It covers classic architectures li… | 32 | 7940 | maintenance |
| shanglianlm0525/PyTorch-Networks A collection of PyTorch implementations of classic and modern CNN architectures, covering classification, detection, segmentation, face, an… | 53 | 2055 | active |
| marcoslucianops/DeepStream-Yolo A collection of configuration files, parsers, and conversion utilities for running YOLO-family object detection models on NVIDIA DeepStream… | 61 | 2054 | active |
| jaywalnut310/vits VITS is the official PyTorch implementation of an end-to-end text-to-speech model based on a conditional variational autoencoder with adver… | 32 | 7889 | maintenance |
| visomaster/VisoMaster VisoMaster is a Python-based desktop application for AI-powered face swapping and face editing in images and videos. It supports multiple s… | 27 | 2052 | active |
| 01-ai/Yi Yi is a family of open-source large language models trained from scratch by 01.AI, including base and chat models in multiple sizes, with b… | 27 | 7836 | maintenance |
| PRIS-CV/DemoFusion DemoFusion is a CVPR 2024 framework that extends open-source latent diffusion models like SDXL to generate high-resolution images without a… | 48 | 2041 | stable |
| jeshraghian/snntorch snnTorch is a Python library for gradient-based deep learning with spiking neural networks, built as an extension of PyTorch. It provides s… | 67 | 2036 | active |
| deep-floyd/IF DeepFloyd IF is an open-source text-to-image model library implementing a cascaded pixel diffusion architecture with a frozen T5 text encod… | 22 | 7804 | maintenance |
| tensorflow/privacy TensorFlow Privacy is a Python library providing TensorFlow optimizers for training machine learning models with differential privacy. It i… | 67 | 2026 | active |
| NUS-HPC-AI-Lab/VideoSys VideoSys is an open-source Python library providing easy and efficient infrastructure for video generation, supporting training, inference,… | 45 | 2022 | active |
| XavierXiao/Dreambooth-Stable-Diffusion An implementation of Google's Dreambooth fine-tuning method applied to Stable Diffusion, enabling personalization of a text-to-image diffus… | 32 | 7738 | maintenance |
| SakanaAI/continuous-thought-machines The Continuous Thought Machine (CTM) is a neural network architecture from Sakana AI that uses neuron-level temporal dynamics and neural sy… | 46 | 2019 | active |
| instantX-research/InstantStyle InstantStyle is a framework for style-preserving text-to-image generation that disentangles style and content from reference images using f… | 26 | 2018 | active |
| NVlabs/SPADE Official PyTorch implementation of SPADE (GauGAN), a CVPR 2019 method for synthesizing photorealistic images from semantic segmentation map… | 32 | 7717 | maintenance |
| tdrussell/diffusion-pipe A Python training script for fine-tuning diffusion models (image and video generation) using DeepSpeed pipeline parallelism across multiple… | 67 | 2015 | active |
| philipperemy/keras-tcn A Keras/TensorFlow implementation of Temporal Convolutional Networks (TCN) with dilated causal convolutions, usable as a drop-in layer alte… | 62 | 2012 | active |
| lyhue1991/torchkeras torchkeras is a lightweight PyTorch model training template library that brings Keras-style compile/fit/evaluate APIs to PyTorch. Its core … | 55 | 2009 | active |
| AntixK/PyTorch-VAE A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f… | 38 | 7665 | maintenance |
| kohya-ss/musubi-tuner Musubi Tuner is a set of Python scripts for training LoRA (Low-Rank Adaptation) adapters for video and image generation model architectures… | 85 | 2002 | active |
| bytetriper/RAE Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u… | 48 | 2001 | active |
| zai-org/GLM-130B GLM-130B is an open bilingual (English and Chinese) 130-billion-parameter dense language model pre-trained with the General Language Model … | 32 | 7651 | maintenance |
| mil-tokyo/webdnn WebDNN is a framework for running deep neural network inference directly in the web browser, accepting ONNX models without Python preproces… | 65 | 1999 | active |
| DEIM DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp… | 62 | 1999 | active |
| clementchadebec/benchmark_VAE Pythae is a PyTorch library that unifies implementations of many Variational Autoencoder (VAE) variants under a common interface, enabling … | 23 | 1995 | active |
| facebookresearch/dino PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th… | 10 | 7611 | maintenance |
| xingyizhou/CenterNet CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi… | 32 | 7573 | maintenance |
| hkchengrex/XMem XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in… | 23 | 1983 | stable |
| lucidrains/titans-pytorch An unofficial PyTorch implementation of the Titans architecture, a neural long-term memory module for transformers that learns to memorize … | 70 | 1980 | active |
| davda54/sam An unofficial PyTorch implementation of Sharpness-Aware Minimization (SAM) and its adaptive variant ASAM, provided as an optimizer wrapper … | 32 | 1980 | stable |
| google-deepmind/alphagenome A Python SDK providing programmatic access to Google DeepMind's AlphaGenome model, which predicts genomic regulatory outputs such as gene e… | 89 | 1979 | active |
| cloneofsimo/lora A Python library for applying Low-Rank Adaptation (LoRA) to quickly fine-tune text-to-image diffusion models like Stable Diffusion. It prod… | 22 | 7550 | maintenance |
| openlm-research/open_llama OpenLLaMA is a permissively licensed (Apache-2.0) open reproduction of Meta AI's LLaMA, releasing 3B, 7B, and 13B pretrained model weights … | 30 | 7528 | maintenance |
| WassimTenachi/PhySO PhySO is a Python library for physical symbolic optimization that uses deep reinforcement learning to discover analytical physical laws fro… | 52 | 1971 | active |
| bigcode-project/starcoder StarCoder is a 15B-parameter code language model trained on 80+ programming languages, and this repository hosts the fine-tuning and infere… | 30 | 7506 | maintenance |
| google-deepmind/tapnet Google DeepMind's official repository for Tracking Any Point (TAP), containing the TAP-Vid and TAPVid-3D benchmarks, the TAPIR and TAPNext … | 74 | 1968 | active |
| meta-recsys/generative-recommenders Meta's research library implementing HSTU and M-FALCON from the ICML'24 paper 'Actions Speak Louder than Words: Trillion-Parameter Sequenti… | 69 | 1966 | active |
| siliconflow/onediff OneDiff is an out-of-the-box acceleration library for diffusion models, providing PyTorch compilation tools and optimized GPU kernels. It i… | 48 | 1964 | active |
| NVIDIA-NeMo/RL NeMo RL is NVIDIA's open-source post-training library for scaling reinforcement learning methods (GRPO, PPO, DPO, SFT, distillation) on LLM… | 81 | 1961 | active |
| OpenMotionLab/MotionGPT MotionGPT is a unified motion-language model that treats 3D human motion as a foreign language by converting motion into discrete motion to… | 32 | 1961 | active |
| 2U1/Qwen-VL-Series-Finetune An open-source Python repository providing training scripts for fine-tuning Alibaba's Qwen-VL series of vision-language models (Qwen2-VL, Q… | 67 | 1960 | active |
| open-mmlab/mmagic MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri… | 23 | 7457 | maintenance |
| SizheAn/PanoHead PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis… | 29 | 1956 | active |
| alibaba/EasyCV EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo… | 32 | 1954 | active |
| eriklindernoren/PyTorch-YOLOv3 A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights… | 32 | 7440 | maintenance |
| Fafa-DL/Awesome-Backbones A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin … | 33 | 1953 | active |
| meta-pytorch/opacus Opacus is a PyTorch library for training neural networks with differential privacy via DP-SGD, requiring minimal code changes through its P… | 79 | 1952 | active |
| LTH14/mar Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive … | 54 | 1949 | stable |
| openai/guided-diffusion OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio… | 32 | 7419 | maintenance |
| Vchitect/Latte Official PyTorch implementation of Latte, a latent diffusion transformer for video generation. It includes model definitions, pre-trained c… | 71 | 1948 | active |
| kyegomez/BitNet A PyTorch implementation of the BitNet architecture from the paper 'BitNet: Scaling 1-bit Transformers for Large Language Models', providin… | 72 | 1945 | active |
| Graph Convolutional Networks (GCN) A TensorFlow implementation of Graph Convolutional Networks (GCN) for semi-supervised node classification on graphs, accompanying the ICLR … | 32 | 7400 | maintenance |
| jd-opensource/JoyAI-Echo JoyAI-Echo is a Python framework for long-horizon audio-visual generation, producing coherent multi-shot videos up to ~5 minutes with paire… | 58 | 1943 | active |
| tensorlayer/TensorLayer TensorLayer is a TensorFlow-based deep learning and reinforcement learning library offering customizable neural layers for researchers and … | 23 | 7381 | maintenance |
| PixArt-alpha/PixArt-sigma PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with… | 25 | 1939 | active |
| google-deepmind/lab DeepMind Lab is a customisable 3D learning environment built on Quake III Arena (ioquake3) that provides navigation and puzzle-solving task… | 23 | 7373 | maintenance |
| NVlabs/RADIO Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro… | 64 | 1933 | active |
| tum-pbs/PhiFlow PhiFlow is an open-source Python simulation toolkit for solving partial differential equations with support for optimization and machine le… | 72 | 1929 | active |
| Audio-AGI/AudioSep AudioSep is the official implementation of the 'Separate Anything You Describe' foundation model for open-domain, language-queried audio so… | 28 | 1929 | active |
| Yuanshi9815/OminiControl OminiControl is a universal control framework for Diffusion Transformer models like FLUX, supporting subject-driven and spatial control (ed… | 62 | 1927 | active |
| OpenTalker/video-retalking VideoReTalking is a Python research system from SIGGRAPH Asia 2022 that edits real-world talking-head videos to match a given audio track, … | 23 | 7280 | maintenance |
| FACEGOOD/FACEGOOD-Audio2Face FACEGOOD Audio2Face is an open-source deep learning framework that converts audio into facial blendshape weights for driving digital humans… | 64 | 1909 | active |
| NVlabs/nvdiffrast Nvdiffrast is a PyTorch library from NVIDIA providing high-performance, GPU-accelerated primitive operations for rasterization-based differ… | 55 | 1905 | stable |
| lucidrains/byol-pytorch A PyTorch library implementing the Bootstrap Your Own Latent (BYOL) self-supervised learning method from DeepMind. It wraps any image-based… | 58 | 1903 | active |
| google-deepmind/penzai Penzai is a JAX research toolkit for building, editing, and visualizing neural networks as legible, functional pytree data structures. It i… | 10 | 1901 | active |
| sapientinc/HRM-Text HRM-Text is a 1B-parameter text generation model based on the hierarchical recurrent HRM architecture, released with a complete pretraining… | 53 | 1899 | active |
| flexflow/flexflow-train FlexFlow Train is a deep learning framework that accelerates distributed DNN training by automatically searching for efficient parallelizat… | 67 | 1898 | active |
| policy-gradient/GRPO-Zero A minimal from-scratch Python implementation of DeepSeek's GRPO (Group Relative Policy Optimization) algorithm for reinforcement learning t… | 27 | 1897 | active |
| LeapLabTHU/Absolute-Zero-Reasoner Official implementation of Absolute Zero Reasoner (AZR), a system that trains LLM reasoning via reinforced self-play with zero external dat… | 36 | 1893 | active |
| qqwweee/keras-yolo3 A Keras (TensorFlow backend) implementation of YOLOv3 for object detection, including Darknet weight conversion, image/video detection scri… | 32 | 7114 | maintenance |
| laugh12321/TensorRT-YOLO A C++/Python deployment toolkit for running YOLO-family models (YOLOv3 through YOLO26) on NVIDIA GPUs using TensorRT, with custom plugins, … | 63 | 1880 | active |
| nitrain/nitrain Nitrain is a framework-agnostic Python library for sampling, augmenting, and training AI models on medical imaging datasets, with support f… | 23 | 1880 | active |
| nndeploy/nndeploy nndeploy is an easy-to-use, high-performance AI deployment framework written in C++ with Python bindings. It provides a visual drag-and-dro… | 87 | 1868 | active |
| we0091234/Chinese_license_plate_detection_recognition A PyTorch-based Chinese license plate detection and recognition system built on YOLOv5 for detection and CRNN for recognition. It supports … | 71 | 1868 | active |
| NVIDIA-AI-IOT/Lidar_AI_Solution NVIDIA's collection of GPU-accelerated Lidar AI inference solutions for autonomous driving, including optimized implementations of PointPil… | 72 | 1867 | active |
| BytedTsinghua-SIA/DAPO DAPO is an open-source reinforcement learning system for large-scale LLM training, released by ByteDance Seed and Tsinghua AIR. It implemen… | 29 | 1861 | active |
| patrick-kidger/jaxtyping A Python library providing type annotations and runtime type-checking for the shape and dtype of arrays and tensors in JAX, PyTorch, NumPy,… | 95 | 1859 | active |
| Omni-Avatar/OmniAvatar OmniAvatar is an audio-driven full-body avatar video generation model built on Wan2.1 text-to-video diffusion models with LoRA-based audio … | 34 | 1859 | active |
| laekov/fastmoe FastMoE is a PyTorch library providing efficient Mixture of Experts (MoE) layers with custom C/CUDA operators. It supports distributed expe… | 26 | 1859 | active |
| OpenNMT OpenNMT is an open-source ecosystem for neural machine translation and sequence learning, with PyTorch (OpenNMT-py) and TensorFlow (OpenNMT… | 44 | 7012 | maintenance |
| KlingAIResearch/ReCamMaster ReCamMaster is a reference implementation of a camera-controlled generative video rendering model that re-renders a single source video alo… | 44 | 1855 | active |
| facebookresearch/MetaCLIP Meta's research code and models for Meta CLIP, a reimplementation and scaling recipe for CLIP-style contrastive vision-language models, inc… | 82 | 1854 | active |
| LuChengTHU/dpm-solver Official PyTorch implementation of DPM-Solver and DPM-Solver++, fast high-order ODE solvers for diffusion probabilistic model sampling that… | 32 | 1852 | stable |
| dotnet/TorchSharp TorchSharp is a .NET library providing bindings to LibTorch, the library that powers PyTorch, with a focus on tensors and a PyTorch-like AP… | 73 | 1850 | active |
| probcomp/Gen.jl Gen.jl is a general-purpose probabilistic programming system embedded in Julia that lets users write generative models as probabilistic pro… | 62 | 1850 | active |
| NVlabs/stylegan3 Official PyTorch implementation of StyleGAN3 (Alias-Free GANs), a state-of-the-art generative adversarial network for high-fidelity image s… | 32 | 6943 | maintenance |
| Tencent-Hunyuan/HunyuanVideo-I2V HunyuanVideo-I2V is Tencent's open-source image-to-video generation framework built on the HunyuanVideo diffusion model, providing PyTorch … | 54 | 1840 | active |
| ytongbai/LVM LVM is a large vision model trained with sequential next-token prediction over 'visual sentences', using no linguistic data. It builds on O… | 30 | 1838 | active |
| NVIDIA/pix2pixHD PyTorch implementation of pix2pixHD, a conditional GAN method for synthesizing and manipulating high-resolution (2048x1024) photorealistic … | 32 | 6923 | maintenance |
| cazala/synaptic Synaptic is an architecture-free neural network library for JavaScript that runs in both Node.js and the browser. It supports building and … | 66 | 6912 | maintenance |
| dauparas/ProteinMPNN ProteinMPNN is a PyTorch-based tool that designs amino acid sequences for given protein backbone structures using a message-passing neural … | 23 | 1834 | stable |
| openai/point-e Point-E is OpenAI's official release of models and code for generating 3D point clouds from text prompts or images using diffusion models. … | 32 | 6895 | maintenance |
| ZFTurbo/Weighted-Boxes-Fusion A Python library implementing several methods for ensembling bounding boxes from multiple object detection models, including Non-maximum Su… | 65 | 1827 | stable |