Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: deep-learning

2771 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
PaddlePaddle/PaddleGAN
PaddleGAN is a Python library providing high-performance implementations of classic and state-of-the-art Generative Adversarial Networks bu…
238048maintenance
ContinualAI/avalanche
Avalanche is an end-to-end continual learning library built on PyTorch, developed by ContinualAI. It provides modules for benchmarks, train…
272088active
vitoplantamura/OnnxStream
A lightweight C++ inference library for ONNX models that streams weights to run large models in very little memory, accelerated by XNNPACK.…
592086active
ali-vilab/In-Context-LoRA
Official repository for In-Context LoRA (IC-LoRA), a framework for adapting Diffusion Transformers to diverse visual generation tasks via L…
222083active
alex-damian/pulse
PULSE is a Python research implementation of a CVPR 2020 paper that upscales low-resolution face photos by searching the latent space of a …
328023maintenance
PKU-YuanGroup/Helios
Helios is a 14B autoregressive diffusion model for real-time, minute-scale video generation supporting text-to-video, image-to-video, and v…
592076active
facebookresearch/ConvNeXt-V2
Official PyTorch implementation of ConvNeXt V2, a family of pure convolutional neural network models co-designed with a fully convolutional…
102069stable
brightmart/text_classification
A collection of deep learning baseline models for text classification in NLP, implemented in TensorFlow. It covers classic architectures li…
327940maintenance
shanglianlm0525/PyTorch-Networks
A collection of PyTorch implementations of classic and modern CNN architectures, covering classification, detection, segmentation, face, an…
532055active
marcoslucianops/DeepStream-Yolo
A collection of configuration files, parsers, and conversion utilities for running YOLO-family object detection models on NVIDIA DeepStream…
612054active
jaywalnut310/vits
VITS is the official PyTorch implementation of an end-to-end text-to-speech model based on a conditional variational autoencoder with adver…
327889maintenance
visomaster/VisoMaster
VisoMaster is a Python-based desktop application for AI-powered face swapping and face editing in images and videos. It supports multiple s…
272052active
01-ai/Yi
Yi is a family of open-source large language models trained from scratch by 01.AI, including base and chat models in multiple sizes, with b…
277836maintenance
PRIS-CV/DemoFusion
DemoFusion is a CVPR 2024 framework that extends open-source latent diffusion models like SDXL to generate high-resolution images without a…
482041stable
jeshraghian/snntorch
snnTorch is a Python library for gradient-based deep learning with spiking neural networks, built as an extension of PyTorch. It provides s…
672036active
deep-floyd/IF
DeepFloyd IF is an open-source text-to-image model library implementing a cascaded pixel diffusion architecture with a frozen T5 text encod…
227804maintenance
tensorflow/privacy
TensorFlow Privacy is a Python library providing TensorFlow optimizers for training machine learning models with differential privacy. It i…
672026active
NUS-HPC-AI-Lab/VideoSys
VideoSys is an open-source Python library providing easy and efficient infrastructure for video generation, supporting training, inference,…
452022active
XavierXiao/Dreambooth-Stable-Diffusion
An implementation of Google's Dreambooth fine-tuning method applied to Stable Diffusion, enabling personalization of a text-to-image diffus…
327738maintenance
SakanaAI/continuous-thought-machines
The Continuous Thought Machine (CTM) is a neural network architecture from Sakana AI that uses neuron-level temporal dynamics and neural sy…
462019active
instantX-research/InstantStyle
InstantStyle is a framework for style-preserving text-to-image generation that disentangles style and content from reference images using f…
262018active
NVlabs/SPADE
Official PyTorch implementation of SPADE (GauGAN), a CVPR 2019 method for synthesizing photorealistic images from semantic segmentation map…
327717maintenance
tdrussell/diffusion-pipe
A Python training script for fine-tuning diffusion models (image and video generation) using DeepSpeed pipeline parallelism across multiple…
672015active
philipperemy/keras-tcn
A Keras/TensorFlow implementation of Temporal Convolutional Networks (TCN) with dilated causal convolutions, usable as a drop-in layer alte…
622012active
lyhue1991/torchkeras
torchkeras is a lightweight PyTorch model training template library that brings Keras-style compile/fit/evaluate APIs to PyTorch. Its core …
552009active
AntixK/PyTorch-VAE
A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f…
387665maintenance
kohya-ss/musubi-tuner
Musubi Tuner is a set of Python scripts for training LoRA (Low-Rank Adaptation) adapters for video and image generation model architectures…
852002active
bytetriper/RAE
Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u…
482001active
zai-org/GLM-130B
GLM-130B is an open bilingual (English and Chinese) 130-billion-parameter dense language model pre-trained with the General Language Model …
327651maintenance
mil-tokyo/webdnn
WebDNN is a framework for running deep neural network inference directly in the web browser, accepting ONNX models without Python preproces…
651999active
DEIM
DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp…
621999active
clementchadebec/benchmark_VAE
Pythae is a PyTorch library that unifies implementations of many Variational Autoencoder (VAE) variants under a common interface, enabling …
231995active
facebookresearch/dino
PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th…
107611maintenance
xingyizhou/CenterNet
CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi…
327573maintenance
hkchengrex/XMem
XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in…
231983stable
lucidrains/titans-pytorch
An unofficial PyTorch implementation of the Titans architecture, a neural long-term memory module for transformers that learns to memorize …
701980active
davda54/sam
An unofficial PyTorch implementation of Sharpness-Aware Minimization (SAM) and its adaptive variant ASAM, provided as an optimizer wrapper …
321980stable
google-deepmind/alphagenome
A Python SDK providing programmatic access to Google DeepMind's AlphaGenome model, which predicts genomic regulatory outputs such as gene e…
891979active
cloneofsimo/lora
A Python library for applying Low-Rank Adaptation (LoRA) to quickly fine-tune text-to-image diffusion models like Stable Diffusion. It prod…
227550maintenance
openlm-research/open_llama
OpenLLaMA is a permissively licensed (Apache-2.0) open reproduction of Meta AI's LLaMA, releasing 3B, 7B, and 13B pretrained model weights …
307528maintenance
WassimTenachi/PhySO
PhySO is a Python library for physical symbolic optimization that uses deep reinforcement learning to discover analytical physical laws fro…
521971active
bigcode-project/starcoder
StarCoder is a 15B-parameter code language model trained on 80+ programming languages, and this repository hosts the fine-tuning and infere…
307506maintenance
google-deepmind/tapnet
Google DeepMind's official repository for Tracking Any Point (TAP), containing the TAP-Vid and TAPVid-3D benchmarks, the TAPIR and TAPNext …
741968active
meta-recsys/generative-recommenders
Meta's research library implementing HSTU and M-FALCON from the ICML'24 paper 'Actions Speak Louder than Words: Trillion-Parameter Sequenti…
691966active
siliconflow/onediff
OneDiff is an out-of-the-box acceleration library for diffusion models, providing PyTorch compilation tools and optimized GPU kernels. It i…
481964active
NVIDIA-NeMo/RL
NeMo RL is NVIDIA's open-source post-training library for scaling reinforcement learning methods (GRPO, PPO, DPO, SFT, distillation) on LLM…
811961active
OpenMotionLab/MotionGPT
MotionGPT is a unified motion-language model that treats 3D human motion as a foreign language by converting motion into discrete motion to…
321961active
2U1/Qwen-VL-Series-Finetune
An open-source Python repository providing training scripts for fine-tuning Alibaba's Qwen-VL series of vision-language models (Qwen2-VL, Q…
671960active
open-mmlab/mmagic
MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri…
237457maintenance
SizheAn/PanoHead
PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis…
291956active
alibaba/EasyCV
EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo…
321954active
eriklindernoren/PyTorch-YOLOv3
A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights…
327440maintenance
Fafa-DL/Awesome-Backbones
A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin …
331953active
meta-pytorch/opacus
Opacus is a PyTorch library for training neural networks with differential privacy via DP-SGD, requiring minimal code changes through its P…
791952active
LTH14/mar
Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive …
541949stable
openai/guided-diffusion
OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio…
327419maintenance
Vchitect/Latte
Official PyTorch implementation of Latte, a latent diffusion transformer for video generation. It includes model definitions, pre-trained c…
711948active
kyegomez/BitNet
A PyTorch implementation of the BitNet architecture from the paper 'BitNet: Scaling 1-bit Transformers for Large Language Models', providin…
721945active
Graph Convolutional Networks (GCN)
A TensorFlow implementation of Graph Convolutional Networks (GCN) for semi-supervised node classification on graphs, accompanying the ICLR …
327400maintenance
jd-opensource/JoyAI-Echo
JoyAI-Echo is a Python framework for long-horizon audio-visual generation, producing coherent multi-shot videos up to ~5 minutes with paire…
581943active
tensorlayer/TensorLayer
TensorLayer is a TensorFlow-based deep learning and reinforcement learning library offering customizable neural layers for researchers and …
237381maintenance
PixArt-alpha/PixArt-sigma
PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with…
251939active
google-deepmind/lab
DeepMind Lab is a customisable 3D learning environment built on Quake III Arena (ioquake3) that provides navigation and puzzle-solving task…
237373maintenance
NVlabs/RADIO
Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro…
641933active
tum-pbs/PhiFlow
PhiFlow is an open-source Python simulation toolkit for solving partial differential equations with support for optimization and machine le…
721929active
Audio-AGI/AudioSep
AudioSep is the official implementation of the 'Separate Anything You Describe' foundation model for open-domain, language-queried audio so…
281929active
Yuanshi9815/OminiControl
OminiControl is a universal control framework for Diffusion Transformer models like FLUX, supporting subject-driven and spatial control (ed…
621927active
OpenTalker/video-retalking
VideoReTalking is a Python research system from SIGGRAPH Asia 2022 that edits real-world talking-head videos to match a given audio track, …
237280maintenance
FACEGOOD/FACEGOOD-Audio2Face
FACEGOOD Audio2Face is an open-source deep learning framework that converts audio into facial blendshape weights for driving digital humans…
641909active
NVlabs/nvdiffrast
Nvdiffrast is a PyTorch library from NVIDIA providing high-performance, GPU-accelerated primitive operations for rasterization-based differ…
551905stable
lucidrains/byol-pytorch
A PyTorch library implementing the Bootstrap Your Own Latent (BYOL) self-supervised learning method from DeepMind. It wraps any image-based…
581903active
google-deepmind/penzai
Penzai is a JAX research toolkit for building, editing, and visualizing neural networks as legible, functional pytree data structures. It i…
101901active
sapientinc/HRM-Text
HRM-Text is a 1B-parameter text generation model based on the hierarchical recurrent HRM architecture, released with a complete pretraining…
531899active
flexflow/flexflow-train
FlexFlow Train is a deep learning framework that accelerates distributed DNN training by automatically searching for efficient parallelizat…
671898active
policy-gradient/GRPO-Zero
A minimal from-scratch Python implementation of DeepSeek's GRPO (Group Relative Policy Optimization) algorithm for reinforcement learning t…
271897active
LeapLabTHU/Absolute-Zero-Reasoner
Official implementation of Absolute Zero Reasoner (AZR), a system that trains LLM reasoning via reinforced self-play with zero external dat…
361893active
qqwweee/keras-yolo3
A Keras (TensorFlow backend) implementation of YOLOv3 for object detection, including Darknet weight conversion, image/video detection scri…
327114maintenance
laugh12321/TensorRT-YOLO
A C++/Python deployment toolkit for running YOLO-family models (YOLOv3 through YOLO26) on NVIDIA GPUs using TensorRT, with custom plugins, …
631880active
nitrain/nitrain
Nitrain is a framework-agnostic Python library for sampling, augmenting, and training AI models on medical imaging datasets, with support f…
231880active
nndeploy/nndeploy
nndeploy is an easy-to-use, high-performance AI deployment framework written in C++ with Python bindings. It provides a visual drag-and-dro…
871868active
we0091234/Chinese_license_plate_detection_recognition
A PyTorch-based Chinese license plate detection and recognition system built on YOLOv5 for detection and CRNN for recognition. It supports …
711868active
NVIDIA-AI-IOT/Lidar_AI_Solution
NVIDIA's collection of GPU-accelerated Lidar AI inference solutions for autonomous driving, including optimized implementations of PointPil…
721867active
BytedTsinghua-SIA/DAPO
DAPO is an open-source reinforcement learning system for large-scale LLM training, released by ByteDance Seed and Tsinghua AIR. It implemen…
291861active
patrick-kidger/jaxtyping
A Python library providing type annotations and runtime type-checking for the shape and dtype of arrays and tensors in JAX, PyTorch, NumPy,…
951859active
Omni-Avatar/OmniAvatar
OmniAvatar is an audio-driven full-body avatar video generation model built on Wan2.1 text-to-video diffusion models with LoRA-based audio …
341859active
laekov/fastmoe
FastMoE is a PyTorch library providing efficient Mixture of Experts (MoE) layers with custom C/CUDA operators. It supports distributed expe…
261859active
OpenNMT
OpenNMT is an open-source ecosystem for neural machine translation and sequence learning, with PyTorch (OpenNMT-py) and TensorFlow (OpenNMT…
447012maintenance
KlingAIResearch/ReCamMaster
ReCamMaster is a reference implementation of a camera-controlled generative video rendering model that re-renders a single source video alo…
441855active
facebookresearch/MetaCLIP
Meta's research code and models for Meta CLIP, a reimplementation and scaling recipe for CLIP-style contrastive vision-language models, inc…
821854active
LuChengTHU/dpm-solver
Official PyTorch implementation of DPM-Solver and DPM-Solver++, fast high-order ODE solvers for diffusion probabilistic model sampling that…
321852stable
dotnet/TorchSharp
TorchSharp is a .NET library providing bindings to LibTorch, the library that powers PyTorch, with a focus on tensors and a PyTorch-like AP…
731850active
probcomp/Gen.jl
Gen.jl is a general-purpose probabilistic programming system embedded in Julia that lets users write generative models as probabilistic pro…
621850active
NVlabs/stylegan3
Official PyTorch implementation of StyleGAN3 (Alias-Free GANs), a state-of-the-art generative adversarial network for high-fidelity image s…
326943maintenance
Tencent-Hunyuan/HunyuanVideo-I2V
HunyuanVideo-I2V is Tencent's open-source image-to-video generation framework built on the HunyuanVideo diffusion model, providing PyTorch …
541840active
ytongbai/LVM
LVM is a large vision model trained with sequential next-token prediction over 'visual sentences', using no linguistic data. It builds on O…
301838active
NVIDIA/pix2pixHD
PyTorch implementation of pix2pixHD, a conditional GAN method for synthesizing and manipulating high-resolution (2048x1024) photorealistic …
326923maintenance
cazala/synaptic
Synaptic is an architecture-free neural network library for JavaScript that runs in both Node.js and the browser. It supports building and …
666912maintenance
dauparas/ProteinMPNN
ProteinMPNN is a PyTorch-based tool that designs amino acid sequences for given protein backbone structures using a message-passing neural …
231834stable
openai/point-e
Point-E is OpenAI's official release of models and code for generating 3D point clouds from text prompts or images using diffusion models. …
326895maintenance
ZFTurbo/Weighted-Boxes-Fusion
A Python library implementing several methods for ensembling bounding boxes from multiple object detection models, including Non-maximum Su…
651827stable

← prev page 8 / 28 next →