Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: deep-learning

2653 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
andabi/deep-voice-conversion
A TensorFlow implementation of deep neural networks for voice conversion (voice style transfer) that converts a source speaker's voice into…
323938maintenance
chengtan9907/OpenSTL
OpenSTL is a comprehensive benchmark and modular framework for spatio-temporal predictive learning, covering video prediction methods acros…
541137active
facebookresearch/watermark-anything
Official PyTorch implementation and pretrained models for the paper 'Watermark Anything with Localized Messages', which embeds multiple loc…
101137active
xzf-thu/Mega-ASR
Mega-ASR is a foundation automatic speech recognition model trained on 2.6M samples spanning 7 atomic acoustic conditions and 54 compound r…
591136active
horseee/LLM-Pruner
LLM-Pruner is a PyTorch library implementing structural pruning of large language models based on gradient information, as published at Neu…
291136active
Dao-AILab/quack
QuACK is a collection of high-performance GPU kernels (RMSNorm, LayerNorm, softmax, cross-entropy, GEMM with epilogues) written in NVIDIA's…
851135active
Kiteretsu77/APISR
APISR is a deep-learning based super-resolution tool that restores and enhances low-quality, low-resolution anime images and videos using t…
371135active
rhymes-ai/Allegro
Allegro is an open-source text-to-video generation model that produces high-quality 720p videos up to 6 seconds at 15 FPS from text prompts…
241135active
Janspiry/Image-Super-Resolution-via-Iterative-Refinement
An unofficial PyTorch implementation of SR3 (Image Super-Resolution via Iterative Refinement), a diffusion-based model for image super-reso…
323923maintenance
zai-org/SCAIL-2
Official implementation of SCAIL-2, an open-source model for end-to-end controlled character animation that drives character videos from re…
581132active
FlagOpen/RoboBrain2.5
RoboBrain 2.5 is an open-source embodied AI foundation model from BAAI that combines multimodal large language model capabilities with 3D s…
501132active
sooftware/conformer
An unofficial PyTorch implementation of the Conformer architecture (convolution-augmented Transformer) from the INTERSPEECH 2020 paper, tar…
721131active
noahcao/OC_SORT
OC-SORT is a pure motion-model-based multi-object tracker for video, improving on SORT by fixing Kalman filter limitations to handle occlus…
671131stable
ikostrikov/pytorch-a2c-ppo-acktr-gail
A PyTorch implementation of several deep reinforcement learning algorithms: A2C, PPO, ACKTR, and GAIL (imitation learning). It works with O…
323903maintenance
Dobiasd/frugally-deep
frugally-deep is a lightweight header-only C++ library for running inference (forward passes) on Keras/TensorFlow models without linking ag…
851128active
BeingBeyond/Being-H
Being-H is a family of human-centric embodied foundation models, including VLA models (Being-H0.5, Being-H0) and latent world-action models…
631126active
OpenGVLab/VideoMamba
VideoMamba is a state space model (Mamba-based) architecture for efficient video understanding, released with code and pretrained models fr…
251125active
geopavlakos/hamer
HaMeR (Hand Mesh Recovery) is a transformer-based model that reconstructs 3D hand meshes from single monocular images using the MANO parame…
561124active
OpenDriveLab/UniVLA
UniVLA is an open-source framework for training cross-embodiment vision-language-action (VLA) robot policies using task-centric latent acti…
431124active
THUDM/SwissArmyTransformer
SwissArmyTransformer (sat) is a PyTorch library for developing custom Transformer model variants where models like BERT, GPT, T5, GLM, and …
231121active
FlagAI-Open/FlagAI
FlagAI is a Python toolkit for training, fine-tuning, and deploying large-scale AI models across NLP, CV, and vision-language tasks. It int…
643869maintenance
PaddlePaddle/PaddleHelix
PaddleHelix is a bio-computing platform built on PaddlePaddle featuring large-scale representation learning and multi-task deep learning fo…
571119active
alibaba-damo-academy/RynnVLA-002
RynnVLA-002 is a unified autoregressive Vision-Language-Action and world model that generates robot actions from text and image observation…
431119active
caiyuanhao1998/MST
A Python toolbox for spectral compressive imaging reconstruction that implements over 15 algorithms including MST, CST, DAUHST, BiSCI, HDNe…
531118active
yangxue0827/RotationDetection
AlphaRotate is a TensorFlow-based benchmark and toolbox for rotated (oriented) object detection, implementing detectors such as R2CNN, Reti…
231118active
Lasagne/Lasagne
Lasagne is a lightweight Python library for building and training neural networks on top of Theano. It supports feed-forward, convolutional…
233857maintenance
HITsz-TMG/Uni-MoE
Uni-MoE is a family of open-source Mixture-of-Experts (MoE) based omnimodal large language models that understand and generate across text,…
681116active
NTMC-Community/MatchZoo
MatchZoo is a Python toolkit for designing, comparing, and sharing deep text matching models. It provides a unified data pipeline, pre-buil…
233849maintenance
yahoo/TensorFlowOnSpark
TensorFlowOnSpark is a Python library that lets existing TensorFlow programs run distributed training and inference on Apache Spark and Had…
233845maintenance
openai/improved-diffusion
The official codebase for OpenAI's Improved Denoising Diffusion Probabilistic Models paper, providing a Python package for training and sam…
323844maintenance
NVIDIA/earth2studio
Earth2Studio is a Python deep-learning framework from NVIDIA for building, exploring, and deploying AI-driven weather and climate workflows…
881112active
baofff/U-ViT
U-ViT is the official PyTorch implementation of a ViT-based backbone architecture for diffusion models from the CVPR 2023 paper 'All are Wo…
321110stable
unitreerobotics/unifolm-world-model-action
UnifoLM-WMA-0 is Unitree's open-source world-model-action framework for general-purpose robot learning across multiple robotic embodiments.…
501109active
princeton-vl/DPVO
DPVO is a deep learning-based visual odometry and SLAM system that estimates camera trajectories from video or image sequences using patch-…
321108active
THU-MIG/RepViT
Official PyTorch implementation of RepViT, a family of lightweight CNNs designed by integrating efficient ViT architectural designs into Mo…
191108stable
open-mmlab/PowerPaint
PowerPaint is a versatile image inpainting model (ECCV 2024) built on diffusion models that handles text-guided object insertion, object re…
701107active
OpenMOSS/MOVA
MOVA is an open-source foundation model and toolkit for joint video-audio generation, synthesizing synchronized video and audio in a single…
601106active
szymanowiczs/splatter-image
Official PyTorch implementation of 'Splatter Image: Ultra-Fast Single-View 3D Reconstruction' (CVPR 2024), which uses an image-to-image net…
261106active
google-research/multinerf
Google Research's official code release for three NeRF papers: Mip-NeRF 360, Ref-NeRF, and RawNeRF, written in JAX. It trains neural radian…
103808maintenance
TencentARC/T2I-Adapter
Official implementation of T2I-Adapter, lightweight adapter models that add controllable conditioning (sketch, canny, lineart, depth, pose)…
313801maintenance
PharMolix/OpenBioMed
OpenBioMed is an open-source toolkit and agent platform for biomedicine and life science, offering multimodal models like BioMedGPT-Mol, 45…
721103active
yangheng95/PyABSA
PyABSA is a PyTorch-based library providing state-of-the-art models for aspect-based sentiment analysis, including aspect term extraction, …
671102active
yzslab/gaussian-splatting-lightning
A PyTorch Lightning implementation of 3D Gaussian Splatting with many derived algorithms (Mip-Splatting, LightGaussian, 2DGS, deformable Ga…
581101active
zai-org/CogView4
CogView4, CogView3-Plus and CogView3 are open-source text-to-image generation models from Zhipu AI, with CogView4 being a 6B-parameter DiT-…
281101active
naturomics/CapsNet-Tensorflow
A TensorFlow implementation of CapsNet (Capsule Networks) based on Geoffrey Hinton's paper 'Dynamic Routing Between Capsules'. It supports …
323786maintenance
yongliang-wu/DFT
DFT (Dynamic Fine-Tuning) is the official implementation of an ICLR 2026 paper that improves Supervised Fine-Tuning of LLMs by dynamically …
611100active
lucidrains/stylegan2-pytorch
A simple PyTorch implementation of StyleGAN2, a state-of-the-art generative adversarial network, trainable entirely from the command line w…
233783maintenance
yandex-research/tabm
TabM is a PyTorch-based deep learning model for tabular data that efficiently imitates an ensemble of MLPs through parameter-efficient ense…
361099active
LAMDA-CL/PyCIL
PyCIL is a PyTorch-based Python toolbox for class-incremental learning, implementing the largest collection of CIL methods for reproducible…
521098active
WangLibo1995/GeoSeg
GeoSeg is an open-source PyTorch-based semantic segmentation toolbox focused on Vision Transformers for remote sensing imagery, featuring t…
321096active
Eyeline-Labs/Go-with-the-Flow
Official implementation of the CVPR 2025 Oral paper 'Go-with-the-Flow', which controls motion in video diffusion models by replacing i.i.d.…
411093active
lizhe00/AnimatableGaussians
Official PyTorch implementation of the CVPR 2024 paper 'Animatable Gaussians', which learns pose-dependent Gaussian maps for high-fidelity …
271093active
lucidrains/tab-transformer-pytorch
A PyTorch implementation of the TabTransformer architecture, an attention-based neural network for tabular data, also including the FT Tran…
701091stable
yerfor/Real3DPortrait
Official PyTorch implementation of Real3D-Portrait, an ICLR 2024 Spotlight paper for one-shot realistic 3D talking portrait synthesis. It g…
261091active
Toni-SM/skrl
skrl is an open-source modular Reinforcement Learning library written in Python, implemented in PyTorch, JAX, and NVIDIA Warp. It supports …
811089active
mymusise/ChatGLM-Tuning
A Python toolkit for fine-tuning the ChatGLM-6B large language model using LoRA (Low-Rank Adaptation) with the Alpaca dataset. It provides …
103740maintenance
rhymes-ai/Aria
Aria is an open multimodal native Mixture-of-Experts (MoE) model with 25.3B total parameters (3.9B activated per token) and a 64K multimoda…
231087active
ndif-team/nnsight
nnsight is a Python library for interpreting and intervening on the internals of deep learning models, built on PyTorch. It lets researcher…
881085active
microsoft/TimeCraft
TimeCraft is a diffusion model-based framework for generating high-quality synthetic time series data across domains. It uses learned seman…
641085active
DSE-MSU/DeepRobust
DeepRobust is a PyTorch library for adversarial robustness research, providing implementations of attack and defense methods for both image…
451085active
mlc-ai/web-stable-diffusion
A project that compiles and runs Stable Diffusion text-to-image models entirely inside web browsers using WebGPU and WebAssembly, with no s…
303721maintenance
Alpha-VLLM/Lumina-mGPT-2.0
Lumina-mGPT 2.0 is a stand-alone decoder-only autoregressive model trained from scratch that unifies a broad range of image generation task…
421084active
GraphSAGE
Reference implementation of the GraphSAGE algorithm for inductive representation learning on large graphs using stochastic graph convolutio…
323719maintenance
bytedance/byteps
BytePS is a high-performance parameter server framework for distributed deep neural network training, supporting TensorFlow, Keras, PyTorch…
103717maintenance
kimiyoung/transformer-xl
Official implementation of Transformer-XL, an attention-based language model architecture that extends context beyond a fixed length via se…
323714maintenance
agi-brain/xuance
XuanCe is an open-source Python library of deep reinforcement learning (DRL) and multi-agent reinforcement learning (MARL) algorithm implem…
951082active
NVlabs/Fast-dLLM
NVIDIA's official implementation of Fast-dLLM, a family of training-free and fine-tuning-based acceleration techniques for diffusion-based …
571082active
charlesq34/pointnet2
Official TensorFlow implementation of PointNet++, a deep neural network that learns hierarchical features on 3D point clouds using metric-s…
323700maintenance
openai/glide-text2im
Official codebase for GLIDE, a diffusion-based text-conditional image synthesis model from OpenAI. It provides pretrained models and notebo…
103684maintenance
facebookresearch/hiera
Hiera is the official PyTorch implementation of a hierarchical vision transformer from Meta AI (ICML 2023 Oral). It achieves state-of-the-a…
201074active
cleardusk/3DDFA
A PyTorch implementation of the TPAMI 2017 paper 'Face Alignment in Full Pose Range: A 3D Total Solution' (3DDFA). It fits a 3D Morphable M…
233677maintenance
zju3dv/InfiniDepth
InfiniDepth is a CVPR 2026 research library for monocular depth estimation that represents depth as neural implicit fields, allowing depth …
531073active
open-gigaai/giga-train
GigaTrain is an efficient and scalable Python training framework for large AI models, supporting distributed multi-GPU/multi-node execution…
621072active
AILab-CVC/UniRepLKNet
UniRepLKNet is a large-kernel ConvNet architecture (CVPR 2024, TPAMI 2025) that provides universal perception across image, audio, video, p…
431072stable
NVIDIA-Merlin/HugeCTR
HugeCTR is a GPU-accelerated deep learning framework from NVIDIA designed for training and inference of large recommender models, especiall…
771071active
memoavatar/memo
MEMO is an open-weight diffusion model for generating expressive, identity-consistent talking videos from a single reference image and an a…
401070active
juntang-zhuang/Adabelief-Optimizer
AdaBelief is a deep learning optimizer that adapts step sizes based on the 'belief' in observed gradients, combining Adam's fast convergenc…
321070stable
yeates/PromptFix
PromptFix is a PyTorch implementation of a diffusion-model-based image restoration model that follows natural language instructions to fix …
241070active
RL-VIG/LibFewShot
LibFewShot is a comprehensive PyTorch library for few-shot learning, implementing many fine-tuning, meta-learning, and metric-learning meth…
541069active
princeton-nlp/SimCSE
SimCSE is a Python library and research codebase implementing simple contrastive learning for sentence embeddings, with pre-trained unsuper…
233654maintenance
guochengqian/PointNeXt
PointNeXt is the official PyTorch implementation of the NeurIPS'22 paper that improves PointNet++ via better training and model scaling str…
931067stable
hujie-frank/SENet
Official Caffe/CUDA implementation of Squeeze-and-Excitation Networks (SENet), channel-attention building blocks for convolutional neural n…
323646maintenance
gangweix/pixel-perfect-depth
Pixel-Perfect Depth is a monocular depth estimation model based on pixel-space diffusion transformers that produces flying-pixel-free depth…
491064active
lucidrains/mlp-mixer-pytorch
A PyTorch implementation of Google AI's MLP-Mixer, an all-MLP architecture for image classification that uses neither convolutions nor atte…
481064active
abertsch72/unlimiformer
Unlimiformer is a method and official implementation for augmenting pretrained encoder-decoder transformers with retrieval-based attention,…
301062stable
vijishmadhavan/ArtLine
ArtLine is a deep learning project that converts portrait photos into line art portraits, with a ControlNet-based variant that adjusts styl…
323630maintenance
NVIDIA/DreamDojo
NVIDIA's official PyTorch codebase for DreamDojo, a generalist robot world model pretrained on 44k hours of human egocentric video and post…
481059active
clovaai/stargan-v2
The official PyTorch implementation of StarGAN v2, a CVPR 2020 paper on diverse image-to-image translation across multiple domains using a …
323617maintenance
mgsalem/Tensorflow-Project-Template
A Python project template that provides a recommended folder structure and object-oriented skeleton (base model, base trainer, data loader,…
323616maintenance
YunYang1994/tensorflow-yolov3
A TensorFlow 1.x implementation of the YOLOv3 real-time object detector, reproducing the 'YOLOv3: An Incremental Improvement' paper. It sup…
233614maintenance
open-gigaai/giga-models
GigaModels is an open-source Python framework providing pipelines for training, inference, deployment, and compression of multi-modal, gene…
621057active
X-LANCE/SLAM-LLM
SLAM-LLM is a deep learning toolkit for training custom multimodal large language models focused on speech, language, audio, and music proc…
551056active
apache/singa
Apache SINGA is a distributed deep learning platform for training neural networks across multiple devices and machines. It provides a C++ c…
643606maintenance
yoyo-nb/Thin-Plate-Spline-Motion-Model
The official PyTorch implementation of the CVPR 2022 paper 'Thin-Plate Spline Motion Model for Image Animation'. It animates a source image…
323604maintenance
BAAI-DCAI/Bunny
Bunny is a family of lightweight multimodal vision-language models that combine plug-and-play vision encoders (EVA-CLIP, SigLIP) with langu…
261053active
Tencent-Hunyuan/HunyuanVideo-Foley
HunyuanVideo-Foley is a multimodal diffusion model from Tencent Hunyuan that generates high-fidelity Foley sound effects synchronized with …
371052active
showlab/MotionDirector
MotionDirector is a research library for customizing text-to-video diffusion models to generate videos with desired motions from a small se…
271050active
drprojects/superpoint_transformer
Official PyTorch implementation of Superpoint Transformer (ICCV'23), SuperCluster (3DV'24), and EZ-SP (ICRA'26) for efficient semantic and …
641049active
arcee-ai/DistillKit
DistillKit is an open-source Python toolkit for knowledge distillation of large language models, supporting both online and offline distill…
601047active
facebookresearch/pytorchvideo
PyTorchVideo is a deep learning library from Facebook Research focused on video understanding research, built on PyTorch. It provides reusa…
593566maintenance

← prev page 12 / 27 next →