Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: machine-learning

5378 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
cambrian-mllm/cambrian
Cambrian-1 is a fully open family of vision-centric multimodal large language models (MLLMs) from NYU's VISIONx group, with training and ev…
472013active
philipperemy/keras-tcn
A Keras/TensorFlow implementation of Temporal Convolutional Networks (TCN) with dilated causal convolutions, usable as a drop-in layer alte…
622012active
tatsu-lab/alpaca_eval
AlpacaEval is an LLM-based automatic evaluation framework for instruction-following language models, producing win rates against a GPT-4 ba…
362012active
uncertainty-toolbox/uncertainty-toolbox
A Python library for predictive uncertainty quantification, providing metrics, visualizations, and recalibration procedures for regression …
272012active
lyhue1991/torchkeras
torchkeras is a lightweight PyTorch model training template library that brings Keras-style compile/fit/evaluate APIs to PyTorch. Its core …
552009active
AntixK/PyTorch-VAE
A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f…
387665maintenance
kohya-ss/musubi-tuner
Musubi Tuner is a set of Python scripts for training LoRA (Low-Rank Adaptation) adapters for video and image generation model architectures…
852002active
WhatDreamsCost/WhatDreamsCost-ComfyUI
A collection of free custom ComfyUI nodes and workflows, centered on LTX Director, a timeline-based tool for directing LTX video generation…
572001active
bytetriper/RAE
Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u…
482001active
zai-org/GLM-130B
GLM-130B is an open bilingual (English and Chinese) 130-billion-parameter dense language model pre-trained with the General Language Model …
327651maintenance
apple/coreai-models
Apple's repository of model export recipes, Python primitives, and Swift runtime utilities for building on-device AI with the Core AI frame…
671999active
mil-tokyo/webdnn
WebDNN is a framework for running deep neural network inference directly in the web browser, accepting ONNX models without Python preproces…
651999active
DEIM
DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp…
621999active
jianchang512/vocal-separate
A minimal local web-based tool for separating vocals from background music in audio or video files, using Spleeter 2stems/4stems/5stems mod…
101998active
elyra-ai/elyra
Elyra is a set of AI-centric extensions for JupyterLab, including a visual pipeline editor for building and executing notebook-based pipeli…
681996active
clementchadebec/benchmark_VAE
Pythae is a PyTorch library that unifies implementations of many Variational Autoencoder (VAE) variants under a common interface, enabling …
231995active
facebookresearch/dino
PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th…
107611maintenance
lightseekorg/tokenspeed
TokenSpeed is a high-performance LLM inference engine designed for agentic workloads, aiming for TensorRT-LLM-level performance with vLLM-l…
681989active
Morizeyao/GPT2-Chinese
A Python library providing GPT-2 training and text generation code tailored for Chinese, built on HuggingFace Transformers with BERT or BPE…
327597maintenance
allenai/scispacy
scispaCy is a Python library providing full spaCy pipelines, models, and custom pipes for processing scientific, biomedical, and clinical t…
541988active
Hyperopt
Hyperopt is a Python library for distributed asynchronous hyperparameter optimization over search spaces with real-valued, discrete, and co…
867592maintenance
Lumiwealth/lumibot
Lumibot is a Python framework for building, backtesting, and deploying algorithmic trading strategies and AI-agent trading teams across sto…
951987active
logpai/logparser
Logparser is a Python machine learning toolkit and benchmark suite for automated log parsing. It extracts event templates from unstructured…
341987active
chaidiscovery/chai-lab
Chai-1 is a state-of-the-art multi-modal foundation model for biomolecular structure prediction, handling proteins, small molecules, DNA, R…
651986active
featureform/featureform
Featureform is a virtual feature store that sits atop your existing data infrastructure and orchestrates it to define, manage, and serve ML…
361985active
xingyizhou/CenterNet
CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi…
327573maintenance
hkchengrex/XMem
XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in…
231983stable
kwuking/TimeMixer
Official PyTorch implementation of TimeMixer, an ICLR 2024 model for time series forecasting using decomposable multiscale mixing. It has s…
461981active
lucidrains/titans-pytorch
An unofficial PyTorch implementation of the Titans architecture, a neural long-term memory module for transformers that learns to memorize …
701980active
davda54/sam
An unofficial PyTorch implementation of Sharpness-Aware Minimization (SAM) and its adaptive variant ASAM, provided as an optimizer wrapper …
321980stable
patrikhuber/eos
A lightweight, header-only 3D Morphable Face Model (3DMM) fitting library written in modern C++11/14, with Python bindings. It provides mod…
311980active
google-deepmind/alphagenome
A Python SDK providing programmatic access to Google DeepMind's AlphaGenome model, which predicts genomic regulatory outputs such as gene e…
891979active
cloneofsimo/lora
A Python library for applying Low-Rank Adaptation (LoRA) to quickly fine-tune text-to-image diffusion models like Stable Diffusion. It prod…
227550maintenance
adobe-research/custom-diffusion
Custom Diffusion is a research codebase for efficiently fine-tuning text-to-image diffusion models like Stable Diffusion on a few example i…
691978stable
JIA-Lab-research/DreamOmni2
DreamOmni2 is the official PyTorch implementation of a CVPR 2026 Highlight model for multimodal instruction-based image editing and generat…
511978active
Linzaer/Ultra-Light-Fast-Generic-Face-Detector-1MB
An ultra-lightweight face detection model (~1MB FP32, ~300KB quantized) designed for edge computing devices, with slim and RFB variants tra…
327542maintenance
PrimeIntellect-ai/prime-rl
prime-rl is a Python framework for large-scale, fully asynchronous reinforcement learning training of language models, built on FSDP2 for t…
871975active
haykgrigo3/TimeCapsuleLLM
TimeCapsuleLLM is a research project training language models from scratch (and fine-tuning small base models) exclusively on text from spe…
751975active
ZhuJHua/moodiary
Moodiary is a fully open-source, cross-platform diary/journaling app built with Flutter and Rust. It supports markdown, plain text, and ric…
691975active
tandpfun/wardrobe
A self-hosted web application that detects garments in photos, extracts clean product cutouts, and generates modeled editorial previews usi…
541975active
openlm-research/open_llama
OpenLLaMA is a permissively licensed (Apache-2.0) open reproduction of Meta AI's LLaMA, releasing 3B, 7B, and 13B pretrained model weights …
307528maintenance
showlab/Show-o
Show-o is a research repository implementing a unified transformer model that combines autoregressive and discrete diffusion modeling for m…
501973active
android/androidify
An open-source Android sample app from Google that lets users create custom Android bot avatars using AI image generation via the Gemini AP…
681972active
WassimTenachi/PhySO
PhySO is a Python library for physical symbolic optimization that uses deep reinforcement learning to discover analytical physical laws fro…
521971active
FireRedTeam/FireRedASR
FireRedASR is a family of open-source industrial-grade automatic speech recognition models supporting Mandarin, Chinese dialects, and Engli…
521971active
theamusing/perfectPixel
A Python library that automatically detects the optimal grid size in AI-generated pixel art images and refines them into clean, perfectly a…
451971active
bigcode-project/starcoder
StarCoder is a 15B-parameter code language model trained on 80+ programming languages, and this repository hosts the fine-tuning and infere…
307506maintenance
google-deepmind/tapnet
Google DeepMind's official repository for Tracking Any Point (TAP), containing the TAP-Vid and TAPVid-3D benchmarks, the TAPIR and TAPNext …
741968active
meta-recsys/generative-recommenders
Meta's research library implementing HSTU and M-FALCON from the ICML'24 paper 'Actions Speak Louder than Words: Trillion-Parameter Sequenti…
691966active
Netflix/void-model
VOID (Video Object and Interaction Deletion) is a research model from Netflix that removes objects from videos along with the physical inte…
541965active
siliconflow/onediff
OneDiff is an out-of-the-box acceleration library for diffusion models, providing PyTorch compilation tools and optimized GPU kernels. It i…
481964active
official-pikafish/Pikafish
Pikafish is a free, open-source UCI xiangqi (Chinese chess) engine derived from Stockfish, using NNUE neural network evaluation to analyze …
761963active
NVIDIA-NeMo/RL
NeMo RL is NVIDIA's open-source post-training library for scaling reinforcement learning methods (GRPO, PPO, DPO, SFT, distillation) on LLM…
811961active
OpenMotionLab/MotionGPT
MotionGPT is a unified motion-language model that treats 3D human motion as a foreign language by converting motion into discrete motion to…
321961active
2U1/Qwen-VL-Series-Finetune
An open-source Python repository providing training scripts for fine-tuning Alibaba's Qwen-VL series of vision-language models (Qwen2-VL, Q…
671960active
open-mmlab/mmagic
MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri…
237457maintenance
CannyLab/tsne-cuda
A CUDA-accelerated implementation of the FIt-SNE t-SNE algorithm with Python bindings, offering up to 1200x speedup over scikit-learn. It e…
951957stable
kadirnar/whisper-plus
A Python library wrapping OpenAI Whisper-family models (including distil-whisper and MLX variants) for fast speech-to-text transcription wi…
601956active
SizheAn/PanoHead
PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis…
291956active
alibaba/EasyCV
EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo…
321954active
eriklindernoren/PyTorch-YOLOv3
A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights…
327440maintenance
Fafa-DL/Awesome-Backbones
A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin …
331953active
meta-pytorch/opacus
Opacus is a PyTorch library for training neural networks with differential privacy via DP-SGD, requiring minimal code changes through its P…
791952active
autodiff/autodiff
autodiff is a C++17 library for automatic differentiation that computes derivatives of functions efficiently using forward mode (dual numbe…
241952active
Yuliang-Liu/Monkey
Monkey is a large multi-modal model (LMM) research project from CVPR 2024 that improves image understanding via higher input resolution and…
651951active
haoheliu/versatile_audio_super_resolution
AudioSR is a Python library and CLI tool that performs versatile audio super-resolution, upsampling any audio (music, speech, sound effects…
451951active
LTH14/mar
Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive …
541949stable
openai/guided-diffusion
OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio…
327419maintenance
Vchitect/Latte
Official PyTorch implementation of Latte, a latent diffusion transformer for video generation. It includes model definitions, pre-trained c…
711948active
cra-ros-pkg/robot_localization
robot_localization is a ROS package of nonlinear state estimation nodes (EKF, UKF, and navsat_transform) for fusing IMU, odometry, and othe…
641947stable
kyegomez/BitNet
A PyTorch implementation of the BitNet architecture from the paper 'BitNet: Scaling 1-bit Transformers for Large Language Models', providin…
721945active
Graph Convolutional Networks (GCN)
A TensorFlow implementation of Graph Convolutional Networks (GCN) for semi-supervised node classification on graphs, accompanying the ICLR …
327400maintenance
jd-opensource/JoyAI-Echo
JoyAI-Echo is a Python framework for long-horizon audio-visual generation, producing coherent multi-shot videos up to ~5 minutes with paire…
581943active
tensorlayer/TensorLayer
TensorLayer is a TensorFlow-based deep learning and reinforcement learning library offering customizable neural layers for researchers and …
237381maintenance
PixArt-alpha/PixArt-sigma
PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with…
251939active
starik222/BooruDatasetTagManager
A desktop tag editor for managing booru-style tagged image and video datasets used to train Stable Diffusion models such as LoRAs, embeddin…
761938active
microsoft/Magma
Magma is Microsoft Research's foundation model for multimodal AI agents, released as an 8B vision-language model that understands images an…
531937active
JuliaAI/MLJ.jl
MLJ.jl is a machine learning framework for Julia providing a common interface to over 200 models, with meta-algorithms for model selection,…
931935stable
NVlabs/RADIO
Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro…
641933active
PAIR-code/facets
Facets is a pair of web-component visualizations (Overview and Dive) for understanding and analyzing machine learning datasets, embeddable …
107336maintenance
Tencent-Hunyuan/HunyuanOCR
HunyuanOCR-1.5 is a lightweight end-to-end OCR vision-language model from Tencent, with a unified inference environment, llama.cpp PC-side …
591930active
tum-pbs/PhiFlow
PhiFlow is an open-source Python simulation toolkit for solving partial differential equations with support for optimization and machine le…
721929active
Audio-AGI/AudioSep
AudioSep is the official implementation of the 'Separate Anything You Describe' foundation model for open-domain, language-queried audio so…
281929active
Yuanshi9815/OminiControl
OminiControl is a universal control framework for Diffusion Transformer models like FLUX, supporting subject-driven and spatial control (ed…
621927active
RightNow-AI/picolm
PicoLM is a pure C11 LLM inference engine that runs 1-billion parameter models in GGUF format on extremely constrained hardware like $10 bo…
461921active
OpenTalker/video-retalking
VideoReTalking is a Python research system from SIGGRAPH Asia 2022 that edits real-world talking-head videos to match a given audio track, …
237280maintenance
GPflow/GPflow
GPflow is a Python library for building Gaussian process models on top of TensorFlow 2 and TensorFlow Probability. It implements modern Gau…
951916active
pymatting/pymatting
PyMatting is a Python library for alpha matting that estimates an alpha matte from an input image and a hand-drawn trimap to extract foregr…
671914active
deepseek-ai/DeepSeek-LLM
DeepSeek LLM is a family of open-source large language models (7B and 67B, Base and Chat variants) trained from scratch on 2 trillion token…
277257maintenance
FACEGOOD/FACEGOOD-Audio2Face
FACEGOOD Audio2Face is an open-source deep learning framework that converts audio into facial blendshape weights for driving digital humans…
641909active
diffgram/diffgram
Diffgram is a self-hosted AI datastore for managing schemas, BLOBs, and predictions, with built-in human supervision (data labeling), data …
621909active
sicxu/Deep3DFaceRecon_pytorch
A PyTorch implementation of Deep3DFaceReconstruction, a weakly-supervised CNN method for reconstructing 3D face geometry from a single imag…
321907stable
NVlabs/nvdiffrast
Nvdiffrast is a PyTorch library from NVIDIA providing high-performance, GPU-accelerated primitive operations for rasterization-based differ…
551905stable
visual-layer/fastdup
fastdup is a free Python tool for rapidly analyzing image and video datasets to surface duplicates, outliers, broken, dark, bright, blurry,…
671904active
Faceplugin-ltd/Open-Source-Face-Recognition-SDK
An open-source face recognition SDK by Faceplugin providing face detection, landmark extraction, feature embedding generation, and face tem…
641903active
lucidrains/byol-pytorch
A PyTorch library implementing the Bootstrap Your Own Latent (BYOL) self-supervised learning method from DeepMind. It wraps any image-based…
581903active
google-deepmind/penzai
Penzai is a JAX research toolkit for building, editing, and visualizing neural networks as legible, functional pytree data structures. It i…
101901active
TheSpaghettiDetective/obico-server
Obico Server is the self-hostable backend of the Obico smart 3D printing platform, providing AI-based print failure detection, webcam strea…
761899active
sapientinc/HRM-Text
HRM-Text is a 1B-parameter text generation model based on the hierarchical recurrent HRM architecture, released with a complete pretraining…
531899active
flexflow/flexflow-train
FlexFlow Train is a deep learning framework that accelerates distributed DNN training by automatically searching for efficient parallelizat…
671898active

← prev page 16 / 54 next →