function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| xLLM-AI/xllm xLLM is a high-performance C++ inference engine for LLM, VLM, DiT and recommendation models, optimized for heterogeneous AI accelerators su… | 78 | 1536 | active |
| hi-primus/optimus Optimus is a Python library for agile data preparation that provides a unified API over pandas, Dask, cuDF, Dask-cuDF, Vaex, and PySpark. I… | 23 | 1536 | active |
| ARISE-Initiative/robomimic robomimic is a modular Python framework for robot learning from demonstration, providing standardized demonstration datasets and offline le… | 69 | 1535 | active |
| RightNow-AI/autokernel AutoKernel is an open-source autoresearch pipeline that takes any PyTorch model, profiles it to find GPU kernel bottlenecks, extracts them … | 48 | 1535 | active |
| Shawn-Shan/fawkes Fawkes is a privacy protection tool from University of Chicago researchers that adds imperceptible adversarial perturbations to photos to p… | 23 | 5595 | maintenance |
| K-Dense-AI/karpathy Karpathy is an open-source agentic Machine Learning Engineer that trains state-of-the-art ML models using the Claude Agent SDK and Google A… | 60 | 1532 | active |
| leela-zero/leela-zero Leela Zero is an open-source Go engine that reimplements AlphaGo Zero, combining Monte Carlo Tree Search with a deep residual convolutional… | 23 | 5586 | maintenance |
| AIGODLIKE/ComfyUI-BlenderAI-node A Blender addon that integrates ComfyUI into Blender by converting ComfyUI nodes into Blender nodes, enabling AI image generation, material… | 47 | 1531 | active |
| hustvl/LightningDiT LightningDiT is a research codebase for latent diffusion models implementing VA-VAE and LightningDiT, achieving FID 1.35 on ImageNet-256 wi… | 47 | 1529 | active |
| BennyKok/comfyui-deploy An open-source, Vercel-like deployment platform for ComfyUI workflows, letting teams share workflows, manage machines (on-premise or server… | 40 | 1529 | active |
| luciddreamer-cvlab/LucidDreamer LucidDreamer is the official implementation of a research method that generates 3D Gaussian Splatting scenes from text prompts, published i… | 69 | 1528 | active |
| skforecast/skforecast Skforecast is a Python library for time series forecasting that turns any scikit-learn compatible estimator (LightGBM, XGBoost, CatBoost, K… | 95 | 1527 | active |
| Equim-chan/Mortal Mortal is a free and open-source AI for Japanese (riichi) mahjong powered by deep reinforcement learning, written in Rust with a Python int… | 52 | 1527 | active |
| cchen156/Learning-to-See-in-the-Dark TensorFlow implementation of 'Learning to See in the Dark' (CVPR 2018), a deep learning model that brightens very dark, short-exposure RAW … | 51 | 5565 | maintenance |
| ByteDance-Seed/Triton-distributed Triton-distributed is a distributed compiler built on OpenAI Triton for computation-communication overlapping on multi-GPU systems. It lets… | 63 | 1526 | active |
| interpretml/DiCE DiCE is a Python library that generates diverse counterfactual explanations for any machine learning model, showing feature-perturbed versi… | 37 | 1525 | active |
| TransparentLC/realesrgan-gui A cross-platform graphical interface for the Real-ESRGAN AI image upscaler (with Real-CUGAN support), built in Python with tkinter. It wrap… | 59 | 1524 | active |
| WenmuZhou/PytorchOCR A PyTorch-based OCR toolkit that ports PaddleOCR models to PyTorch, supporting common text detection and recognition algorithms like the PP… | 59 | 1523 | active |
| keras-rl/keras-rl A Python library implementing state-of-the-art deep reinforcement learning algorithms (DQN, DDPG, SARSA, and more) that integrates seamless… | 23 | 5547 | maintenance |
| NirAharon/BoT-SORT BoT-SORT is a state-of-the-art multi-object tracker that combines motion and appearance information with camera motion compensation and an … | 32 | 1522 | active |
| o19s/elasticsearch-learning-to-rank An Elasticsearch plugin that integrates Learning to Rank (machine-learned relevance) into Elasticsearch. It stores feature query templates,… | 90 | 1521 | active |
| Tencent/TFace TFace is a research platform from Tencent Youtu Lab for trusty face analysis, covering face recognition, face security (anti-spoofing), fac… | 57 | 1521 | active |
| anishathalye/neural-style A Python command-line tool implementing the neural style transfer algorithm (Gatys et al.) in TensorFlow, applying the style of one image t… | 67 | 5541 | maintenance |
| espressif/esp-csi Espressif's official examples and tools for Wi-Fi Channel State Information (CSI) on ESP32-series chips, enabling non-contact wireless sens… | 67 | 1520 | active |
| allenzren/open-pi-zero An open-source re-implementation of the pi0 vision-language-action (VLA) model from Physical Intelligence, built on a pre-trained PaliGemma… | 25 | 1518 | active |
| Phantom-video/Phantom Phantom is a subject-consistent video generation model from ByteDance that preserves reference subject identity via cross-modal alignment. … | 39 | 1517 | active |
| microsoft/Mage Mage is a family of lightweight 4B-parameter multimodal models from Microsoft, including Mage-VL for image and video understanding and Mage… | 57 | 1516 | active |
| NVlabs/describe-anything Describe Anything Model (DAM) is a vision-language model that generates detailed descriptions of user-specified regions in images and video… | 32 | 1514 | active |
| bytedance/SALMONN SALMONN is a family of open-source multi-modal large language models from ByteDance and Tsinghua that unify speech, audio, music, and video… | 73 | 1513 | active |
| geomstats/geomstats Geomstats is an open-source Python package for computations, statistics, and machine learning on manifolds with geometric structures. It pr… | 67 | 1513 | active |
| ZFTurbo/Music-Source-Separation-Training A Python training framework for music source separation models, supporting many architectures such as MDX23C, Demucs, Band Split RoFormer, … | 83 | 1512 | active |
| ATH-MaaS/Ovis Ovis is an open-source Multimodal Large Language Model (MLLM) architecture that structurally aligns visual and textual embeddings, with rel… | 65 | 1512 | active |
| google/meridian Meridian is Google's open-source marketing mix modeling (MMM) framework built on Bayesian causal inference, letting advertisers run in-hous… | 70 | 1511 | active |
| decoderesearch/SAELens SAELens is a Python library for training sparse autoencoders (SAEs) on language model activations and analyzing them for mechanistic interp… | 88 | 1510 | active |
| HiDream-ai/HiDream-O1-Image HiDream-O1-Image is an open-weights 8B image generation foundation model built on a Pixel-level Unified Transformer (UiT) that natively enc… | 53 | 1510 | active |
| azuwis/pianotrans A simple GUI wrapper and packaging for ByteDance's Piano Transcription with Pedals, a PyTorch system that converts piano audio recordings i… | 67 | 1509 | active |
| kyutai-labs/hibiki Hibiki is a decoder-only model for streaming (simultaneous) speech-to-speech translation, built on the multistream Moshi architecture. It p… | 28 | 1509 | active |
| imoneoi/openchat OpenChat is a library of open-source large language models fine-tuned with C-RLFT, an offline reinforcement learning strategy that learns f… | 20 | 5490 | maintenance |
| wb14123/seq2seq-couplet A deep learning project that generates Chinese couplets (对联) using a seq2seq model built with TensorFlow. It includes training scripts, a w… | 32 | 5487 | maintenance |
| facebookexperimental/Robyn Robyn is Meta Marketing Science's open-source, semi-automated Marketing Mix Modeling (MMM) package available in R and Python. It uses ridge… | 52 | 1508 | active |
| hkchengrex/Tracking-Anything-with-DEVA DEVA is a decoupled video segmentation framework that combines task-specific image-level segmentation models with a universal bi-directiona… | 27 | 1508 | stable |
| sail-sg/envpool EnvPool is a C++-based batched environment pool with pybind11 bindings and a thread pool for high-performance parallel RL environment execu… | 94 | 1506 | active |
| sepandhaghighi/pycm PyCM is a Python library for computing multi-class confusion matrices and a wide range of per-class and overall evaluation statistics. It a… | 86 | 1506 | stable |
| mbzuai-oryx/Video-ChatGPT Video-ChatGPT is a video conversation model that combines large language models with a pretrained visual encoder adapted for spatiotemporal… | 45 | 1506 | active |
| Flashlight wav2letter++ is Facebook AI Research's end-to-end automatic speech recognition (ASR) toolkit written in C++. It has been consolidated into … | 62 | 5466 | maintenance |
| czczup/ViT-Adapter Official PyTorch implementation of ViT-Adapter, an ICLR 2023 Spotlight paper introducing a pre-training-free adapter that lets plain Vision… | 34 | 1503 | stable |
| charlesq34/pointnet Reference implementation of PointNet, a neural network architecture that directly consumes unordered 3D point clouds for classification and… | 32 | 5459 | maintenance |
| leggedrobotics/ocs2 OCS2 is a C++ toolbox for formulating and solving nonlinear optimal control problems, with an emphasis on real-time Model Predictive Contro… | 65 | 1501 | active |
| meiqua/shape_based_matching A C++ library implementing Halcon-style shape-based matching (equivalent to LINE-MOD) using gradient orientation templates for robust 2D ob… | 32 | 1500 | active |
| TencentARC/MotionCtrl MotionCtrl is the official implementation of a SIGGRAPH 2024 paper providing a unified and flexible motion controller for video generation … | 30 | 1500 | active |
| microsoft/ai-dev-gallery AI Dev Gallery is a Windows application from Microsoft that lets developers explore over 25 interactive samples powered by local AI models … | 66 | 1497 | active |
| QwenAudio/Fun-ASR Fun-ASR is a family of open-source LLM-based end-to-end speech recognition models from Tongyi Lab, covering Chinese, dialects, accents, and… | 81 | 1496 | active |
| uccl-project/uccl UCCL is a high-performance GPU communication library written in C++ that provides collectives (as a drop-in NCCL/RCCL replacement), P2P tra… | 71 | 1496 | active |
| piddnad/DDColor DDColor is the official PyTorch implementation of an ICCV 2023 paper on photo-realistic automatic image colorization using dual decoders an… | 59 | 1496 | active |
| pygod-team/pygod PyGOD is a Python library for graph outlier detection (anomaly detection) built on PyTorch and PyTorch Geometric. It provides 10+ graph-bas… | 23 | 1496 | active |
| mattmireles/gemma-tuner-multimodal A Python tool for LoRA fine-tuning of Gemma 4 and 3n models on text, images, and audio using Apple Silicon's Metal Performance Shaders. It … | 62 | 1495 | active |
| ibab/tensorflow-wavenet A TensorFlow implementation of DeepMind's WaveNet generative neural network architecture for raw audio waveform generation. It provides tra… | 32 | 5428 | maintenance |
| open-thought/reasoning-gym Reasoning Gym is a Python library of procedural dataset generators and algorithmically verifiable reasoning environments for training LLMs … | 66 | 1494 | active |
| WecoAI/aideml AIDE ML is an open-source Python package implementing an LLM-driven agent that uses agentic tree search to autonomously write, debug, and i… | 65 | 1494 | active |
| open-mmlab/mmengine MMEngine is the foundational training engine library for OpenMMLab projects, providing a unified training loop, config system, registry, ho… | 71 | 1492 | active |
| bojone/bert4keras A lightweight, clean reimplementation of BERT and other transformer models (RoBERTa, ALBERT, T5, GPT, ELECTRA, NEZHA) for Keras/tf.keras. I… | 23 | 5415 | maintenance |
| jonaswinkler/paperless-ng Paperless-ng is a self-hosted document management system that scans, OCRs, indexes, and archives physical documents with full-text search a… | 10 | 5414 | maintenance |
| google-deepmind/graph_nets DeepMind's library for building graph networks (graph neural networks) in TensorFlow and Sonnet, based on the 'Relational inductive biases,… | 32 | 5406 | maintenance |
| graphdeco-inria/diff-gaussian-rasterization A CUDA-based differentiable rasterization engine for 3D Gaussian Splatting, used in the SIGGRAPH 2023 paper '3D Gaussian Splatting for Real… | 29 | 1489 | stable |
| rail-berkeley/hil-serl HIL-SERL is a Python library suite for training reinforcement learning policies for precise robotic manipulation using human demonstrations… | 44 | 1487 | active |
| CUT3R/CUT3R CUT3R is the official PyTorch implementation of 'Continuous 3D Perception Model with Persistent State' (CVPR 2025 Oral), a stateful recurre… | 38 | 1486 | active |
| ModelOriented/DALEX DALEX (moDel Agnostic Language for Exploration and eXplanation) is a library for exploring, explaining, and visualizing the behavior of pre… | 64 | 1485 | active |
| Aceinna/gnss-ins-sim A Python library for simulating GNSS/INS integrated navigation systems. It generates reference trajectories and synthetic IMU, GPS, odomete… | 23 | 1485 | active |
| k2-fsa/icefall Icefall is a collection of speech recognition (ASR) and TTS training recipes built on the k2 and lhotse libraries, implemented in Python wi… | 64 | 1482 | active |
| hustvl/DiffusionDrive DiffusionDrive is a truncated diffusion model for real-time end-to-end autonomous driving, released as the official PyTorch implementation … | 44 | 1480 | active |
| salesforce/CodeTF CodeTF is a Python transformer library for code large language models, providing unified interfaces for training, fine-tuning, and inferenc… | 10 | 1480 | active |
| MaxHalford/prince Prince is a Python library for multivariate exploratory data analysis, implementing PCA, CA, MCA, MFA, FAMD, GPA, and PGA with a scikit-lea… | 74 | 1479 | stable |
| Soul-AILab/SoulX-FlashTalk SoulX-FlashTalk is a 14B audio-driven talking avatar model that streams infinite real-time video from a reference image and audio, achievin… | 58 | 1479 | active |
| alxndrTL/mamba.py A simple, readable pure-PyTorch (plus MLX) implementation of the Mamba state-space model architecture with a parallel scan for efficient tr… | 53 | 1478 | active |
| SciSharp/NumSharp NumSharp is a .NET library providing NumPy-shaped N-dimensional arrays with broadcasting, slicing views, dtype-aware math, and runtime-gene… | 92 | 1477 | active |
| Nutlope/napkins Napkins.dev is an open-source web application that turns screenshots or wireframes of website designs into working React + Tailwind code us… | 63 | 1477 | active |
| Om-Alve/smolGPT A minimal pure-PyTorch implementation for training small GPT-style LLMs from scratch, featuring flash attention, RMSNorm, SwiGLU, RoPE, and… | 23 | 1475 | active |
| sb-ai-lab/LightAutoML LightAutoML (LAMA) is a Python framework for automatic machine learning model creation (AutoML) supporting tabular, time series, image, and… | 60 | 1472 | active |
| mit-han-lab/torchsparse TorchSparse is a high-performance PyTorch library for sparse convolution on 3D point clouds, with optimized GPU kernels for both training a… | 26 | 1472 | active |
| reczoo/FuxiCTR FuxiCTR is an open-source Python library for click-through rate (CTR) prediction built on PyTorch and TensorFlow. It offers a configurable,… | 77 | 1470 | active |
| sczhou/Upscale-A-Video Upscale-A-Video is a diffusion-based model for real-world video super-resolution that takes low-resolution videos and text prompts as input… | 26 | 1470 | active |
| Lightning-AI/lightning-thunder Lightning Thunder is a source-to-source deep learning compiler for PyTorch that optimizes models for training and inference. It provides a … | 72 | 1469 | active |
| google/GNM GNM is an open ecosystem of parametric statistical human models and perception stacks from Google, starting with GNM Head, a high-fidelity … | 59 | 1469 | active |
| code-kern-ai/refinery An open-source tool for scaling, assessing, and maintaining natural language training data, treating datasets like software artifacts. It p… | 23 | 1469 | active |
| QuantaAlpha/QuantaAlpha QuantaAlpha is an LLM-driven framework for mining quantitative alpha factors using a trajectory-based self-evolving paradigm. Users describ… | 56 | 1466 | active |
| autonomousvision/mip-splatting Mip-Splatting is a research implementation of alias-free 3D Gaussian Splatting, introducing a 3D smoothing filter and 2D Mip filter to elim… | 27 | 1466 | active |
| hojonathanho/diffusion The official reference implementation of Denoising Diffusion Probabilistic Models (DDPM) from the 2020 paper by Jonathan Ho et al., written… | 32 | 5300 | maintenance |
| Meta-Harness Meta-Harness is a Python framework for automated end-to-end search over task-specific model harnesses — the code surrounding a fixed base L… | 56 | 1464 | active |
| NVIDIA/tacotron2 NVIDIA's PyTorch implementation of the Tacotron 2 text-to-speech model, which synthesizes mel spectrograms from text for vocoder-based audi… | 32 | 5296 | maintenance |
| yunjey/stargan Official PyTorch implementation of StarGAN, a unified generative adversarial network for multi-domain image-to-image translation (CVPR 2018… | 32 | 5296 | maintenance |
| Turing-Project/WriteGPT WriteGPT is a generative text-creation AI framework built on GPT-2 and other models (EAST, CRNN, BERT), fine-tuned to generate Chinese exam… | 23 | 5289 | maintenance |
| FeiYull/TensorRT-Alpha A C++/CUDA library providing TensorRT-accelerated deployment for 30+ popular computer vision models including YOLOv3-v8, YOLOv8-Pose/Seg/Cl… | 32 | 1460 | active |
| supermaven-inc/supermaven-nvim The official Neovim plugin for Supermaven, an AI-powered code completion service. It provides inline suggestions in Neovim with configurabl… | 24 | 1460 | active |
| tensorflow/tpu A collection of reference models and tools for training machine learning models on Google Cloud TPUs, maintained as a public mirror by the … | 72 | 5278 | maintenance |
| facebookresearch/MobileLLM Meta's training code for MobileLLM, a family of sub-billion parameter language models optimized for on-device use, published at ICML 2024. … | 59 | 1459 | active |
| jax-md/jax-md JAX MD is a Python library for molecular dynamics simulations built on JAX, making them hardware accelerated on CPU, GPU, and TPU and end-t… | 86 | 1458 | active |
| robfiras/loco-mujoco LocoMuJoCo is an imitation learning benchmark for whole-body locomotion control built on MuJoCo, featuring humanoid, quadruped, and musculo… | 76 | 1453 | active |
| microsoft/KBLaM Official implementation of KBLaM, a method for augmenting pre-trained LLMs with external knowledge by encoding a knowledge base into contin… | 64 | 1451 | active |
| dleemiller/WordLlama WordLlama is a fast, lightweight Python NLP toolkit built on LLM token embeddings for tasks like similarity computation, ranking, fuzzy ded… | 49 | 1451 | active |
| dbolya/yolact YOLACT is a PyTorch implementation of a fully convolutional model for real-time instance segmentation, accompanying the YOLACT and YOLACT++… | 51 | 5241 | maintenance |