function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| vietanhdev/anylabeling AnyLabeling is a desktop image annotation tool that combines LabelImg/Labelme-style manual labeling with AI-assisted auto-labeling. It runs… | 85 | 3463 | active |
| NVlabs/Eagle Eagle is NVIDIA's family of frontier vision-language models (Eagle, Eagle 2, Eagle 2.5) built with data-centric training strategies, plus L… | 64 | 3462 | active |
| facebookresearch/sam-3d-body SAM 3D Body is a promptable model for single-image full-body 3D human mesh recovery (HMR), estimating body, feet, and hand pose using the M… | 48 | 3461 | active |
| nihui/waifu2x-ncnn-vulkan A portable command-line tool implementing the waifu2x anime-style image upscaler and denoiser using the ncnn inference framework with the V… | 63 | 3456 | active |
| petercorke/robotics-toolbox-python A Python library providing tools for robot manipulator kinematics, dynamics, Jacobians, and trajectory generation, built on NumPy and SciPy… | 99 | 3454 | active |
| PaddlePaddle/PARL PARL is a flexible, high-performance reinforcement learning framework built on PaddlePaddle, providing Model/Algorithm/Agent abstractions a… | 41 | 3453 | active |
| shenweichen/DeepCTR-Torch DeepCTR-Torch is a PyTorch library providing easy-to-use, modular, and extendable implementations of deep-learning-based CTR (click-through… | 78 | 3450 | active |
| ob-f/OpenBot OpenBot is an open-source project that turns Android smartphones into the brains of low-cost robots, paired with a ~$50 electric vehicle bo… | 67 | 3442 | active |
| PennyLaneAI/pennylane PennyLane is a cross-platform Python library for quantum computing, quantum machine learning, and quantum chemistry. It lets users build an… | 92 | 3441 | stable |
| XinJingHao/DRL-Pytorch A unified PyTorch implementation collection of popular deep reinforcement learning algorithms including DQN variants, PPO, DDPG, TD3, SAC, … | 44 | 3436 | active |
| NVlabs/GR00T-WholeBodyControl NVIDIA's unified platform for developing, training, and deploying whole-body controllers for humanoid robots, including the decoupled WBC m… | 61 | 3428 | active |
| NVlabs/stylegan The official TensorFlow implementation of StyleGAN, NVIDIA's style-based generator architecture for generative adversarial networks from th… | 32 | 14416 | maintenance |
| continue-revolution/sd-webui-animatediff An AUTOMATIC1111 Stable Diffusion WebUI extension that integrates AnimateDiff motion modules to generate animated GIFs and videos from diff… | 28 | 3425 | active |
| Qwen3-ASR Qwen3-ASR is a family of open-source speech recognition models from Alibaba's Qwen team, supporting ASR and language identification across … | 55 | 3423 | active |
| aqlaboratory/openfold OpenFold is a faithful, trainable PyTorch reproduction of DeepMind's AlphaFold 2 for protein structure prediction. It is memory-efficient a… | 48 | 3420 | active |
| microsoft/nni NNI (Neural Network Intelligence) is an open-source AutoML toolkit from Microsoft that automates hyperparameter tuning, neural architecture… | 10 | 14361 | maintenance |
| davidsandberg/facenet A TensorFlow implementation of the FaceNet face recognizer that generates 128-dimensional face embeddings, including face detection via MTC… | 32 | 14343 | maintenance |
| embeddings-benchmark/mteb MTEB (Massive Text Embedding Benchmark) is a Python library and CLI for benchmarking embedding models across 1000+ tasks, languages, and mo… | 95 | 3406 | active |
| jsvine/markovify Markovify is a simple, extensible Python library for building Markov chain models from text corpora and generating random semi-plausible se… | 32 | 3403 | stable |
| luigifreda/pyslam pySLAM is a hybrid Python/C++ Visual SLAM pipeline supporting monocular, stereo, and RGB-D cameras with a wide range of local and global fe… | 77 | 3401 | active |
| NovaSky-AI/SkyThought SkyThought is the open-source repository behind Sky-T1, a family of reasoning language models trained for under $450, including training sc… | 25 | 3399 | active |
| IQA-PyTorch A pure Python/PyTorch toolbox for image quality assessment (IQA) providing GPU-accelerated reimplementations of many full-reference and no-… | 82 | 3380 | active |
| deepseek-ai/DeepSeek-OCR-2 DeepSeek-OCR 2 is an open-source vision-language model and inference toolkit implementing 'Visual Causal Flow' for optical character recogn… | 44 | 3379 | active |
| WongKinYiu/yolov7 Official PyTorch implementation of the YOLOv7 paper, a state-of-the-art real-time object detector with trainable bag-of-freebies techniques… | 23 | 14139 | maintenance |
| CompVis/latent-diffusion The official research code and pretrained model zoo for Latent Diffusion Models (LDM), the paper behind Stable Diffusion, enabling high-res… | 32 | 14133 | maintenance |
| nv-tlabs/kimodo Kimodo is NVIDIA's official implementation of a kinematic motion diffusion model trained on 700 hours of motion capture data to generate hi… | 56 | 3365 | active |
| Lagrange-Labs/deep-prove DeepProve is a Rust framework for generating zero-knowledge proofs of neural network inference, including end-to-end proving of full LLM fo… | 60 | 3357 | active |
| libAudioFlux/audioFlux audioFlux is a C-based library with Python bindings for audio and music analysis and feature extraction. It supports dozens of time-frequen… | 53 | 3351 | active |
| VainF/Torch-Pruning Torch-Pruning is a PyTorch framework for structural neural network pruning based on the DepGraph algorithm from CVPR 2023. It automatically… | 50 | 3348 | active |
| shankarpandala/lazypredict Lazy Predict is a Python library that trains dozens of machine learning models with minimal code to quickly identify which algorithms perfo… | 78 | 3347 | active |
| magenta/ddsp DDSP is a Python library of differentiable digital signal processing components (synthesizers, filters, waveshapers) that can be embedded i… | 64 | 3344 | active |
| google-ai-edge/LiteRT LiteRT is Google's successor to TensorFlow Lite, an on-device runtime for high-performance ML and GenAI inference on edge platforms. It pro… | 85 | 3339 | active |
| xenova/whisper-web A browser-based speech recognition app that runs OpenAI's Whisper models entirely client-side using Transformers.js. It transcribes audio w… | 30 | 3338 | active |
| HKUDS/VideoRAG VideoRAG is a retrieval-augmented generation framework for chatting with and understanding extremely long-context videos, using a dual-chan… | 53 | 3337 | active |
| lakehq/sail Sail is an open-source, Rust-native multimodal compute engine that serves as a drop-in replacement for Apache Spark, unifying batch process… | 93 | 3333 | active |
| opengeos/geoai GeoAI is a Python package that integrates artificial intelligence with geospatial data analysis, built on PyTorch, Transformers, and segmen… | 89 | 3327 | active |
| pgmpy/pgmpy pgmpy is a Python library for causal and probabilistic reasoning with graphical models such as Bayesian Networks, Dynamic Bayesian Networks… | 81 | 3318 | stable |
| RKNN-Toolkit2 RKNN-Toolkit2 is Rockchip's SDK for converting trained neural network models into RKNN format and deploying them on Rockchip NPU chips like… | 36 | 3313 | active |
| modelscope/evalscope EvalScope is a Python framework from the ModelScope community for evaluating large language models, vision-language models, embedding model… | 93 | 3312 | active |
| Peterande/D-FINE D-FINE is the official PyTorch implementation of an ICLR 2025 Spotlight paper that redefines the regression task in DETR-style detectors as… | 67 | 3305 | active |
| Farama-Foundation/HighwayEnv HighwayEnv is a collection of Gymnasium environments for autonomous driving and tactical decision-making tasks, covering scenarios like hig… | 94 | 3298 | active |
| microsoft/LoRA loralib is the official PyTorch implementation of LoRA (Low-Rank Adaptation), which fine-tunes large language models by injecting trainable… | 23 | 13767 | maintenance |
| ResearAI/DeepScientist DeepScientist is a local-first autonomous research studio that runs the full scientific research loop on your machine, from baselines and e… | 74 | 3296 | active |
| xiph/opus libopus is the reference implementation of the Opus audio codec, an IETF-standardized (RFC 6716) royalty-free codec for interactive speech … | 79 | 3293 | stable |
| mljar/mljar-supervised MLJAR AutoML is a Python package for automated machine learning on tabular data, providing feature engineering, hyperparameter tuning, mode… | 91 | 3287 | active |
| ltdrdata/ComfyUI-Impact-Pack A custom node pack for ComfyUI that enhances Stable Diffusion image generation through detectors, detailers, upscalers, and pipeline utilit… | 65 | 3283 | active |
| hbb1/2d-gaussian-splatting Official implementation of 2D Gaussian Splatting (2DGS), a SIGGRAPH 2024 method that represents scenes as 2D oriented Gaussian disks for ge… | 69 | 3279 | stable |
| timerring/bilive BILIVE is a Python application that records Bilibili live streams and danmaku 24/7, then automatically renders danmaku and AI-generated sub… | 61 | 3275 | active |
| jixiaozhong/Sonic Sonic is the official PyTorch implementation of the CVPR 2025 paper 'Sonic: Shifting Focus to Global Audio Perception in Portrait Animation… | 49 | 3273 | active |
| thomasahle/sunfish Sunfish is a simple but strong chess engine written in Python in roughly 111 lines of code, using MTD-bi search with piece-square table eva… | 87 | 3271 | active |
| mani-skill/ManiSkill ManiSkill is an open-source GPU-parallelized robotics simulation framework and benchmark built on SAPIEN, focused on manipulation skills. I… | 91 | 3264 | active |
| vladmandic/human Human is a JavaScript/TypeScript library built on TensorFlow.js that combines multiple ML models for 3D face detection and recognition, bod… | 48 | 3264 | active |
| Tencent-Hunyuan/HunyuanImage-3.0 HunyuanImage-3.0 is Tencent's open-source native multimodal model for text-to-image and image-to-image generation, with inference code and … | 57 | 3253 | active |
| MAIF/shapash Shapash is a Python library that makes machine learning models interpretable and understandable through clear visualizations, a webapp, and… | 90 | 3251 | active |
| deepdoctection/deepdoctection deepdoctection is a Python library for Document AI that orchestrates document layout analysis, table recognition, OCR, and document/token c… | 98 | 3248 | active |
| cocktailpeanut/fluxgym FluxGym is a simple web UI for training FLUX LoRA models with low VRAM support (12GB/16GB/20GB). It combines the AI-Toolkit Gradio frontend… | 65 | 3246 | active |
| determined-ai/determined Determined is an open-source deep learning platform that combines distributed training, hyperparameter tuning, experiment tracking, and GPU… | 39 | 3236 | active |
| Beckschen/TransUNet Official PyTorch implementation of TransUNet, a U-Net-style architecture that uses a Vision Transformer encoder for medical image segmentat… | 63 | 3234 | stable |
| mit-han-lab/bevfusion BEVFusion is a PyTorch-based multi-task multi-sensor fusion framework that unifies camera and LiDAR features in a shared bird's-eye view re… | 10 | 3230 | stable |
| Jittor/jittor Jittor is a high-performance deep learning framework from Tsinghua University based on just-in-time (JIT) compilation and meta-operators, w… | 67 | 3229 | active |
| onnx/onnx-tensorrt A C++ parser library and backend that converts ONNX models into TensorRT engines for high-performance GPU inference. It is maintained by NV… | 92 | 3228 | active |
| MisoLabsAI/MisoTTS Miso TTS 8B is an open-source text-to-speech model based on an RVQ Transformer architecture with a Llama 3.2-style 8B backbone, designed fo… | 52 | 3224 | active |
| Mayandev/notion-avatar Notion Avatar Maker is an AI-powered web application for creating Notion-style illustrated avatars from photos or text descriptions, with a… | 71 | 3219 | active |
| MzeroMiko/VMamba VMamba is a PyTorch implementation of a visual state space model (SSM) vision backbone based on Mamba, featuring 2D Selective Scan (SS2D) f… | 21 | 3219 | active |
| kerlomz/captcha_trainer A deep learning training tool for image CAPTCHA recognition built on TensorFlow, using CNN/ResNet/DenseNet backbones with GRU/LSTM recurren… | 55 | 3213 | active |
| jy0205/Pyramid-Flow Pyramid Flow is the official PyTorch implementation of a training-efficient autoregressive video generation model based on pyramidal flow m… | 22 | 3208 | active |
| xianfei/SysMocap SysMocap is a cross-platform, video-driven real-time motion capture system that animates 3D virtual characters from webcam footage. It rend… | 78 | 3199 | active |
| NVIDIA/physicsnemo NVIDIA PhysicsNeMo is an open-source Python deep-learning framework for building, training, fine-tuning, and inferring physics AI models us… | 89 | 3198 | active |
| prs-eth/Marigold Marigold is a family of diffusion-based models and a fine-tuning protocol that adapts pretrained latent diffusion models like Stable Diffus… | 52 | 3198 | active |
| Pointcept Pointcept is a PyTorch-based research codebase for point cloud perception, providing implementations of state-of-the-art 3D scene understan… | 76 | 3196 | active |
| facebookresearch/dinov2 PyTorch implementation and pretrained models for DINOv2, a self-supervised vision transformer method from Meta AI that learns robust visual… | 68 | 13266 | maintenance |
| LeelaChessZero/lc0 Lc0 is an open-source, UCI-compliant chess engine that plays chess using neural networks trained via AlphaZero-style self-play reinforcemen… | 65 | 3193 | active |
| stepfun-ai/Step-Video-T2V Step-Video-T2V is an open-source text-to-video generation model from StepFun, released with inference code and pretrained weights (includin… | 25 | 3187 | active |
| Nerogar/OneTrainer OneTrainer is a GUI and CLI application for fine-tuning diffusion image models, supporting full fine-tuning, LoRA, and embeddings across ma… | 74 | 3184 | active |
| ARM-software/ComputeLibrary Arm's Compute Library is a C++ collection of over 100 low-level machine learning and computer vision functions optimized for Arm Cortex-A/N… | 96 | 3183 | active |
| MiniMax-AI/MiniMax-M1 MiniMax-M1 is an open-weight, large-scale hybrid-attention reasoning language model released by MiniMax under Apache-2.0. The repository pr… | 32 | 3180 | active |
| Rudrabha/Wav2Lip Wav2Lip is the official research code for the ACM Multimedia 2020 paper 'A Lip Sync Expert Is All You Need for Speech to Lip Generation In … | 45 | 13182 | maintenance |
| tslearn-team/tslearn tslearn is a Python machine learning library dedicated to time series analysis, built on numpy and compatible with scikit-learn. It provide… | 88 | 3174 | stable |
| facebookresearch/tribev2 TRIBE v2 is a multimodal deep learning model from Meta AI that predicts fMRI brain responses to naturalistic video, audio, and text stimuli… | 55 | 3172 | active |
| rmurai0610/MASt3R-SLAM MASt3R-SLAM is a real-time monocular dense SLAM system built on the MASt3R two-view 3D reconstruction prior, producing globally consistent … | 43 | 3167 | active |
| nesaorg/nesa Nesa is a privacy-preserving AI inference platform that runs models like Llama, Mistral, and Stable Diffusion with end-to-end encryption us… | 24 | 3167 | active |
| qdrant/fastembed FastEmbed is a lightweight Python library for generating text embeddings using ONNX Runtime instead of PyTorch, requiring no GPU and minima… | 83 | 3166 | active |
| tekaratzas/RustGPT A transformer-based large language model implemented entirely in pure Rust with no external ML frameworks, using only ndarray for matrix op… | 38 | 3157 | active |
| stemrollerapp/stemroller StemRoller is a free open-source desktop app that separates vocals, drums, bass, and other stems from any song with a single click, using F… | 84 | 3156 | active |
| parrt/dtreeviz dtreeviz is a Python library for visualizing decision trees and interpreting tree-based machine learning models. It supports scikit-learn, … | 69 | 3155 | active |
| ali-vilab/VGen VGen is the official repository for a holistic video generation ecosystem built on diffusion models, including the I2VGen-XL cascaded image… | 27 | 3155 | active |
| SciML/DifferentialEquations.jl A Julia suite of high-performance numerical solvers for differential equations, covering ODEs, SDEs, DDEs, DAEs, RODEs, jumps, and (S)PDEs,… | 95 | 3151 | active |
| cleardusk/3DDFA_V2 3DDFA_V2 is the official PyTorch implementation of the ECCV 2020 paper 'Towards Fast, Accurate and Stable 3D Dense Face Alignment'. It regr… | 23 | 3149 | stable |
| megvii-research/NAFNet NAFNet is the official PyTorch implementation of a state-of-the-art image restoration network that removes nonlinear activation functions. … | 32 | 3148 | stable |
| Djdefrag/QualityScaler QualityScaler is a Windows GUI application that uses AI deep-learning models to upscale, enhance, and de-noise images and videos. It is wri… | 95 | 3138 | active |
| scikit-learn-contrib/hdbscan A high-performance Python implementation of the HDBSCAN hierarchical density-based clustering algorithm, part of the scikit-learn-contrib e… | 85 | 3136 | stable |
| modelscope/3D-Speaker 3D-Speaker is an open-source Python toolkit for single- and multi-modal speaker verification, speaker recognition, and speaker diarization,… | 56 | 3121 | active |
| junyanz/CycleGAN A Torch (Lua) implementation of CycleGAN and pix2pix for unpaired image-to-image translation using cycle-consistent adversarial networks. I… | 32 | 12870 | maintenance |
| jina-ai/clip-as-service CLIP-as-service is a low-latency, high-scalability server for embedding images and text into fixed-length vectors using OpenAI's CLIP model… | 23 | 12836 | maintenance |
| MakazhanAlpamys/Soup Soup is a Python CLI that fine-tunes and post-trains LLMs from a single YAML config, supporting 23 methods (SFT, DPO, ORPO, SimPO, KTO, etc… | 77 | 3102 | active |
| ddangelov/Top2Vec Top2Vec is a Python library that learns jointly embedded topic, document, and word vectors for topic modeling and semantic search. It suppo… | 23 | 3102 | stable |
| thuml/Time-Series-Library TSLib is an open-source Python library providing a unified codebase of advanced deep learning models for general time series analysis. It s… | 66 | 12785 | maintenance |
| ridgerchu/matmulfreellm A Python implementation of MatMul-Free LM, a language model architecture that eliminates matrix multiplication operations using ternary wei… | 49 | 3089 | active |
| naver/mast3r MASt3R is the official PyTorch implementation of 'Grounding Image Matching in 3D with MASt3R' (ECCV 2024), a model that performs dense 3D r… | 37 | 3088 | active |
| guillaume-be/rust-bert A Rust-native library providing ready-to-use NLP pipelines and transformer-based models (BERT, DistilBERT, GPT-2, RoBERTa, BART, etc.), por… | 60 | 3076 | active |