function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| acados/acados acados is a modular C library providing fast embedded solvers for nonlinear optimal control problems, designed for real-time applications l… | 93 | 1450 | active |
| zylo117/Yet-Another-EfficientDet-Pytorch A PyTorch re-implementation of Google's EfficientDet object detection model with pretrained weights matching near-SOTA accuracy at real-tim… | 23 | 5238 | maintenance |
| NVIDIA-ISAAC-ROS/isaac_ros_visual_slam Isaac ROS Visual SLAM is a ROS 2 package providing GPU-accelerated visual simultaneous localization and mapping (VSLAM) using stereo visual… | 94 | 1448 | active |
| amdegroot/ssd.pytorch A PyTorch implementation of the Single Shot MultiBox Detector (SSD) object detection model from the 2016 paper by Wei Liu et al. It include… | 32 | 5221 | maintenance |
| zama-ai/concrete-ml Concrete ML is a privacy-preserving machine learning library built on top of Zama's Concrete FHE compiler. It lets data scientists convert … | 69 | 1445 | active |
| ByteDance-Seed/m3-agent M3-Agent is a multimodal agent framework from ByteDance Seed that processes real-time visual and auditory inputs to build entity-centric lo… | 48 | 1445 | active |
| do-mpc/do-mpc do-mpc is an open-source Python toolbox for robust model predictive control (MPC) and moving horizon estimation (MHE) of nonlinear systems.… | 61 | 1443 | active |
| zsyOAOA/InvSR InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff… | 51 | 1443 | active |
| AIM-Harvard/pyradiomics PyRadiomics is an open-source Python package for extracting radiomics features from 2D and 3D medical images and binary masks. It provides … | 45 | 1441 | active |
| ThoughtfulDev/EagleEye EagleEye is a Python-based OSINT tool that identifies social media profiles (Instagram, Facebook, Twitter, YouTube) of a person using face … | 32 | 5197 | maintenance |
| valentinfrlch/ha-llmvision LLM Vision is a Home Assistant integration (installed via HACS) that uses multimodal large language models to analyze images, videos, live … | 90 | 1440 | active |
| voice-cloning-app/Voice-Cloning-App A Python/PyTorch desktop application for cloning and synthesizing human voices from audio datasets. It handles the full pipeline from autom… | 23 | 1440 | active |
| google-deepmind/rlax RLax is a JAX-based library of building blocks for implementing reinforcement learning agents, providing mathematical operations like value… | 83 | 1439 | active |
| Walter0807/MotionBERT Official PyTorch implementation of MotionBERT (ICCV 2023), a unified pretrained model for learning human motion representations from 2D ske… | 65 | 1439 | active |
| MiroShark/MiroShark MiroShark is a universal swarm-intelligence engine that turns any document or scenario into a simulated world: it builds a Neo4j knowledge … | 59 | 1439 | active |
| tensorspace-team/tensorspace TensorSpace is a neural network 3D visualization framework built on TensorFlow.js, Three.js, and Tween.js. It provides Keras-like APIs to b… | 23 | 5191 | maintenance |
| aeon-toolkit/aeon aeon is a scikit-learn compatible Python toolkit for machine learning on time series, covering classification, regression, clustering, fore… | 91 | 1438 | active |
| tianweiy/DMD2 DMD2 is the official PyTorch implementation of Improved Distribution Matching Distillation, a NeurIPS 2024 method that distills diffusion m… | 28 | 1438 | active |
| neuralchen/SimSwap SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It… | 23 | 5188 | maintenance |
| sql-machine-learning/sqlflow SQLFlow is a compiler that extends SQL with AI-oriented syntax (training, prediction, evaluation, explanation, and mathematical programming… | 23 | 5188 | maintenance |
| Topdu/OpenOCR OpenOCR is an open-source Python toolkit for general OCR research and applications, covering text detection and recognition, formula and ta… | 58 | 1437 | active |
| vanloctech/youwee Youwee is a cross-platform desktop GUI for yt-dlp built with Tauri, supporting video downloads from YouTube, TikTok, Instagram, and 1800+ s… | 82 | 1436 | active |
| fan-ziqi/rl_sar A C++ framework for simulation verification and physical deployment of reinforcement learning policies for robots, supporting quadruped, wh… | 79 | 1436 | active |
| ucb-bar/gemmini Gemmini is Berkeley's open-source generator for parameterizable systolic-array DNN hardware accelerators, written in Chisel (Scala) and int… | 63 | 1436 | active |
| SakanaAI/evolutionary-model-merge Official repository for SakanaAI's Evolutionary Model Merge research, providing code and resources to reproduce paper evaluations of models… | 16 | 1436 | active |
| salesforce/CodeGen CodeGen is a family of open-source large language models (350M to 16B parameters) from Salesforce AI Research for program synthesis, genera… | 70 | 5179 | maintenance |
| HongyuanLuke/frequencylaw Official code repository for the paper 'Textual Frequency Law on Large Language Models', implementing TFL, TFD, and CTFT methods for studyi… | 49 | 1435 | active |
| SociallyIneptWeeb/AICoverGen AICoverGen is a Python WebUI (with CLI support) that generates AI song covers by converting vocals in a song to any RVC v2 trained voice, s… | 31 | 1435 | active |
| ymcui/Chinese-ELECTRA Chinese-ELECTRA is a collection of pre-trained Chinese ELECTRA language models released by the HIT-iFLYTEK joint lab, built on the official… | 67 | 1434 | stable |
| NVlabs/Fast-FoundationStereo Fast-FoundationStereo is NVIDIA's official PyTorch implementation of a real-time zero-shot stereo matching model family, accepted to CVPR 2… | 54 | 1432 | active |
| winedarksea/AutoTS AutoTS is a Python library for automated time series forecasting, offering dozens of sklearn-style models (statistical, ML, deep learning) … | 95 | 1430 | active |
| qupath/qupath QuPath is an open-source desktop application for bioimage analysis, aimed especially at digital pathology and whole-slide imaging. It provi… | 79 | 1430 | active |
| jasonmayes/Real-Time-Person-Removal A browser-based demo that removes people from complex video backgrounds in real time using TensorFlow.js. It learns the static background o… | 23 | 5154 | maintenance |
| YanjieZe/3D-Diffusion-Policy 3D Diffusion Policy (DP3) is a visual imitation learning algorithm that combines compact 3D point cloud representations with diffusion poli… | 46 | 1429 | active |
| FireRedTeam/FireRedTTS2 FireRedTTS-2 is a long-form streaming text-to-speech system for multi-speaker dialogue generation, built in PyTorch with a dual-transformer… | 40 | 1428 | active |
| zsyOAOA/ResShift ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi… | 62 | 1427 | active |
| zuruoke/watermark-removal A machine learning tool that removes watermarks from images using deep learning image inpainting, based on Contextual Attention and Gated C… | 83 | 5139 | maintenance |
| kootenpv/whereami A Python CLI tool that uses WiFi access point signal strengths and a scikit-learn RandomForest model to predict your indoor location. It le… | 32 | 5139 | maintenance |
| tianweiy/CausVid CausVid is a research codebase implementing a fast autoregressive video diffusion model distilled from a bidirectional diffusion transforme… | 36 | 1426 | active |
| mdbloice/Augmentor Augmentor is a standalone Python library for image augmentation in machine learning, providing a pipeline of stochastic operations like rot… | 32 | 5133 | maintenance |
| nilearn/nilearn Nilearn is a Python library providing statistical and machine-learning tools for analyzing brain imaging data such as fMRI and MRI volumes … | 88 | 1425 | active |
| microsoft/rStar Microsoft's research repository for rStar2-Agent, a 14B math reasoning model trained with agentic reinforcement learning that autonomously … | 41 | 1425 | active |
| paul-buerkner/brms brms is an R package providing a formula-based interface (similar to lme4) for fitting Bayesian generalized (non-)linear multivariate multi… | 67 | 1424 | stable |
| apple/ml-aim Apple's official repository for AIM (Autoregressive Image Models), providing code and pretrained checkpoints for AIMv1 and AIMv2 large visi… | 42 | 1424 | active |
| deepseek-ai/EPLB EPLB is DeepSeek's open-source Expert Parallelism Load Balancer for Mixture-of-Experts models. It computes balanced expert replication and … | 26 | 1424 | active |
| rzru/nightingale Nightingale is an open-source karaoke application that turns any song in your music library into a karaoke track using neural networks. It … | 81 | 1423 | active |
| NVIDIA/DLSS NVIDIA DLSS is the public SDK repository for NVIDIA's RTX Deep Learning Super Sampling, a neural network that boosts game frame rates and g… | 87 | 1421 | active |
| facebookresearch/vggsfm VGGSfM is a deep learning-based Structure from Motion pipeline from Meta AI and Oxford VGG that recovers camera poses and 3D point clouds f… | 29 | 1421 | active |
| lyst/lightfm LightFM is a Python library implementing hybrid recommendation algorithms that combine collaborative filtering with user and item metadata … | 23 | 5111 | maintenance |
| SagiPolaczek/NeuralSVG Official PyTorch implementation of NeuralSVG, an ICCV 2025 paper that generates layered, editable SVG vector graphics from text prompts. It… | 47 | 1419 | active |
| roshan-research/hazm Hazm is a Python library for natural language processing on Persian (Farsi) text, offering normalization, tokenization, stemming, lemmatiza… | 76 | 1417 | active |
| oracle/tribuo Tribuo is a Java machine learning library from Oracle Labs providing classification, regression, clustering, anomaly detection, and multi-l… | 59 | 1417 | active |
| FIND (Framework for Internal Navigation and Discovery, v2) FIND (Framework for Internal Navigation and Discovery) is a self-hosted server that determines indoor position using WiFi fingerprints from… | 23 | 5098 | maintenance |
| jrzaurin/pytorch-widedeep A PyTorch library for multimodal deep learning that combines tabular data with text and images using Wide and Deep model architectures. It … | 62 | 1416 | active |
| fmind/mlops-python-package A Python package template that provides a production-grade code base with MLOps best practices for building and deploying machine learning … | 91 | 1415 | active |
| InternScience/InternAgent InternAgent-1.5 is a Python-based unified agentic framework for long-horizon autonomous scientific discovery, orchestrating multi-agent sys… | 62 | 1415 | active |
| XPandora/PhysGaussian PhysGaussian is a research library that integrates Material Point Method (MPM) physics simulation with 3D Gaussian Splatting representation… | 55 | 1414 | active |
| argilla-io/argilla Argilla is an open-source collaboration tool for AI engineers and domain experts to build, annotate, and curate high-quality datasets for N… | 79 | 5085 | maintenance |
| affinelayer/pix2pix-tensorflow A TensorFlow implementation of pix2pix, a conditional GAN that learns a mapping from input images to output images. It is a faithful port o… | 32 | 5081 | maintenance |
| CSAILVision/semantic-segmentation-pytorch A PyTorch implementation of semantic segmentation (scene parsing) models for the MIT ADE20K dataset, including pretrained model zoo and tra… | 32 | 5078 | maintenance |
| Lyken17/pytorch-OpCounter THOP (PyTorch-OpCounter) is a Python library that counts the MACs/FLOPs and parameters of PyTorch models. It profiles arbitrary nn.Modules … | 32 | 5078 | maintenance |
| lucidrains/self-rewarding-lm-pytorch A PyTorch library implementing the Self-Rewarding Language Model training framework from MetaAI, along with the SPIN training method. It pr… | 16 | 1411 | active |
| NVlabs/BundleSDF BundleSDF is a CVPR 2023 research implementation from NVIDIA for near real-time 6-DoF pose tracking of unknown rigid objects from monocular… | 65 | 1410 | stable |
| explosion/spacy-transformers A spaCy v3 extension package that provides pipeline components for using pretrained transformer models like BERT, RoBERTa, XLNet, and GPT-2… | 71 | 1409 | stable |
| nv-tlabs/GEN3C GEN3C is NVIDIA's research codebase for a generative video model that achieves precise camera control and temporal 3D consistency using a 3… | 59 | 1409 | active |
| gnobitab/InstaFlow InstaFlow is a one-step text-to-image generation model based on Rectified Flow, enabling ultra-fast Stable Diffusion inference without iter… | 28 | 1409 | active |
| primihub/primihub PrimiHub is an open-source privacy-preserving computing platform built by a team of cryptography experts, supporting secure multi-party com… | 44 | 1408 | active |
| koaning/scikit-lego scikit-lego is a Python library of extra building blocks—custom transformers, models, metrics, and datasets—that are fully compatible with … | 87 | 1407 | active |
| mratsim/Arraymancer Arraymancer is a fast, ergonomic N-dimensional tensor (ndarray) library written in Nim, inspired by NumPy and PyTorch. It provides CPU, CUD… | 61 | 1407 | active |
| OpenBMB/AgentCPM-GUI AgentCPM-GUI is an open-source 8B-parameter on-device GUI agent built on MiniCPM-V that takes Android screenshots as input and autonomously… | 46 | 1407 | active |
| XiaoMi/mace MACE (Mobile AI Compute Engine) is a deep learning inference framework optimized for mobile heterogeneous computing on Android, iOS, Linux … | 23 | 5046 | maintenance |
| dexmal/dexbotic Dexbotic is an open-source PyTorch-based toolbox for developing Vision-Language-Action (VLA) models for embodied intelligence. It unifies p… | 72 | 1403 | active |
| dailenson/SDT Official PyTorch implementation of the CVPR 2023 paper 'Disentangling Writer and Character Styles for Handwriting Generation' (SDT). It gen… | 43 | 1403 | active |
| McGill-NLP/webllama WebLlama is a framework for building Llama-3-powered agents that browse the web by following natural language instructions and dialogue. It… | 59 | 1402 | active |
| alibaba/Logics-Parsing Logics-Parsing is an end-to-end document parsing model from Alibaba that converts document images into structured output using a single mul… | 54 | 1402 | active |
| HugoTini/DeepBump DeepBump is a machine-learning tool that generates normal, height, and curvature maps from single pictures, using a U-Net (MobileNetV2) mod… | 28 | 1399 | active |
| Zejun-Yang/AniPortrait AniPortrait is a Python framework from Tencent that generates photorealistic portrait animations from an audio clip and a reference image, … | 25 | 5021 | maintenance |
| open-gigaai/giga-world-policy GigaWorld-Policy is a World Action Model (WAM) for robot policy learning that jointly models actions and future visual observations during … | 59 | 1398 | active |
| yfeng95/PRNet PRNet is a Python/TensorFlow implementation of the ECCV 2018 Position Map Regression Network for joint 3D face reconstruction and dense ali… | 32 | 5013 | maintenance |
| lucidrains/transfusion-pytorch A PyTorch implementation of Transfusion, MetaAI's approach to predicting the next token and diffusing images with a single multi-modal mode… | 75 | 1395 | active |
| om-ai-lab/OmDet OmDet-Turbo is a PyTorch implementation of a transformer-based open-vocabulary object detection model that detects arbitrary user-defined o… | 57 | 1393 | active |
| wenet-e2e/wespeaker WeSpeaker is a research and production-oriented toolkit for speaker embedding learning, supporting speaker verification, recognition, and d… | 64 | 1392 | active |
| logtd/ComfyUI-Fluxtapoz A set of ComfyUI custom nodes for editing and stylizing images with Flux models, implementing techniques like RF-Inversion, RF-Edit, Firefl… | 23 | 1392 | active |
| SebastienZh/StockTradebyZ A semi-automated stock screening application for China's A-share market that fetches daily K-line data via Tushare, applies quantitative pr… | 51 | 1390 | active |
| ARahim3/mlx-tune A Python library for fine-tuning LLMs, vision-language, audio (TTS/STT), embedding, OCR, and JEPA models natively on Apple Silicon Macs usi… | 75 | 1389 | active |
| zju3dv/street_gaussians Street Gaussians is a research implementation of the ECCV 2024 paper 'Modeling Dynamic Urban Scenes with Gaussian Splatting', which reconst… | 40 | 1388 | active |
| Junyi42/monst3r MonST3R is the official PyTorch implementation of an ICLR 2025 paper that estimates per-timestep geometry (pointmaps) from dynamic videos i… | 36 | 1386 | active |
| cszn/BSRGAN BSRGAN is a PyTorch implementation of a practical degradation model for deep blind image super-resolution, presented at ICCV 2021. It provi… | 32 | 1386 | stable |
| zjunlp/KnowLM KnowLM is an open-source framework for building knowledgeable large language models, covering data processing, pre-training, fine-tuning, k… | 30 | 1386 | active |
| lorey/mlscraper mlscraper is a Python library that automatically extracts structured data from HTML pages using machine learning. Instead of writing CSS se… | 23 | 1385 | active |
| Denys88/rl_games RL Games is a high-performance reinforcement learning library built on PyTorch, focused on training agents in massively parallel GPU-based … | 78 | 1383 | active |
| PyO3/rust-numpy Rust bindings for the NumPy C-API built on PyO3, enabling zero-copy conversion between Rust ndarray types and NumPy arrays. It supports wri… | 91 | 1382 | active |
| OpenGVLab/DragGAN An unofficial full-featured Python implementation of DragGAN, the interactive point-based image manipulation method built on StyleGAN gener… | 29 | 4946 | maintenance |
| autonomousvision/unimatch UniMatch is a PyTorch research library implementing a unified transformer-based model for optical flow, stereo matching, and depth estimati… | 32 | 1379 | stable |
| yanx27/Pointnet_Pointnet2_pytorch A pure PyTorch implementation of the PointNet and PointNet++ deep learning architectures for point cloud processing. It includes training a… | 32 | 4936 | maintenance |
| AlmondGod/tinyworlds A minimal Python implementation of DeepMind's Genie autoregressive world model, including a video tokenizer, action tokenizer, and dynamics… | 54 | 1378 | active |
| QwenAudio/ThinkSound ThinkSound is a PyTorch implementation of a NeurIPS 2025 framework that generates and edits audio from video, text, or audio inputs using C… | 52 | 1378 | active |
| siyuanliii/masa Official PyTorch implementation of MASA (CVPR 2024 Highlight), a universal instance appearance model that learns to match any objects acros… | 33 | 1377 | active |
| keyu-tian/SparK SparK is the official PyTorch implementation of an ICLR 2023 Spotlight paper that applies BERT/MAE-style masked image modeling to convoluti… | 22 | 1376 | stable |
| qubvel/segmentation_models A Python library providing neural network architectures for image segmentation (Unet, FPN, Linknet, PSPNet) built on Keras and TensorFlow K… | 23 | 4923 | maintenance |