function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| pq-yang/MatAnyone MatAnyone is a CVPR 2025 human video matting framework that extracts alpha mattes of target people from video using consistent memory propa… | 54 | 1605 | active |
| meituan/YOLOv6 YOLOv6 is a single-stage object detection framework implemented in PyTorch, designed for industrial applications with a family of pretraine… | 23 | 5895 | maintenance |
| apache/opennlp Apache OpenNLP is a machine learning based Java toolkit for processing natural language text, supporting tasks like tokenization, sentence … | 97 | 1602 | stable |
| THUDM/WebGLM WebGLM is an efficient web-enhanced question answering system (KDD 2023) that combines a large language model with web search retrieval and… | 35 | 1602 | active |
| ml4a/ml4a ml4a is a Python library and collection of Jupyter notebooks for making art with machine learning. It wraps popular deep learning models li… | 32 | 1602 | active |
| JustGlowing/minisom MiniSom is a minimalistic, NumPy-based Python implementation of Self Organizing Maps (SOM), a type of neural network for unsupervised learn… | 75 | 1601 | stable |
| time-series-foundation-models/lag-llama Lag-Llama is the first open-source foundation model for probabilistic time series forecasting, built on a transformer architecture. It prov… | 37 | 1601 | active |
| bilibili/ailab Bilibili's AI lab repository, best known for Real-CUGAN, a deep learning model for anime image super-resolution (upscaling). It provides pr… | 23 | 5881 | maintenance |
| autonomousvision/transfuser Official PyTorch implementation of TransFuser, a transformer-based multi-modal sensor fusion model for end-to-end autonomous driving, publi… | 53 | 1599 | stable |
| X-LANCE/AniTalker AniTalker is the official PyTorch implementation of an ACM MM 2024 paper that animates a single static portrait into a vivid talking-face v… | 24 | 1598 | active |
| dfm/emcee emcee is a stable, well-tested Python implementation of the affine-invariant ensemble sampler for Markov chain Monte Carlo (MCMC) proposed … | 67 | 1597 | stable |
| SonyResearch/micro_diffusion Official implementation of Sony Research's micro-budget approach to training large-scale text-to-image diffusion transformer models from sc… | 24 | 1594 | active |
| UbiquitousLearning/mllm MLLM is a fast, lightweight multimodal LLM inference engine written in C++ for mobile and edge devices, with backends for ARM CPU, Qualcomm… | 73 | 1593 | active |
| facebookresearch/fast3r Fast3R is the official PyTorch implementation of a CVPR 2025 model from Meta FAIR that reconstructs 3D scenes and estimates camera poses fr… | 10 | 1593 | active |
| tphakala/birdnet-go BirdNET-Go is a self-hosted, 24/7 realtime soundscape analyser that classifies birds, bats, and other wildlife using local AI models like B… | 94 | 1592 | active |
| google/uncertainty-baselines A library of high-quality, minimal-dependency implementations of standard and state-of-the-art uncertainty and robustness methods for deep … | 77 | 1592 | active |
| alibaba/Pai-Megatron-Patch Pai-Megatron-Patch is an open-source deep learning training toolkit from Alibaba Cloud for large-scale training and inference of LLMs and V… | 56 | 1591 | active |
| spotify/voyager Voyager is an in-memory approximate nearest-neighbor search library implementing the HNSW algorithm, with bindings for Python and Java (and… | 55 | 1591 | stable |
| Tencent-Hunyuan/HunyuanWorld-Voyager HunyuanWorld-Voyager is a video diffusion framework from Tencent Hunyuan that generates world-consistent RGBD video and 3D point-cloud sequ… | 52 | 1590 | active |
| samim23/polymath Polymath is a Python CLI tool that uses machine learning to convert any music library into a searchable music production sample-library. It… | 31 | 1589 | active |
| intel/auto-round AutoRound is Intel's advanced quantization toolkit for LLMs and vision-language models, using sign-gradient descent to achieve high accurac… | 91 | 1588 | active |
| microsoft/Semi-supervised-learning USB (Unified Semi-supervised learning Benchmark) is a PyTorch-based codebase from Microsoft for semi-supervised learning across computer vi… | 65 | 1588 | active |
| databricks/megablocks MegaBlocks is a lightweight Python library for efficient training of mixture-of-experts (MoE) models, built around its dropless-MoE (dMoE) … | 62 | 1588 | active |
| MoonshotAI/Kimi-Linear Kimi Linear is a hybrid linear attention architecture (Kimi Delta Attention, based on Gated DeltaNet) released by Moonshot AI with 48B-para… | 40 | 1588 | active |
| FoundationVision/Infinity Infinity is a bitwise autoregressive text-to-image generation model (CVPR 2025 Oral) with released training and inference code, checkpoints… | 56 | 1587 | active |
| Drexubery/ViewCrafter ViewCrafter is a research codebase that uses video diffusion models to synthesize high-fidelity novel views of scenes from a single or spar… | 49 | 1587 | active |
| zju3dv/EasyVolcap EasyVolcap is a PyTorch-based library for accelerating neural volumetric video research, covering volumetric video capture, reconstruction,… | 27 | 1587 | active |
| DeepGraphLearning/torchdrug TorchDrug is a PyTorch-based machine learning library for drug discovery, covering graph neural networks, deep generative models, and reinf… | 23 | 1587 | active |
| yakhyo/uniface UniFace is a unified Python library for face analysis that bundles detection, recognition, landmark localization, face parsing, gaze estima… | 88 | 1586 | active |
| Tencent-Hunyuan/HY-WorldPlay HY-WorldPlay (HY-World 1.5) is Tencent Hunyuan's open-source framework for interactive 3D world modeling, generating explorable 3D scenes f… | 55 | 1586 | active |
| royshil/obs-localvocal LocalVocal is an OBS Studio plugin that performs real-time, fully local speech recognition and translation using Whisper models running on … | 81 | 1585 | active |
| pytorch/FBGEMM FBGEMM is a collection of highly optimized low-precision matrix multiplication and convolution kernels for server-side deep learning infere… | 93 | 1584 | active |
| scikit-learn-contrib/MAPIE MAPIE is a scikit-learn-compatible Python library for quantifying uncertainty in machine learning models via conformal prediction. It compu… | 93 | 1584 | active |
| shibing624/similarity A Java toolkit for computing text similarity at word, phrase, sentence, and paragraph levels, offering algorithms like Cilin-based similari… | 64 | 1584 | active |
| XPixelGroup/HAT HAT (Hybrid Attention Transformer) is a PyTorch implementation of a state-of-the-art transformer model for image super-resolution and resto… | 32 | 1583 | stable |
| microsoft/MMdnn MMdnn is a Microsoft toolkit for converting, visualizing, and diagnosing deep learning models across frameworks such as TensorFlow, PyTorch… | 39 | 5805 | maintenance |
| meta-pytorch/torchtune Torchtune is a PyTorch-native library for authoring, post-training, and experimenting with large language models. It provides hackable trai… | 69 | 5801 | maintenance |
| pangxiaobin/image-matting A free open-source desktop AI image tool built with pywebview and Vue that performs local background removal (matting) using the RMBG-1.4 m… | 86 | 1580 | active |
| AlibabaResearch/DAMO-ConvAI The official codebase for Alibaba DAMO Academy's Conversational AI research, containing implementations of models from their published pape… | 71 | 1580 | active |
| SalesforceAIResearch/uni2ts Uni2TS is a PyTorch library for unified pre-training, fine-tuning, inference, and evaluation of universal time series forecasting transform… | 60 | 1580 | active |
| MatrAIx-ai/MatrAIx-Persona-8B MatrAIx is a population-scale, persona-driven evaluation framework that instantiates sampled persona records as LLM agents to simulate hete… | 57 | 1580 | active |
| kotaro-kinoshita/yomitoku YomiToku is an AI-powered document image analysis engine specialized for Japanese, providing full-text OCR, layout analysis, table structur… | 87 | 1579 | active |
| RWKV/rwkv.cpp A C++ port of the RWKV language model to the ggml tensor library, providing FP32, FP16, and quantized INT4/INT5/INT8 inference focused on C… | 37 | 1579 | active |
| tensorflow/model-optimization The TensorFlow Model Optimization Toolkit (tfmot) is a Python library providing tools to optimize machine learning models for deployment, i… | 82 | 1578 | stable |
| capitalone/DataProfiler DataProfiler is a Python library that loads CSV, AVRO, Parquet, JSON, text, or URL data into a pandas-compatible DataFrame and profiles it … | 79 | 1578 | active |
| CodeReclaimers/neat-python A pure-Python implementation of NEAT (NeuroEvolution of Augmenting Topologies), an algorithm for evolving neural network topologies and wei… | 79 | 1578 | active |
| menyifang/MIMO MIMO is the official PyTorch implementation of a CVPR 2025 paper on controllable character video synthesis using spatially decomposed model… | 34 | 1578 | active |
| DiffEqML/torchdyn Torchdyn is a PyTorch library dedicated to numerical deep learning, providing tools for neural differential equations, implicit models, and… | 23 | 1578 | active |
| NVlabs/sionna Sionna is an open-source, GPU-accelerated, differentiable Python library from NVIDIA for research on communication systems. It comprises Si… | 83 | 1577 | active |
| stanfordnlp/pyreft pyreft is Stanford NLP's Python library for Representation Finetuning (ReFT), which adapts frozen language models by learning task-specific… | 54 | 1577 | active |
| edbeeching/godot_rl_agents Godot RL Agents is an open-source Python package that bridges games built in the Godot Engine with reinforcement learning algorithms, enabl… | 64 | 1575 | active |
| BloodAxe/pytorch-toolbelt A Python library of PyTorch extensions providing building blocks for fast R&D prototyping, including encoder-decoder architectures, special… | 44 | 1574 | active |
| Tencent/DepthCrafter DepthCrafter is a diffusion-based video depth estimation model from Tencent AI Lab that generates temporally consistent long depth sequence… | 38 | 1574 | active |
| photosynthesis-team/piq PyTorch Image Quality (PIQ) is a collection of measures and metrics for image quality assessment in image-to-image tasks, written in pure P… | 23 | 1574 | stable |
| ArztSamuel/Applying_EANNs A 2D Unity simulation where cars learn to navigate courses using a feedforward neural network trained by a modified genetic algorithm. It s… | 43 | 1571 | stable |
| gcorso/DiffDock DiffDock is a deep learning implementation of a diffusion generative model for molecular docking, predicting how small molecule ligands bin… | 31 | 1569 | active |
| FluxML/Zygote.jl Zygote.jl is a source-to-source automatic differentiation library for Julia that hooks into the Julia compiler to generate gradient (backwa… | 97 | 1568 | active |
| Babyhamsta/Aimmy Aimmy is a universal AI-based aim alignment mechanism (aim assist) for gamers with impairments, built in C# using YOLOv8 models run via ONN… | 78 | 1568 | active |
| Xilinx/brevitas Brevitas is a PyTorch library for neural network quantization supporting both post-training quantization (PTQ) and quantization-aware train… | 91 | 1567 | active |
| Chinese-Text-Classification-Pytorch A PyTorch-based collection of ready-to-run Chinese text classification models including TextCNN, TextRNN, FastText, TextRCNN, BiLSTM with a… | 32 | 5728 | maintenance |
| sdv-dev/CTGAN CTGAN is a Python library of deep learning based synthetic data generators for single-table tabular data, implementing the CTGAN conditiona… | 80 | 1562 | active |
| IDEA-Research/Rex-Omni Rex-Omni is a 3B-parameter multimodal large language model that unifies object detection, OCR, pointing, keypoint detection, and visual pro… | 47 | 1561 | active |
| evo-design/evo Evo is a 7-billion-parameter DNA foundation model built on the StripedHyena architecture, trained on ~300 billion tokens of prokaryotic who… | 71 | 1560 | active |
| real-stanford/universal_manipulation_interface Universal Manipulation Interface (UMI) is a data collection and policy learning framework that transfers in-the-wild human demonstrations i… | 63 | 1560 | active |
| yeemachine/kalidokit KalidoKit is a TypeScript library that converts 3D landmark outputs from Mediapipe/Tensorflow.js face, pose, and hand tracking models into … | 49 | 5699 | maintenance |
| newaetech/chipwhisperer ChipWhisperer is an open-source toolchain for hardware security research, providing capture hardware designs, FPGA/USB firmware, and a Pyth… | 79 | 1557 | active |
| google-deepmind/bsuite bsuite (Behaviour Suite for Reinforcement Learning) is a collection of carefully-designed experiments from DeepMind that investigate core c… | 82 | 1555 | stable |
| bashtage/arch A Python library for financial econometrics providing ARCH/GARCH volatility models, unit root tests, cointegration analysis, bootstrapping,… | 71 | 1555 | stable |
| kritiksoman/GIMP-ML GIMP-ML is a set of Python plugins that bring computer vision and deep learning models into the GNU Image Manipulation Program (GIMP). It p… | 23 | 1553 | active |
| ZeyueT/AudioX AudioX is a unified multimodal framework for anything-to-audio generation, producing audio and music conditioned on text, video, image, or … | 52 | 1552 | active |
| tianrun-chen/SAM-Adapter-PyTorch A PyTorch library that adapts Meta AI's Segment Anything Model (SAM, SAM2, SAM3) to underperforming downstream segmentation tasks using lig… | 67 | 1551 | active |
| Enemyx-net/VibeVoice-ComfyUI A ComfyUI custom node integration for Microsoft's VibeVoice text-to-speech model, providing single and multi-speaker voice synthesis with v… | 53 | 1549 | active |
| kijai/ComfyUI-CogVideoXWrapper A ComfyUI custom node wrapper for CogVideoX and related video generation models (including Fun variants, CogVideoX 1.5, and Go-with-the-Flo… | 39 | 1549 | active |
| nubank/fklearn fklearn is a Python machine learning library from Nubank that applies functional programming principles to build, validate, and deploy mode… | 89 | 1547 | active |
| Tencent/AngelSlim AngelSlim is a Python toolkit from Tencent for compressing large language models and related architectures (VLMs, diffusion, audio models) … | 72 | 1547 | active |
| mosaicml/streaming StreamingDataset is a Python library from MosaicML for fast, accurate streaming of training data from cloud object storage (S3, GCS, Azure,… | 70 | 1547 | active |
| baaivision/Emu3.5 Emu3.5 is BAAI's native multimodal foundation model that jointly predicts next states across vision and language, trained on 10T+ interleav… | 43 | 1547 | active |
| Memento-Teams/Memento-Skills Memento-Skills is a generalist, continually-learnable LLM agent framework where reusable skills stored as markdown files serve as persisten… | 77 | 1546 | active |
| baijiuyang/collision-avoidance2 A computational modeling framework for pedestrian moving-obstacle avoidance behavior, implementing Fajen-style steering and Cohen-style avo… | 73 | 1546 | active |
| lucidrains/soundstorm-pytorch A PyTorch implementation of SoundStorm, Google DeepMind's efficient parallel audio generation model that applies MaskGiT-style masked gener… | 39 | 1546 | active |
| baichuan-inc/Baichuan-7B Baichuan-7B is an open-source, commercially usable 7-billion-parameter pretrained language model built on the Transformer architecture, tra… | 29 | 5649 | maintenance |
| KEV0143/Adaptive-forecasting-of-electricity-consumption-24-168-720-h-with-load-regime-conditioning An open-source tool for adaptive forecasting of hourly electricity consumption over 24, 168, and 720 hour horizons, using load regime condi… | 59 | 1545 | active |
| combust/mleap MLeap is a serialization format (Bundle.ML) and portable execution engine for machine learning pipelines, implemented in Scala with Python … | 95 | 1543 | active |
| ShineChen1024/MagicClothing Official PyTorch implementation of Magic Clothing, a diffusion-based model for controllable garment-driven image synthesis (virtual try-on)… | 26 | 1543 | active |
| facebookresearch/mmf MMF is a modular PyTorch framework for vision and language multimodal research from Facebook AI Research. It ships reference implementation… | 64 | 5633 | maintenance |
| google-ai-edge/model-explorer Model Explorer is a model graph visualizer and debugger from Google AI Edge that renders neural network graphs hierarchically with expandab… | 93 | 1542 | active |
| vibevoice-community/VibeVoice VibeVoice is a community-maintained fork of Microsoft's long-form conversational text-to-speech model, generating expressive multi-speaker … | 60 | 1542 | active |
| hustvl/MapTR MapTR is an end-to-end transformer-based framework for online vectorized HD map construction from camera imagery in autonomous driving. It … | 27 | 1542 | active |
| DrTimothyAldenDavis/SuiteSparse SuiteSparse is a collection of sparse matrix algorithm libraries written in C/C++, including factorization packages (UMFPACK, CHOLMOD, SPQR… | 98 | 1541 | stable |
| VNCCS VNCCS is a ComfyUI custom node suite providing an end-to-end pipeline for generating visual novel character sprites with consistent appeara… | 83 | 1541 | active |
| tensorflow/gnn TensorFlow GNN is a Python library for building Graph Neural Networks on TensorFlow, including a GraphTensor type for heterogeneous graphs,… | 66 | 1541 | active |
| KruxAI/ragbuilder RagBuilder is a Python toolkit that automatically builds an optimal, production-ready Retrieval-Augmented Generation (RAG) pipeline for you… | 35 | 1541 | active |
| RLHFlow/RLHF-Reward-Modeling A collection of training recipes for reward models used in RLHF, covering Bradley-Terry reward models, pairwise preference models, ArmoRM, … | 33 | 1541 | active |
| sobri909/LocoKit LocoKit is a Swift framework for iOS that combines Core Location and Core Motion recording with machine learning based activity type detect… | 29 | 1541 | active |
| lucidrains/DALLE-pytorch A PyTorch implementation/replication of OpenAI's DALL-E, a text-to-image transformer, including a discrete VAE and optional CLIP for rankin… | 23 | 5627 | maintenance |
| MoonshotAI/Moonlight Moonlight is a 3B/16B Mixture-of-Experts LLM trained with the Muon optimizer, released by Moonshot AI along with a memory- and communicatio… | 36 | 1540 | active |
| alibaba/FederatedScope FederatedScope is a comprehensive federated learning platform built on PyTorch with an event-driven architecture. It provides convenient us… | 23 | 1540 | active |
| ATH-MaaS/Marco-o1 Marco-o1 is an open large reasoning model from Alibaba International Digital Commerce, designed for o1-like chain-of-thought reasoning acro… | 61 | 1538 | active |
| pipecat-ai/smart-turn Smart Turn is an open-source, audio-native turn detection model that decides when a voice agent should respond to human speech, using proso… | 49 | 1538 | active |
| Arthur151/ROMP ROMP is a PyTorch-based library and pip-installable API (simple-romp) for real-time monocular multi-person 3D human mesh recovery, implemen… | 23 | 1538 | stable |