function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| stared/livelossplot A Python library that draws live training loss and metric plots inside Jupyter Notebooks for Keras, PyTorch, and other deep learning framew… | 91 | 1319 | stable |
| sanchit-gandhi/whisper-jax An optimized JAX implementation of OpenAI's Whisper speech recognition model, built on Hugging Face Transformers, offering up to 70x faster… | 30 | 4682 | maintenance |
| open-edge-platform/geti Geti is an open-source, end-to-end Vision AI application from Intel that takes users from raw images to deployed computer vision models, ru… | 98 | 1317 | active |
| lean-dojo/LeanCopilot Lean Copilot integrates large language models natively into the Lean 4 theorem prover for proof automation. It provides tactic suggestion, … | 90 | 1316 | active |
| jishengpeng/WavTokenizer WavTokenizer is a state-of-the-art discrete neural audio codec that compresses speech, music, and general audio into only 40 or 75 discrete… | 27 | 1316 | active |
| peterdsharpe/AeroSandbox AeroSandbox is a Python library for designing and optimizing aircraft and other engineered systems using automatic differentiation. It prov… | 73 | 1315 | active |
| sicara/easy-few-shot-learning A Python library (easyfsl) with ready-to-use code and tutorial notebooks for few-shot image classification and meta-learning, built on PyTo… | 23 | 1313 | stable |
| yeyupiaoling/VoiceprintRecognition-Pytorch A PyTorch-based voiceprint recognition (speaker recognition) framework implementing models such as ECAPA-TDNN, ResNetSE, ERes2Net, and CAM+… | 58 | 1312 | active |
| ali-vilab/MimicBrush MimicBrush is the official implementation of a zero-shot image editing method that lets users mask a region in a source image and provide a… | 24 | 1311 | active |
| kwai/DouZero DouZero is a deep reinforcement learning framework that masters the Chinese card game DouDizhu through self-play, published at ICML 2021 by… | 23 | 4652 | maintenance |
| soda-inria/tabicl TabICLv2 is an open-source tabular foundation model that performs classification and regression via a single forward pass through a pre-tra… | 81 | 1309 | active |
| albermax/innvestigate iNNvestigate is a Python toolbox providing a common interface and out-of-the-box implementations of many neural network explanation methods… | 30 | 1309 | active |
| ChenmienTan/RL2 RL2 (Ray Less Reinforcement Learning) is a concise Python library for post-training large language models with reinforcement learning, SFT,… | 57 | 1307 | active |
| Alibaba-NLP/ZeroSearch ZeroSearch is a reinforcement learning framework from Alibaba's Tongyi Lab that trains LLMs to use search by simulating search engine resul… | 36 | 1307 | active |
| derrian-distro/LoRA_Easy_Training_Scripts A PySide6 desktop GUI that wraps Kohya's sd-scripts to simplify training LoRA, LoCon, and other LoRA-type models for Stable Diffusion. It s… | 33 | 1307 | active |
| huawei-noah/Efficient-Computing A collection of efficient deep learning methods from Huawei Noah's Ark Lab, covering model compression, knowledge distillation, pruning, qu… | 32 | 1307 | active |
| DAMO-NLP-SG/VideoLLaMA2 VideoLLaMA 2 is an open-source video large language model that adds spatial-temporal modeling and audio understanding to video-LLMs. It pro… | 25 | 1307 | active |
| seetaface/SeetaFaceEngine SeetaFace Engine is an open-source C++ face recognition engine comprising face detection, face alignment, and face identification modules. … | 32 | 4636 | maintenance |
| inference-labs-inc/JSTprove JSTprove is a Rust CLI toolkit that generates zero-knowledge proofs of machine learning inference on ONNX models, built on Polyhedra Networ… | 70 | 1306 | active |
| unitreerobotics/unitree_rl_lab A collection of reinforcement learning environments for Unitree robots (Go2, H1, G1) built on NVIDIA IsaacLab. It enables training locomoti… | 56 | 1306 | active |
| Albert-Weasker/niubi_guard An open-source defense system that protects GitHub repositories from spam, harassment, and coordinated abuse via configurable detection sig… | 65 | 1305 | active |
| Vincentqyw/image-matching-webui A Gradio-based web UI that matches keypoints between two images using many state-of-the-art image matching algorithms (LoFTR, SuperGlue, Li… | 91 | 1302 | active |
| arkflow-rs/arkflow ArkFlow is a high-performance stream processing engine written in Rust on top of Tokio, connecting configurable inputs (Kafka, MQTT, HTTP, … | 70 | 1302 | active |
| vye16/shape-of-motion Shape of Motion is a Python research codebase for 4D reconstruction of dynamic scenes from a single monocular video, based on the ICCV 2025… | 30 | 1302 | active |
| luosiallen/latent-consistency-model Official implementation of Latent Consistency Models (LCM), a diffusion-based approach for synthesizing high-resolution images with few-ste… | 27 | 4615 | maintenance |
| openai/simple-evals A lightweight Python library from OpenAI for evaluating language models against benchmarks like MMLU, GPQA, MATH, HumanEval, SimpleQA, Heal… | 60 | 4612 | maintenance |
| OpenTeleVision/TeleVision Open-TeleVision is an open-source immersive robot teleoperation system that streams stereoscopic visual feedback to VR headsets (Apple Visi… | 24 | 1301 | active |
| turboderp-org/exllamav2 ExLlamaV2 is a fast Python inference library for running large language models locally on modern consumer GPUs, with support for EXL2/GPTQ … | 61 | 4611 | maintenance |
| SakanaAI/text-to-lora Text-to-LoRA (T2L) is a hypernetwork that generates LoRA adapters for large language models in a single forward pass, using only a natural … | 30 | 1300 | active |
| okfn-brasil/serenata-de-amor Operação Serenata de Amor is an open-source data science project that uses machine learning to audit Brazilian congresspeople's public expe… | 32 | 4603 | maintenance |
| STVIR/pysot PySOT is a Python research platform by SenseTime for single object visual tracking, implementing algorithms such as SiamRPN, SiamRPN++, DaS… | 45 | 4600 | maintenance |
| Uminosachi/sd-webui-inpaint-anything A Stable Diffusion Web UI extension that performs inpainting and outpainting using masks generated by Segment Anything models (SAM 2, SAM-H… | 30 | 1298 | active |
| deepseek-ai/DeepSeek-Prover-V2 DeepSeek-Prover-V2 is an open-source large language model for formal theorem proving in Lean 4, trained via reinforcement learning with sub… | 34 | 1297 | active |
| bambinos/bambi Bambi is a high-level Bayesian model-building interface for Python built on top of the PyMC probabilistic programming framework. It uses a … | 93 | 1296 | active |
| donydchen/mvsplat MVSplat is a PyTorch implementation of an ECCV 2024 Oral model that predicts 3D Gaussians from sparse multi-view images in a single feed-fo… | 61 | 1296 | active |
| robodhruv/visualnav-transformer Official code and pre-trained checkpoints for the GNM, ViNT, and NoMaD family of general-purpose goal-conditioned visual navigation policie… | 19 | 1294 | active |
| braindecode/braindecode Braindecode is an open-source Python toolbox built on PyTorch for decoding raw electrophysiological brain signals such as EEG, ECoG, and ME… | 98 | 1293 | active |
| W2GenAI-Lab/LucidFlux LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a… | 55 | 1293 | active |
| Parskatt/RoMa RoMa (romatch) is a Python library for robust dense feature matching between image pairs, estimating pixel-dense warps and reliable certain… | 52 | 1293 | active |
| PantoMatrix/PantoMatrix PantoMatrix is an open-source research project that generates 3D face and body animation from speech audio, including the EMAGE model for h… | 33 | 1293 | active |
| plaidml/plaidml PlaidML is a portable tensor compiler that enables deep learning on hardware (especially GPUs and embedded devices) not well supported by m… | 10 | 4566 | maintenance |
| Farama-Foundation/Minari Minari is a Python library providing a standard format for offline reinforcement learning datasets, with popular reference datasets and uti… | 66 | 1290 | active |
| BishopFox/eyeballer Eyeballer is a convolutional neural network tool that classifies screenshots of web hosts taken during large-scope penetration tests. It la… | 55 | 1290 | active |
| ClownsharkBatwing/RES4LYF RES4LYF is a ComfyUI custom node collection providing advanced diffusion samplers (RES samplers) that achieve high-quality image generation… | 67 | 1289 | active |
| xinychen/transdim transdim is a Python/Jupyter Notebook project providing machine learning models for transportation data imputation and spatiotemporal time … | 63 | 1289 | active |
| RoyalVane/CLAN Official PyTorch implementation of CLAN, a CVPR 2019 (oral) / TPAMI 2022 method for unsupervised domain adaptation in semantic segmentation… | 32 | 1289 | stable |
| TencentQQGYLab/ELLA ELLA is an Efficient Large Language Model Adapter that equips text-to-image diffusion models with LLM-based text understanding via a Timest… | 25 | 1289 | active |
| buoyancy99/diffusion-forcing Official research code for the NeurIPS paper 'Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion', implementing a metho… | 65 | 1288 | active |
| bytedance/Bernini Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer perf… | 57 | 1287 | active |
| DBraun/DawDreamer DawDreamer is a Python framework that provides digital audio workstation features, letting users compose graphs of audio processors includi… | 86 | 1286 | active |
| nicodv/kmodes A Python library implementing the k-modes and k-prototypes clustering algorithms for categorical and mixed numerical/categorical data. It f… | 23 | 1286 | stable |
| locuslab/TCN PyTorch implementation of Temporal Convolutional Networks (TCN) with benchmarks from the paper 'An Empirical Evaluation of Generic Convolut… | 32 | 4549 | maintenance |
| inukshuk/anystyle AnyStyle is a fast machine-learning-based parser that splits bibliographic references into structured segments like author, title, and publ… | 42 | 1285 | active |
| InternRobotics/InternUtopia InternUtopia (formerly GRUtopia) is a Python simulation platform built on NVIDIA Isaac Sim for embodied AI research, supporting diverse rob… | 41 | 1285 | active |
| huanngzh/MV-Adapter MV-Adapter is a plug-and-play adapter that turns pre-trained text-to-image diffusion models (e.g., SDXL, SD2.1) into multi-view consistent … | 34 | 1285 | active |
| jbarrow/commonforms CommonForms is a Python package and CLI that uses trained object-detection models (FFDNet-S/L) to automatically detect form fields in a PDF… | 65 | 1284 | active |
| Xiangyue-Zhang/auto-deep-researcher-24x7 An open-source Python framework where LLM agents autonomously run deep learning experiments 24/7, covering hypothesis formation, code imple… | 52 | 1283 | active |
| Phantom-video/HuMo HuMo is a research model and Python codebase from Tsinghua University and ByteDance for human-centric video generation using collaborative … | 46 | 1283 | active |
| poetiq-ai/poetiq-arc-agi-solver A Python research codebase that reproduces Poetiq's record-breaking, leaderboard-topping submissions to the ARC-AGI-1 and ARC-AGI-2 abstrac… | 42 | 1282 | active |
| e3nn/e3nn e3nn is a Python/PyTorch library for building E(3)-equivariant neural networks that respect 3D rotation, translation, and mirror symmetries… | 70 | 1280 | active |
| ZhengyiLuo/PHC Official implementation of the ICCV 2023 paper 'Perpetual Humanoid Control for Real-time Simulated Avatars'. It provides a Python codebase … | 44 | 1280 | active |
| Tianxiaomo/pytorch-YOLOv4 A minimal PyTorch implementation of YOLOv4 (and YOLOv4-tiny) supporting inference and training, with tools to convert Darknet weights to Py… | 32 | 4521 | maintenance |
| zju3dv/MatchAnything MatchAnything is a deep learning model for universal cross-modality image matching, released as research code accompanying a TPAMI 2026 pap… | 64 | 1279 | active |
| AaronJackson/vrn Research code for the ICCV 2017 paper 'Large Pose 3D Face Reconstruction from a Single Image via Direct Volumetric CNN Regression'. It uses… | 32 | 4517 | maintenance |
| Tencent-Hunyuan/SRPO SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion… | 54 | 1278 | active |
| city-super/Scaffold-GS Scaffold-GS is a research implementation of a structured 3D Gaussian splatting method that uses anchor points on a sparse voxel grid to dis… | 27 | 1278 | active |
| opendatalab/LabelLLM LabelLLM is an open-source data annotation platform designed to streamline the labeling workflows needed for LLM development. It offers con… | 65 | 1277 | active |
| suragnair/alpha-zero-general A clean, flexible implementation of the AlphaZero self-play reinforcement learning algorithm that can be adapted to any two-player turn-bas… | 32 | 4506 | maintenance |
| PrunaAI/pruna Pruna is an open-source Python model optimization framework that makes AI models faster, smaller, cheaper, and greener via caching, quantiz… | 84 | 1275 | active |
| studio-dots-ai/dots.tts dots.tts is a 2B-parameter fully continuous, end-to-end autoregressive text-to-speech system, distributed as a Python library with pretrain… | 79 | 1275 | active |
| argmin-rs/argmin argmin is a numerical optimization library written entirely in pure Rust, offering a wide range of optimization algorithms behind a consist… | 60 | 1275 | active |
| Ma-Lab-Berkeley/CRATE CRATE is the official PyTorch implementation of the Coding RAte reduction TransformEr, a family of 'white-box' transformer architectures de… | 29 | 1275 | active |
| google-research/simclr Google Research's official implementation of SimCLR and SimCLRv2, a framework for contrastive learning of visual representations, with 65 p… | 10 | 4502 | maintenance |
| flutter-ml/google_ml_kit_flutter A set of Flutter plugins wrapping Google's ML Kit on-device machine learning APIs for Android and iOS. It provides vision APIs (barcode, fa… | 76 | 1274 | active |
| Stable-X/Stable3DGen Stable3DGen is a modular Python framework for generating 3D assets from images, adapted from Microsoft's TRELLIS with NVIDIA library depend… | 33 | 1274 | active |
| DreamTechAI/Direct3D-S2 Direct3D-S2 is a research framework for high-resolution 3D shape generation from images, built on sparse volumetric representations and a n… | 29 | 1274 | active |
| dcharatan/pixelsplat pixelSplat is a PyTorch implementation of a feed-forward model that reconstructs 3D radiance fields parameterized by 3D Gaussian primitives… | 27 | 1274 | stable |
| nianticlabs/monodepth2 Monodepth2 is the reference PyTorch implementation of the ICCV 2019 paper 'Digging into Self-Supervised Monocular Depth Prediction'. It tra… | 32 | 4497 | maintenance |
| SimonBlanke/Gradient-Free-Optimizers A Python library for gradient-free optimization of black-box functions, offering 23 algorithms (hill climbing, Bayesian optimization, parti… | 94 | 1273 | active |
| Renumics/spotlight Renumics Spotlight is an open-source tool for interactively exploring unstructured datasets (images, audio, text, video, time-series, meshe… | 93 | 1272 | active |
| microsoft/BioGPT BioGPT is Microsoft's domain-specific generative Transformer language model pre-trained on biomedical text, with implementation code and pr… | 32 | 4488 | maintenance |
| NVlabs/stylegan2-ada-pytorch Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi… | 32 | 4487 | maintenance |
| Visual-Agent/DeepEyes DeepEyes is a research project that trains multimodal vision-language models to 'think with images' using end-to-end reinforcement learning… | 44 | 1271 | active |
| leoxiaobin/deep-high-resolution-net.pytorch Official PyTorch implementation of HRNet (Deep High-Resolution Representation Learning for Human Pose Estimation, CVPR 2019), which maintai… | 32 | 4480 | maintenance |
| water8394/flink-recommandSystem-demo A real-time product recommendation system built on Apache Flink, demonstrating streaming computation of product popularity, user profiles, … | 32 | 4480 | maintenance |
| Nixtla/mlforecast mlforecast is a Python framework for time series forecasting using any machine learning model with fit/predict methods, providing efficient… | 96 | 1269 | active |
| MilaNLProc/contextualized-topic-models A Python library implementing Contextualized Topic Models (CTM), which combine pre-trained contextual embeddings like BERT with neural topi… | 47 | 1269 | active |
| luo3300612/Visualizer A lightweight Python library that extracts attention maps and other local variables from deep inside PyTorch models for visualization. It w… | 32 | 1269 | stable |
| lucidrains/flamingo-pytorch A PyTorch implementation of DeepMind's Flamingo visual language model architecture, providing the Perceiver Resampler and Gated Cross-Atten… | 23 | 1269 | active |
| BachiLi/diffvg diffvg is a differentiable rasterizer for 2D vector graphics that bridges the raster and vector domains via backpropagation. It computes pi… | 42 | 1268 | active |
| amaiya/ktrain ktrain is a lightweight Python wrapper around TensorFlow Keras that provides pre-canned, low-code models for text, vision, graph, and tabul… | 25 | 1268 | active |
| vipshop/cache-dit Cache-DiT is a PyTorch-native inference engine that accelerates Diffusion Transformer (DiT) models with hybrid caching, parallelism, quanti… | 82 | 1267 | active |
| facebookresearch/DrQA DrQA is a PyTorch implementation of a system for open-domain question answering that combines document retrieval over Wikipedia with a neur… | 10 | 4468 | maintenance |
| thunlp/OpenNRE OpenNRE is an open-source Python toolkit for neural relation extraction, extracting relation triples between entities from plain text. It u… | 32 | 4467 | maintenance |
| magenta/magenta-studio Magenta Studio is a collection of five MIDI music generation plugins (Continue, Generate, Interpolate, Groove, Drumify) for Ableton Live, b… | 62 | 1266 | active |
| Bjarten/early-stopping-pytorch A small PyTorch utility package providing an EarlyStopping class that monitors validation loss during training and stops when it stops impr… | 34 | 1266 | stable |
| ToniRV/NeRF-SLAM NeRF-SLAM is a real-time dense monocular SLAM system that combines neural radiance fields (Instant-NGP) with probabilistic volumetric fusio… | 32 | 1266 | active |
| nv-tlabs/Difix3D Difix3D+ is a research codebase from NVIDIA implementing a single-step diffusion model pipeline that removes artifacts from NeRF and 3D Gau… | 32 | 1266 | active |
| tensorflow/model-analysis TensorFlow Model Analysis (TFMA) is a Python library for evaluating TensorFlow models on large datasets in a distributed manner. It compute… | 83 | 1265 | active |
| DreamLM/Dream Dream 7B is an open diffusion large language model (dLLM) with base and instruct checkpoints, plus inference and training code built on Hug… | 44 | 1265 | active |