function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| OpenTSLM/OpenTSLM OpenTSLM is a family of Time-Series Language Models that integrate time series as a native modality into pretrained LLMs (Llama, Gemma), en… | 61 | 1211 | active |
| GuitarML/NeuralPi NeuralPi is a DIY Raspberry Pi 4-based guitar pedal that uses neural networks to emulate real amplifiers and distortion/overdrive pedals in… | 23 | 1211 | active |
| MyoHub/myosuite MyoSuite is a collection of musculoskeletal environments and tasks simulated with the MuJoCo physics engine and wrapped in the OpenAI gym A… | 93 | 1210 | active |
| Zefan-Cai/R-KV R-KV is a training-free, redundancy-aware KV cache compression method for reasoning LLMs, discarding repetitive tokens on-the-fly during de… | 60 | 1209 | active |
| datawhalechina/torch-rechub Torch-RecHub is a lightweight PyTorch framework for building recommendation system models with 30+ out-of-the-box algorithms covering ranki… | 94 | 1207 | active |
| rohanpsingh/LearningHumanoidWalking A Python research codebase for training humanoid robots to walk using deep reinforcement learning (PPO) in MuJoCo simulation. It provides e… | 70 | 1207 | active |
| ardha27/AI-Song-Cover-RVC A collection of Google Colab and Kaggle notebooks that form an all-in-one toolkit for creating AI song covers with RVC (Retrieval-based Voi… | 68 | 1207 | active |
| DachunKai/EvTexture Official PyTorch implementation of EvTexture and EvTexture++, event-driven video super-resolution models that use event-camera signals to e… | 54 | 1207 | active |
| willisma/SiT Official PyTorch implementation of Scalable Interpolant Transformers (SiT), a family of generative models built on Diffusion Transformers t… | 53 | 1206 | active |
| aTrainTranscription/aTrain aTrain is a desktop GUI application for offline transcription of speech recordings using Whisper-based machine learning models, with speake… | 79 | 1205 | active |
| frotms/PaddleOCR2Pytorch A PyTorch port of PaddleOCR that lets you run PaddleOCR-trained models (detection, recognition, and document structure parsing) without the… | 73 | 1205 | active |
| dexsuite/dex-retargeting A Python library of retargeting optimizers that translate human hand motion (from video or pose datasets) into robot dexterous hand joint c… | 37 | 1205 | active |
| ICT-FinD-Lab/alphagen AlphaGen is a Python research library that automatically generates formulaic alpha (predictive) stock factors using reinforcement learning,… | 70 | 1204 | active |
| graphframes/graphframes GraphFrames is a package for Apache Spark that provides DataFrame-based graph processing with distributed graph algorithms like PageRank, c… | 95 | 1203 | stable |
| metavoiceio/metavoice-src MetaVoice-1B is a 1.2B parameter foundational text-to-speech model trained on 100K hours of speech, focused on emotional rhythm and tone in… | 26 | 4205 | maintenance |
| dendenxu/fast-gaussian-rasterization A drop-in replacement for diff-gaussian-rasterization that renders 3D Gaussian Splatting scenes using a geometry-shader-based GPU pipeline … | 19 | 1202 | active |
| NVIDIA/kvpress kvpress is a Python library from NVIDIA that implements multiple KV cache compression methods and benchmarks for long-context LLM inference… | 86 | 1201 | active |
| Chevron7Locked/kima-hub Kima Hub is a self-hosted, on-demand audio streaming platform that brings a Spotify-like experience to your personal music library, includi… | 83 | 1200 | active |
| basetenlabs/truss Truss is a Python CLI and packaging framework for deploying and serving AI/ML models in production, primarily on the Baseten platform. It h… | 96 | 1199 | active |
| xorbitsai/xorbits Xorbits is an open-source distributed computing framework that scales Python data science and machine learning workloads from a laptop to l… | 64 | 1199 | active |
| Artelnics/opennn OpenNN is an open-source C++ library for building, training, and deploying neural networks for advanced analytics. It is dependency-free, o… | 97 | 1198 | active |
| cp2k/cp2k CP2K is an open-source quantum chemistry and solid state physics package for atomistic simulations of molecular, liquid, periodic, and biol… | 89 | 1198 | stable |
| JuliaStats/Distributions.jl A Julia package providing a comprehensive collection of probability distributions and associated functions. It implements moments, entropy,… | 98 | 1197 | stable |
| openvinotoolkit/nncf NNCF is Intel's Neural Network Compression Framework, a Python library providing post-training and training-time compression algorithms (qu… | 94 | 1197 | active |
| Calamari-OCR/calamari Calamari is a Python-based OCR engine for line-based automatic text recognition, built on OCRopy and Kraken with a TensorFlow deep-learning… | 74 | 1197 | active |
| toshas/torch-fidelity A PyTorch library providing accurate and efficient implementations of generative model evaluation metrics such as FID, Inception Score, KID… | 70 | 1197 | active |
| autonomousvision/stylegan-t Official training code for StyleGAN-T, an ICML 2023 paper on fast large-scale text-to-image synthesis using GANs. It provides dataset prepa… | 31 | 1197 | active |
| DeepRec-AI/DeepRec DeepRec is a high-performance deep learning framework for recommendation models, built on TensorFlow 1.15 with Intel and NVIDIA TensorFlow … | 24 | 1197 | active |
| martinpacesa/BindCraft BindCraft is a Python-based computational pipeline for de novo protein binder design that combines AlphaFold2 backpropagation, ProteinMPNN,… | 78 | 1196 | active |
| NVlabs/alpasim AlpaSim is an open-source, Python-based autonomous vehicle simulation platform for developing and testing end-to-end AV policies in closed … | 75 | 1196 | active |
| aydinnyunus/ai-captcha-bypass A Python command-line tool that uses multimodal LLMs (GPT-4o, Gemini) to automatically solve various CAPTCHA types, including text, reCAPTC… | 61 | 1196 | active |
| haidog-yaqub/MeanFlow An unofficial PyTorch implementation of MeanFlow and iMF, one-step generative modeling methods based on flow matching. It provides config-d… | 59 | 1196 | active |
| deepseek-ai/DeepSeek-VL DeepSeek-VL is an open-source vision-language foundation model for real-world multimodal understanding, released with model weights and inf… | 25 | 4175 | maintenance |
| qualcomm/ai-hub-models Qualcomm AI Hub Models is a curated collection of 300+ state-of-the-art machine learning models (vision, audio, speech, generative AI) pre-… | 89 | 1195 | active |
| cvs-health/uqlm UQLM is a Python library for detecting LLM hallucinations using uncertainty quantification techniques. It provides five categories of score… | 86 | 1195 | active |
| EvolvingLMMs-Lab/LLaVA-OneVision-2 A fully open framework for training multimodal large language models, releasing models, datasets, and training recipes for the LLaVA-OneVis… | 72 | 1195 | active |
| facebookresearch/esm Meta FAIR's Evolutionary Scale Modeling (ESM) library providing Transformer protein language models with pretrained weights, including ESM-… | 10 | 4170 | maintenance |
| compdemocracy/polis Polis is an open-source, AI-powered sentiment gathering platform for large-scale open-ended feedback, mapping high-dimensional opinion spac… | 75 | 1194 | active |
| zai-org/CogAgent CogAgent is an open-source vision-language model (VLM) based GUI agent that understands screen captures and natural language to automate in… | 33 | 1194 | active |
| omicverse/omicverse OmicVerse is a Python library for multi-omics analysis covering bulk, single-cell, and spatial RNA-seq workflows. It is part of the scverse… | 94 | 1193 | active |
| Cerebras/modelzoo Cerebras Model Zoo is a collection of reference deep learning model implementations (Llama, Mixtral, DINOv2, Llava, etc.) with configs and … | 77 | 1193 | active |
| deepfence/FlowMeter FlowMeter is a Go utility that analyzes network packet headers, groups packets into flows, and uses machine learning to classify flows as b… | 10 | 1193 | active |
| FEniCS/dolfinx DOLFINx is the next-generation computational environment of the FEniCS Project, implemented in C++ with Python bindings, for solving partia… | 95 | 1191 | active |
| Anserini Pyserini is a Python toolkit for reproducible information retrieval research supporting both sparse (via Anserini/Lucene) and dense (via Fa… | 77 | 1191 | active |
| Tencent-Hunyuan/HunyuanWorld-Mirror HunyuanWorld-Mirror is a feed-forward 3D reconstruction model from Tencent that predicts camera poses, intrinsics, depth maps, point clouds… | 54 | 1191 | active |
| trianglesplatting/triangle-splatting Official implementation of 'Triangle Splatting for Real-Time Radiance Field Rendering' (3DV 2026), which uses 3D triangles as rendering pri… | 43 | 1191 | active |
| shallowdream204/DreamClear DreamClear is a diffusion-transformer based real-world image restoration model for high-fidelity super-resolution, published at NeurIPS 202… | 27 | 1191 | active |
| zai-org/VisualGLM-6B VisualGLM-6B is an open-source multimodal conversational language model supporting images, Chinese, and English, built on ChatGLM-6B with a… | 30 | 4154 | maintenance |
| a-r-j/graphein Graphein is a Python library for constructing graph and mesh representations of proteins, RNA, molecules, and biological interaction networ… | 76 | 1190 | active |
| sair-lab/AirSLAM AirSLAM is an efficient, illumination-robust point-line visual SLAM system supporting stereo visual odometry/VIO, offline map optimization,… | 47 | 1190 | active |
| getkeops/keops KeOps (pykeops) is a Python library for computing kernel reductions over large arrays on CPUs and GPUs using efficient C++/CUDA routines wi… | 65 | 1189 | active |
| pqpo/SmartCropper An Android library for smart image cropping that automatically detects document borders using OpenCV (with an optional TensorFlow Lite HED … | 66 | 4132 | maintenance |
| pymc-labs/CausalPy CausalPy is a Python package for causal inference in quasi-experimental settings, offering methods like difference-in-differences, syntheti… | 94 | 1185 | active |
| SHI-Labs/Neighborhood-Attention-Transformer Official PyTorch implementation of the Neighborhood Attention Transformer (NAT/DiNAT), a family of efficient vision transformers with local… | 32 | 1184 | stable |
| chakki-works/seqeval seqeval is a Python library for evaluating sequence labeling tasks such as named-entity recognition, part-of-speech tagging, and semantic r… | 23 | 1184 | stable |
| mlmed/torchxrayvision TorchXRayVision is an open-source PyTorch library providing pre-trained deep learning models and a unified interface for publicly available… | 91 | 1183 | active |
| googleapis/go-genai The official Google Gen AI SDK for Go, providing a client interface to Google's generative models such as Gemini via the Gemini Developer A… | 85 | 1183 | active |
| functime-org/functime functime is a Python library for production-ready global forecasting and time-series feature extraction on large panel datasets, built on l… | 70 | 1183 | active |
| GAIR-NLP/ASI-Arch A multi-agent framework that lets an LLM autonomously conduct end-to-end research on neural network architecture discovery, specifically li… | 43 | 1183 | active |
| chensjtu/GaussianObject GaussianObject is a research framework for high-quality 3D object reconstruction from as few as four input images using Gaussian splatting,… | 26 | 1183 | active |
| orpatashnik/StyleCLIP Official implementation of StyleCLIP, a method for text-driven manipulation of StyleGAN-generated imagery using CLIP. It provides three app… | 32 | 4121 | maintenance |
| NVIDIA/audio-flamingo NVIDIA's PyTorch implementation of the Audio Flamingo series of large audio-language models (AF1, AF2, AF3, and Music Flamingo) for audio u… | 50 | 1182 | active |
| XPixelGroup/DiffBIR DiffBIR is a blind image restoration framework that uses generative diffusion priors to restore degraded real-world images. It provides pre… | 34 | 4119 | maintenance |
| mlfoundations/open_flamingo OpenFlamingo is an open-source PyTorch implementation of DeepMind's Flamingo, a large multimodal vision-language model that interleaves ima… | 23 | 4118 | maintenance |
| DAMO-NLP-SG/VideoLLaMA3 VideoLLaMA 3 is a frontier multimodal foundation model for image and video understanding, released with checkpoints, inference code, and de… | 37 | 1179 | active |
| ddlBoJack/emotion2vec Official PyTorch implementation of emotion2vec, a self-supervised pre-trained model for speech emotion representation. It provides code for… | 27 | 1179 | active |
| higgsfield-ai/higgsfield Higgsfield is an open-source GPU orchestration and machine learning framework for fault-tolerant, distributed training of very large models… | 23 | 4106 | maintenance |
| ubicomplab/rPPG-Toolbox rPPG-Toolbox is an open-source Python toolbox for camera-based physiological sensing (remote photoplethysmography), enabling heart rate and… | 51 | 1178 | active |
| NVlabs/Deep_Object_Pose NVIDIA's Deep Object Pose Estimation (DOPE), a deep learning system for detecting known objects and estimating their 6-DoF pose from RGB ca… | 48 | 1178 | active |
| balancap/SSD-Tensorflow A TensorFlow re-implementation of the Single Shot MultiBox Detector (SSD) for object detection, including VGG-based SSD-300 and SSD-512 net… | 32 | 4101 | maintenance |
| AlgRUC/JittorGeometric JittorGeometric is a graph machine learning library built on the Jittor deep learning framework, providing implementations of 40+ Graph Neu… | 62 | 1177 | active |
| Tencent-Hunyuan/MixGRPO MixGRPO is a research framework from Tencent Hunyuan implementing a mixed ODE-SDE GRPO algorithm for efficient reinforcement learning fine-… | 58 | 1177 | active |
| Soul-AILab/SoulX-LiveAct SoulX-LiveAct is the official inference code for a real-time human animation framework that generates lifelike, audio/multimodal-controlled… | 54 | 1176 | active |
| tjiiv-cprg/EPro-PnP EPro-PnP is a probabilistic Perspective-n-Points (PnP) layer for end-to-end 6DoF monocular object pose estimation networks, built on PyTorc… | 41 | 1175 | stable |
| warmshao/FasterLivePortrait A real-time portrait animation application based on LivePortrait that animates still photos or videos using a driving video, image, audio, … | 38 | 1174 | active |
| baichuan-inc/Baichuan2 Baichuan 2 is a family of open large language models (7B and 13B, Base and Chat variants with 4-bit quantized versions) trained by Baichuan… | 28 | 4084 | maintenance |
| MIC-DKFZ/batchgenerators A Python framework for data augmentation of 2D and 3D images, developed by the German Cancer Research Center for medical image classificati… | 71 | 1173 | stable |
| chongzhou96/EdgeSAM EdgeSAM is the official PyTorch implementation of a distilled, accelerated variant of the Segment Anything Model (SAM) designed for on-devi… | 37 | 1173 | active |
| princeton-nlp/MeZO MeZO is a memory-efficient zeroth-order optimizer that fine-tunes language models using only forward passes, with the same memory footprint… | 29 | 1173 | stable |
| NVlabs/imaginaire NVIDIA's PyTorch library containing optimized implementations of image and video synthesis methods, including GAN-based image-to-image tran… | 32 | 4082 | maintenance |
| bytedance/1d-tokenizer A research repository from ByteDance containing code and pretrained model weights for 1D visual tokenizers (TiTok, TA-TiTok, FlowTok) and i… | 29 | 1172 | active |
| csguoh/MambaIR MambaIR and MambaIRv2 are PyTorch-based image restoration models built on Mamba state-space models, published at ECCV 2024 and CVPR 2025. T… | 54 | 1171 | active |
| magicleap/SuperGluePretrainedNetwork SuperGlue is a PyTorch implementation of a graph neural network with an optimal matching layer that matches sparse image features between t… | 32 | 4072 | maintenance |
| onnx/onnxmltools ONNXMLTools is a Python library that converts machine learning models from various toolkits (Keras/TensorFlow, scikit-learn, Core ML, XGBoo… | 78 | 1170 | active |
| 8080labs/ppscore ppscore is a Python library implementing the Predictive Power Score (PPS), an asymmetric, data-type-agnostic metric that detects linear and… | 55 | 1170 | stable |
| baidu-research/warp-ctc A fast parallel implementation of the Connectionist Temporal Classification (CTC) loss function for CPU and CUDA GPU, with a simple C inter… | 32 | 4069 | maintenance |
| jenly1314/MLKit MLKit is an easy-to-use Kotlin wrapper library around Google ML Kit for Android, exposing text recognition, barcode scanning, image labelin… | 79 | 1168 | active |
| CarlGao4/Demucs-Gui A graphical desktop application wrapping the Demucs AI music source-separation model, letting users split songs into stems (vocals, drums, … | 46 | 1168 | active |
| iver56/torch-audiomentations A PyTorch library for fast audio data augmentation, inspired by audiomentations. It provides GPU-accelerated, differentiable audio transfor… | 58 | 1167 | active |
| SuperBruceJia/EEG-DL EEG-DL is a deep learning library built on TensorFlow for classifying EEG signals, supporting many architectures including CNNs, RNNs, GCNs… | 47 | 1167 | active |
| Object Detection Metrics A Python toolkit implementing the most popular metrics (AP, mAP, precision-recall curves) used to evaluate object detection algorithms, wit… | 50 | 1166 | stable |
| GaParmar/clean-fid Clean-FID is a PyTorch library for computing the Frechet Inception Distance (FID) with correct image resizing and quantization steps, fixin… | 48 | 1166 | stable |
| cure-lab/MagicDrive MagicDrive is the official PyTorch implementation of an ICLR 2024 paper for controllable street view generation using diffusion models with… | 35 | 1166 | active |
| nv-tlabs/LLaMA-Mesh LLaMA-Mesh is a fine-tuned large language model from NVIDIA Research that generates and understands 3D meshes by representing vertex coordi… | 28 | 1166 | active |
| sksq96/pytorch-summary A PyTorch library providing a Keras-style model.summary() that prints layer types, output shapes, parameter counts, and memory estimates. I… | 32 | 4053 | maintenance |
| facebookresearch/VideoPose3D A PyTorch implementation of CVPR 2019 research on 3D human pose estimation in video using temporal convolutions over 2D keypoint trajectori… | 10 | 4052 | maintenance |
| yosinski/deep-visualization-toolbox A GUI toolbox for visualizing and understanding deep neural networks, showing per-unit activations, backprop/deconv, and regularized-optimi… | 32 | 4051 | maintenance |
| lmstudio-ai/mlx-engine mlx-engine is the Apple MLX-based LLM inference engine that powers LM Studio on Mac, built on mlx-lm with support for vision models via mlx… | 67 | 1165 | active |
| Text-to-Audio/AudioLCM AudioLCM is a PyTorch implementation of an ACM-MM'24 paper for efficient, high-quality text-to-audio generation using latent consistency mo… | 37 | 1165 | active |
| thunlp/OpenKE OpenKE is an open-source PyTorch-based toolkit for knowledge graph embedding (knowledge representation learning), with C++ accelerated data… | 32 | 4047 | maintenance |