function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| muscriptor/muscriptor MuScriptor is a multi-instrument music transcription model by Kyutai and Mirelo that converts audio recordings into MIDI and sheet music. I… | 79 | 1263 | active |
| meta-pytorch/data TorchData is a PyTorch library providing scalable, performant data loading utilities, including StatefulDataLoader, a drop-in replacement f… | 67 | 1263 | active |
| bryandlee/animegan2-pytorch A PyTorch implementation of AnimeGANv2, a GAN-based image-to-image style transfer model that converts photos into anime-style images. It pr… | 32 | 4452 | maintenance |
| rlawjdghek/StableVITON StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent… | 47 | 1261 | stable |
| MemeMeow-Studio/MemeMeow MemeMeow is a self-hosted meme/sticker management and retrieval application that lets users find images by describing the desired scene in … | 65 | 1260 | active |
| Fictionarry/ER-NeRF ER-NeRF is the official PyTorch implementation of an ICCV 2023 paper on region-aware Neural Radiance Fields for high-fidelity talking portr… | 24 | 1260 | stable |
| ShiftHackZ/Stable-Diffusion-KMP SDAI is an open-source, cross-platform Stable Diffusion client app for Android and iOS built with Kotlin Multiplatform and Jetpack Compose.… | 91 | 1259 | active |
| nv-tlabs/GET3D GET3D is NVIDIA's PyTorch implementation of a generative model that synthesizes high-quality 3D textured meshes (cars, chairs, animals, bui… | 32 | 4435 | maintenance |
| mjpost/sacrebleu SacreBLEU is a Python library and CLI tool for computing shareable, comparable, and reproducible BLEU, chrF, and TER scores for machine tra… | 76 | 1258 | active |
| Francis-Rings/StableAvatar StableAvatar is an end-to-end video diffusion transformer that generates infinite-length, high-quality talking avatar videos from a referen… | 46 | 1258 | active |
| plurai-ai/intellagent IntellAgent is a Python framework for diagnosing and optimizing conversational LLM agents through simulated, realistic synthetic user inter… | 54 | 1257 | active |
| kubeai-project/kubeai KubeAI is a Kubernetes operator for serving machine learning models in production, supporting LLMs via vLLM and Ollama, vector embeddings, … | 89 | 1256 | active |
| Yuliang-Liu/MonkeyOCRv2 MonkeyOCRv2 is a document-native vision encoder and visual-text foundation model for Document AI, pretrained on the 113M-image MonkeyDoc v2… | 58 | 1256 | active |
| aiming-lab/Agent0 Agent0 Series is a research framework for training self-evolving LLM agents from zero external data via tool-integrated reasoning and co-ev… | 58 | 1256 | active |
| BytedTsinghua-SIA/CUDA-Agent CUDA-Agent is a large-scale agentic reinforcement learning system from ByteDance Seed and Tsinghua that trains LLMs to generate high-perfor… | 56 | 1256 | active |
| stardist/stardist StarDist is a Python library for object detection and instance segmentation in 2D and 3D microscopy images using star-convex shapes, built … | 60 | 1255 | stable |
| JonathonLuiten/TrackEval TrackEval is a Python library for evaluating multi-object tracking (MOT) algorithms, implementing metrics such as HOTA, CLEARMOT, IDF1, VAC… | 32 | 1255 | stable |
| ingra14m/Deformable-3D-Gaussians Official PyTorch implementation of the CVPR 2024 paper 'Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene Reconstruction'. … | 18 | 1255 | stable |
| facebookresearch/llm-transparency-tool An interactive toolkit from Meta Research for analyzing the internal workings of Transformer-based language models. It visualizes contribut… | 10 | 1254 | active |
| VainF/pytorch-msssim A PyTorch library providing fast, differentiable SSIM and MS-SSIM image quality metrics using separable Gaussian filtering for speed. It ca… | 23 | 1253 | stable |
| pymc-labs/pymc-marketing PyMC-Marketing is a Python library of Bayesian marketing analytics models built on PyMC, including Marketing Mix Modeling (MMM), Customer L… | 100 | 1252 | active |
| Linketic/CityGaussian Official implementation of the CityGaussian series (ECCV 2024, ICLR 2025) for high-quality large-scale 3D scene reconstruction with Gaussia… | 66 | 1251 | active |
| PyDMD/PyDMD PyDMD is a Python package implementing Dynamic Mode Decomposition (DMD) and its many variants (mrDMD, HODMD, etc.) for extracting spatiotem… | 60 | 1251 | active |
| ModelCloud/GPTQModel GPTQModel is a production-ready Python toolkit for quantizing (compressing) large language models using GPTQ, AWQ, and related methods, wit… | 91 | 1248 | active |
| sentinel-hub/eo-learn eo-learn is a collection of open-source Python packages for accessing and processing spatio-temporal satellite imagery, built around modula… | 51 | 1247 | active |
| gitmylo/audio-webui An all-in-one web UI for audio-related neural networks, bundling text-to-speech (Bark), voice conversion/cloning (RVC), and text-to-audio/m… | 30 | 1247 | active |
| withoutbg/withoutbg-python A Python SDK (pip install withoutbg) for removing image backgrounds, offering a free local open-weights ONNX model and an optional paid clo… | 80 | 1246 | active |
| thu-ml/Motus Motus is the official implementation of a unified latent action world model for robotics, combining a video generation model, a vision-lang… | 43 | 1246 | active |
| lpiccinelli-eth/UniDepth UniDepth is a Python library and research codebase for universal monocular metric depth estimation from single images, based on CVPR 2024 a… | 35 | 1246 | active |
| automl/SMAC3 SMAC3 is a Python library for Bayesian Optimization used to tune hyperparameters of machine learning algorithms and configure arbitrary alg… | 81 | 1244 | active |
| unum-cloud/UForm UForm is a compact multimodal AI library providing tiny image-text embedding models (64-768 dimensions, Matryoshka-style) and small generat… | 55 | 1244 | active |
| fpgaminer/joycaption JoyCaption is an open, free, and uncensored image captioning Visual Language Model (VLM) with released weights and training scripts. It gen… | 53 | 1244 | active |
| LTH14/fractalgen A PyTorch implementation of Fractal Generative Models (FractalGen), enabling pixel-by-pixel high-resolution image generation. It includes p… | 24 | 1244 | active |
| 3DTopia/OpenLRM OpenLRM is an open-source PyTorch implementation of Large Reconstruction Models (LRM) that reconstruct 3D objects (meshes and rendered vide… | 17 | 1244 | active |
| abewley/sort SORT is a barebones Python implementation of a simple online and realtime multiple object tracking algorithm for 2D video sequences, based … | 32 | 4373 | maintenance |
| zinggAI/zingg Zingg is an ML-based tool for scalable master data management, entity resolution, identity resolution, and record deduplication. It runs on… | 88 | 1243 | active |
| HJYao00/Mulberry Mulberry is a research implementation of an o1-like multimodal large language model (MLLM) that performs step-by-step reasoning and reflect… | 49 | 1243 | active |
| cvg/depthsplat DepthSplat is a PyTorch research library implementing a CVPR 2025 model that connects Gaussian splatting with single/multi-view depth estim… | 56 | 1242 | active |
| showlab/Tune-A-Video Tune-A-Video is the official PyTorch implementation of an ICCV 2023 paper that fine-tunes pre-trained text-to-image diffusion models (like … | 31 | 4364 | maintenance |
| alibaba/Tora Tora is Alibaba's official implementation of a trajectory-oriented Diffusion Transformer (DiT) for controllable video generation, integrati… | 64 | 1241 | active |
| Roblox/cube Cube is Roblox's open-source family of foundation models for 3D intelligence, including text-to-3D shape generation and part-controllable m… | 58 | 1241 | active |
| XPixelGroup/HYPIR Official PyTorch implementation of HYPIR, a SIGGRAPH 2025 method that harnesses diffusion-yielded score priors for image restoration. It pr… | 39 | 1241 | active |
| jcjohnson/fast-neural-style A Torch (Lua) implementation of feedforward neural style transfer from the ECCV 2016 paper 'Perceptual Losses for Real-Time Style Transfer … | 32 | 4359 | maintenance |
| google-deepmind/android_env AndroidEnv is a Python library from DeepMind that exposes an Android device (real or emulated) as a Reinforcement Learning environment. Age… | 85 | 1240 | active |
| wannesm/dtaidistance A Python library for computing time series distances, most notably Dynamic Time Warping (DTW), with a fast C implementation exposed via Cyt… | 64 | 1240 | stable |
| marian42/mesh_to_sdf A Python library that computes approximate signed distance fields (SDFs) for arbitrary triangle meshes, including non-watertight, self-inte… | 32 | 1240 | stable |
| ZHKKKe/MODNet MODNet is a deep learning model for real-time portrait matting (background removal) that requires only an RGB image as input, with no trima… | 32 | 4355 | maintenance |
| facebookresearch/deit Official PyTorch repository for DeiT and related vision transformer architectures (CaiT, ResMLP, PatchConvnet, DeiT III), providing trainin… | 10 | 4355 | maintenance |
| declare-lab/tango Tango is a family of latent diffusion models for text-to-audio generation, with Tango 2 improving prompt alignment via DPO-based fine-tunin… | 45 | 1239 | active |
| apchenstu/TensoRF TensoRF is a PyTorch implementation of the ECCV 2022 paper 'TensoRF: Tensorial Radiance Fields', which models and reconstructs radiance fie… | 44 | 1239 | stable |
| sh-lee-prml/HierSpeechpp Official PyTorch implementation of HierSpeech++, a fast zero-shot speech synthesizer for text-to-speech and voice conversion based on hiera… | 28 | 1238 | active |
| pytorch/serve TorchServe is a flexible, production-ready model server for serving, optimizing, and scaling PyTorch models over HTTP with support for CPU,… | 10 | 4346 | maintenance |
| X-Square-Robot/wall-x Wall-X is the open-source training and inference stack for X Square Robot's WALL series of embodied foundation models (VLAs) for general-pu… | 61 | 1236 | active |
| FenTechSolutions/CausalDiscoveryToolbox A Python library for causal inference in graphs and pairwise settings, implementing many algorithms for graph structure recovery from obser… | 53 | 1236 | active |
| open-mmlab/playground OpenMMLab Playground is a central hub collecting and showcasing community projects that extend OpenMMLab libraries with Segment Anything Mo… | 30 | 1236 | active |
| metadriverse/metadrive MetaDrive is an open-source, lightweight driving simulator built for AI and autonomy research, supporting compositional scene synthesis and… | 39 | 1235 | active |
| bytedance/USO USO is ByteDance's open-source unified style- and subject-driven image generation model based on diffusion (FLUX), combining any subject wi… | 36 | 1235 | active |
| meta-pytorch/attention-gym Attention Gym is a collection of tools, examples, and reference implementations for working with PyTorch's FlexAttention API. It provides a… | 91 | 1234 | active |
| xtreme1-io/xtreme1 Xtreme1 is an open-source, self-hosted data labeling and annotation platform for multimodal training data, supporting images, 3D LiDAR poin… | 62 | 1234 | active |
| facebookresearch/home-robot HomeRobot is an open-source robotics stack from Meta AI for mobile manipulation tasks on low-cost hardware like the Hello Robot Stretch. It… | 23 | 1234 | active |
| Equim-chan/mjai-reviewer A Rust CLI tool that reviews riichi mahjong game logs using mjai-compatible AI engines such as Mortal and akochan. It fetches logs from Ten… | 26 | 1233 | active |
| Acly/comfyui-inpaint-nodes A set of custom nodes for ComfyUI that improve image inpainting and outpainting workflows. It integrates the Fooocus inpaint model for SDXL… | 64 | 1232 | active |
| VAST-AI-Research/TripoSplat TripoSplat is an inference-only Python library from TripoAI that converts a single 2D image into high-quality 3D Gaussian splats with a var… | 57 | 1232 | active |
| lucidrains/deep-daze Deep Daze is a simple command line tool for text-to-image generation that combines OpenAI's CLIP with a Siren implicit neural representatio… | 23 | 4315 | maintenance |
| stereolabs/zed-sdk The ZED SDK is a cross-platform spatial perception library for Stereolabs ZED stereo cameras, providing depth sensing, SLAM, 3D reconstruct… | 91 | 1229 | active |
| CSSLab/maia-chess Maia is a collection of human-like neural network chess engines trained on millions of human games, targeting skill levels from ELO 1100 to… | 61 | 1228 | active |
| python-adaptive/adaptive Adaptive is an open-source Python library for parallel active learning of mathematical functions. It intelligently selects the most informa… | 91 | 1227 | active |
| Tencent-Hunyuan/HunyuanCustom HunyuanCustom is a multimodal-driven customized video generation framework built on HunyuanVideo, supporting image, text, audio, and video … | 40 | 1227 | active |
| NVIDIA/BigVGAN BigVGAN is NVIDIA's official PyTorch implementation of a universal neural vocoder (ICLR 2023) that generates high-fidelity raw audio wavefo… | 23 | 1227 | stable |
| alibaba/x-deeplearning X-DeepLearning (XDL) is an industrial deep learning framework from Alibaba optimized for high-dimension sparse data scenarios such as adver… | 23 | 4304 | maintenance |
| microsoft/MInference MInference is a Microsoft library that accelerates long-context LLM inference using dynamic sparse attention, reducing pre-fill latency by … | 49 | 1226 | active |
| pythongosssss/ComfyUI-WD14-Tagger A ComfyUI custom node extension that interrogates images to extract booru-style tags using WD 1.4 tagger models (ONNX-based). It integrates… | 43 | 1226 | active |
| apple/python-apple-fm-sdk Python bindings for Apple's Foundation Models framework, giving access to the on-device foundation model behind Apple Intelligence on macOS… | 75 | 1225 | active |
| Bolin97/GongBU GongBU is a self-hosted, no-code web platform for fine-tuning, evaluating, and deploying large language models, built on Transformers and P… | 52 | 1225 | active |
| run-house/kubetorch Kubetorch is a Python library that lets you distribute and run ML workloads (training, inference, data processing) on Kubernetes directly f… | 83 | 1224 | active |
| 3dg1luk43/ha_washdata A Home Assistant custom integration that monitors smart-plug-connected appliances like washing machines, dryers, and dishwashers. It learns… | 83 | 1224 | active |
| mcmonkeyprojects/sd-dynamic-thresholding A Stable Diffusion extension that enables using higher CFG scale values without color artifacts by clamping latents between sampling steps.… | 36 | 1224 | active |
| MoonshotAI/Kimi-VL Kimi-VL is an open-source Mixture-of-Experts vision-language model (VLM) with a 2.8B activated parameter language decoder, offering multimo… | 33 | 1224 | active |
| yeyupiaoling/Whisper-Finetune A toolkit for fine-tuning OpenAI's Whisper speech recognition models using LoRA, supporting training with or without timestamps and even wi… | 66 | 1223 | active |
| ElectricAlexis/NotaGen NotaGen is a symbolic music generation model that produces high-quality classical sheet music using LLM-style training paradigms: pre-train… | 32 | 1223 | active |
| mrousavy/react-native-fast-tflite A high-performance TensorFlow Lite library for React Native built on Nitro Modules, using the low-level C/C++ TFLite core API with zero-cop… | 84 | 1222 | active |
| eduardoleao052/js-pytorch JS-PyTorch is a deep learning library for JavaScript that closely mirrors PyTorch's syntax, providing tensor operations, automatic differen… | 16 | 1222 | active |
| brian-team/brian2 Brian2 is a free, open-source, clock-driven simulator for spiking neural networks written in Python. It lets scientists define neuron and s… | 75 | 1221 | stable |
| fastgs/FastGS FastGS is a general acceleration framework for 3D Gaussian Splatting that trains scenes in roughly 100 seconds using multi-view consistent … | 49 | 1221 | active |
| SciML/NeuralPDE.jl NeuralPDE.jl is a Julia library of physics-informed neural network (PINN) solvers for ordinary, stochastic, and partial differential equati… | 99 | 1220 | active |
| kellyvv/PhoneClaw PhoneClaw is a mobile-native local AI agent framework that turns phones into on-device agent runtimes, running Gemma models via LiteRT and … | 77 | 1220 | active |
| MotrixLab/SMPLer-X Official code for SMPLer-X, a family of foundation models for expressive human pose and shape estimation (EHPS) that unifies body, hand, an… | 59 | 1220 | stable |
| Project-MONAI/research-contributions A collection of peer-reviewed research prototype implementations built on the MONAI framework for medical imaging AI. It serves as a fast-t… | 43 | 1220 | active |
| gabrielchua/RAGxplorer RAGxplorer is a Python library and hosted Streamlit app for visualizing Retrieval Augmented Generation pipelines. It chunks PDFs, embeds th… | 16 | 1220 | active |
| zkonduit/ezkl EZKL is a Rust-based library and command-line tool that converts deep learning models and arbitrary computational graphs (exported as ONNX)… | 75 | 1219 | active |
| Aratako/Irodori-TTS Irodori-TTS is a Flow Matching-based text-to-speech model with training and inference code, built on a Rectified Flow Diffusion Transformer… | 58 | 1219 | active |
| microsoft/malmo Project Malmo is a platform for artificial intelligence experimentation and research built on top of Minecraft, providing a gym-like enviro… | 10 | 4270 | maintenance |
| nfstream/nfstream NFStream is a multiplatform Python framework for fast, flexible network flow data analysis from live interfaces or pcap files. It provides … | 81 | 1218 | stable |
| bowang-lab/MedRAX MedRAX is a medical reasoning agent framework that integrates chest X-ray analysis tools (segmentation, grounding, report generation, disea… | 42 | 1218 | active |
| simondlevy/TinyEKF TinyEKF is a lightweight, header-only C/C++ implementation of the Extended Kalman Filter designed for microcontrollers like Arduino and STM… | 66 | 1217 | stable |
| lucidrains/perceiver-pytorch A PyTorch implementation of the Perceiver architecture (General Perception with Iterative Attention) and its follow-up Perceiver IO. It pro… | 62 | 1217 | active |
| fudan-generative-vision/champ Champ is a research framework for controllable and consistent human image animation using 3D parametric guidance (SMPL-based depth, normal,… | 25 | 4261 | maintenance |
| Picsart-AI-Research/Text2Video-Zero Official implementation of Text2Video-Zero, a zero-shot text-to-video generation method that adapts text-to-image diffusion models like Sta… | 30 | 4245 | maintenance |
| ifzhang/FairMOT FairMOT is a research implementation of a one-shot multi-object tracking model that jointly performs object detection and re-identification… | 32 | 4244 | maintenance |
| hasktorch/hasktorch Hasktorch is a Haskell library for tensor math and neural networks, built on bindings to the C++ libtorch libraries that power PyTorch. It … | 76 | 1211 | active |