function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| majianjia/nnom NNoM is a high-level neural network inference library written in C for microcontrollers. It converts Keras models into optimized on-device … | 23 | 1164 | stable |
| facebookresearch/encodec EnCodec is a deep learning based neural audio codec from Meta AI that compresses mono 24 kHz and stereo 48 kHz audio to bitrates from 1.5 t… | 32 | 4041 | maintenance |
| ucla-mobility/OpenCDA OpenCDA is an open-source Python framework for prototyping and evaluating full-stack cooperative driving automation (CDA) applications in a… | 67 | 1162 | active |
| Linaom1214/TensorRT-For-YOLO-Series A Python and C++ toolkit for running YOLO-series object detection models (YOLOv3 through YOLOv12, YOLOX) with NVIDIA TensorRT, including ON… | 44 | 1162 | active |
| chengxuxin/extreme-parkour Official code for 'Extreme Parkour with Legged Robots' (ICRA 2024), a reinforcement learning framework for training quadruped robots to per… | 27 | 1161 | active |
| minivision-ai/photo2cartoon A Python deep-learning project from Minivision that converts real portrait photos into cartoon-style avatars using unpaired image translati… | 32 | 4029 | maintenance |
| DepthAnything/PromptDA Prompt Depth Anything is a Python library implementing a CVPR 2025 method for high-resolution (up to 4K) accurate metric depth estimation. … | 50 | 1159 | active |
| kijai/ComfyUI-IC-Light A ComfyUI custom node that provides a native implementation of IC-Light models for image relighting. It lets users run IC-Light workflows i… | 35 | 1159 | active |
| LibCity/Bigscity-LibCity LibCity is an open-source PyTorch library for urban spatial-temporal data mining, providing a unified pipeline for traffic prediction resea… | 23 | 1157 | active |
| fundamentalvision/Deformable-DETR Official PyTorch implementation of Deformable DETR, an efficient end-to-end object detector that uses deformable attention to fix DETR's sl… | 32 | 4015 | maintenance |
| JackHopkins/factorio-learning-environment An open-source framework for developing and evaluating LLM agents in the game of Factorio, providing an open-ended, non-saturating benchmar… | 82 | 1156 | active |
| sirius-ai/LPRNet_Pytorch A PyTorch implementation of LPRNet, a lightweight deep neural network for license plate recognition. It ships with pretrained weights focus… | 32 | 1156 | stable |
| gemelo-ai/vocos Vocos is a fast neural vocoder that synthesizes audio waveforms from acoustic features such as mel-spectrograms or EnCodec tokens. It uses … | 65 | 1155 | stable |
| wuyoscar/Internal-Safety-Collapse ISC-Bench/TVD is a research framework for studying 'Internal Safety Collapse' in frontier LLMs, where agents placed in adversarial codespac… | 59 | 1155 | active |
| deepmodeling/Uni-Mol Uni-Mol is a collection of 3D molecular representation learning frameworks and pretrained models for tasks like molecule property predictio… | 34 | 1155 | active |
| JunMa11/SegLossOdyssey A curated collection of loss functions for medical image segmentation, accompanying the 'Loss Odyssey in Medical Image Segmentation' survey… | 32 | 4007 | maintenance |
| SystemErrorWang/White-box-Cartoonization Official TensorFlow implementation of the CVPR 2020 paper 'Learning to Cartoonize Using White-box Cartoon Representations', which converts … | 61 | 4001 | maintenance |
| TJU-Aerial-Robotics/YOPO YOPO is a learning-based one-stage planner for quadrotor autonomous navigation in obstacle-dense environments, integrating perception, mapp… | 79 | 1153 | active |
| apple/ml-clara CLaRa is Apple's open-source end-to-end Retrieval-Augmented Generation model that compresses documents into continuous latent representatio… | 43 | 1153 | active |
| OpenGVLab/VisionLLM VisionLLM is a series of open-source multimodal large language models from OpenGVLab that unify vision-centric tasks under language instruc… | 33 | 1153 | active |
| Libr-AI/OpenFactVerification Loki is an open-source Python tool that automates fact verification by decomposing texts into claims, retrieving evidence via search, and u… | 15 | 1153 | active |
| quark0/darts DARTS is the official PyTorch implementation of the ICLR 2019 paper 'DARTS: Differentiable Architecture Search', which performs neural arch… | 32 | 3997 | maintenance |
| shyamsn97/mario-gpt MarioGPT is a Python library and finetuned GPT2 model that generates playable Super Mario Bros levels from text prompts. It accompanies the… | 21 | 1152 | active |
| mangdangroboticsclub/QuadrupedRobot Mini Pupper is an open-source ROS-based quadruped robot dog kit built around Raspberry Pi, with software for SLAM, navigation, and OpenCV-b… | 67 | 1150 | active |
| amazon-science/mm-cot Official PyTorch implementation of the paper 'Multimodal Chain-of-Thought Reasoning in Language Models', which adds vision features to a tw… | 31 | 3985 | maintenance |
| lhotse-speech/lhotse Lhotse is a Python library for flexible, scalable preparation of multimodal (speech, audio, video, image, text) data for machine learning, … | 89 | 1149 | active |
| PKU-Alignment/omnisafe OmniSafe is a PyTorch-based infrastructural framework for safe reinforcement learning research, providing a unified modular toolkit and com… | 27 | 1149 | active |
| brightmart/albert_zh A repository providing pre-trained ALBERT models for Chinese language, implemented in TensorFlow with PyTorch and Keras conversions. It inc… | 32 | 3982 | maintenance |
| JDAI-CV/fast-reid FastReID is a PyTorch-based research platform implementing state-of-the-art re-identification algorithms for persons, vehicles, and faces. … | 23 | 3981 | maintenance |
| bodaay/HuggingFaceModelDownloader A Go-based CLI utility for downloading models and datasets from the HuggingFace Hub with parallel, resumable downloads. It includes an inte… | 87 | 1147 | active |
| open-gigaai/giga-world-1 GigaWorld-1 is an open-source framework providing training, inference, data processing, checkpoint conversion, and LoRA merge workflows for… | 54 | 1147 | active |
| simpler-env/SimplerEnv SIMPLER (SimplerEnv) is a collection of simulated environments built on SAPIEN/ManiSkill for evaluating real-world robot manipulation polic… | 51 | 1147 | active |
| snipsco/snips-nlu Snips NLU is a Python library (with a Rust core) that extracts structured meaning from natural language text by detecting user intents and … | 23 | 3973 | maintenance |
| google/lyra Lyra is a very low-bitrate speech codec from Google that combines traditional codec techniques with generative machine learning models to c… | 23 | 3973 | maintenance |
| ShiqiYu/OpenGait OpenGait is a flexible and extensible Python framework for gait recognition research, providing implementations of state-of-the-art models … | 67 | 1146 | active |
| midas-research/audino Audino is an open-source, self-hosted web application for annotating audio, supporting transcription, labeling, and speaker-related tasks. … | 52 | 1146 | active |
| okooo5km/HiPixel HiPixel is a native macOS app for AI-powered image super-resolution, built with SwiftUI and using Upscayl's AI models. It offers batch upsc… | 76 | 1145 | active |
| lquesada/ComfyUI-Inpaint-CropAndStitch A set of ComfyUI custom nodes that crop an image around a masked area before sampling and stitch the inpainted result back afterward. This … | 68 | 1145 | active |
| sgl-project/SpecForge SpecForge is a Python framework from the SGLang team for training speculative decoding models such as EAGLE/EAGLE3 draft heads. Trained mod… | 64 | 1145 | active |
| Codium-ai/AlphaCodium Official implementation of the AlphaCodium paper, a test-based, multi-stage, iterative flow for LLM code generation on competitive programm… | 26 | 3965 | maintenance |
| extropic-ai/thrml THRML is a JAX library for building and sampling probabilistic graphical models, focused on efficient block Gibbs sampling of energy-based … | 71 | 1144 | active |
| ScorpioLea/AiCE AiCE is a Python tool that predicts high-fitness protein mutations by sampling sequences from protein inverse folding models such as Protei… | 40 | 1144 | active |
| facebookresearch/fairseq2 fairseq2 is a PyTorch-based sequence modeling toolkit from Meta FAIR for training custom models for content generation tasks such as langua… | 89 | 1143 | active |
| callous-youth/BOAT BOAT is a PyTorch-based library providing a compositional, operation-level toolbox for gradient-based bi-level optimization (BLO). It decom… | 87 | 1143 | active |
| IDEA-Research/Grounding-DINO-1.5-API Python examples and API client for Grounding DINO 1.5/1.6, IDEA Research's open-world (open-set) object detection model series hosted on De… | 25 | 1143 | active |
| chrishayuk/larql LARQL decompiles transformer models into a queryable 'vindex' (vector index) format and provides LQL, a SQL-like query language, to browse,… | 82 | 1142 | active |
| nnaisense/evotorch EvoTorch is an open-source evolutionary computation library built on top of PyTorch, developed at NNAISENSE. It provides distribution-based… | 78 | 1142 | active |
| MIT-SPARK/Hydra Hydra is a C++ system that incrementally builds hierarchical 3D Scene Graphs from sensor data in real time. It is developed by MIT SPARK as… | 68 | 1142 | active |
| facebookresearch/StarSpace StarSpace is a general-purpose neural model from Facebook Research that learns entity embeddings for classification, retrieval, ranking, an… | 10 | 3952 | maintenance |
| HengyiWang/spann3r Spann3R is a transformer-based model for dense 3D reconstruction from ordered or unordered image collections, built on the DUSt3R paradigm.… | 26 | 1141 | active |
| cvg/glue-factory Glue Factory is a PyTorch-based library for training and evaluating deep neural networks that detect and match local visual features (point… | 69 | 1140 | active |
| Tavish9/any4lerobot Any4LeRobot is a curated collection of Python utilities for the Hugging Face LeRobot robotics ecosystem, including dataset conversion scrip… | 64 | 1140 | active |
| smthemex/ComfyUI_Sonic A ComfyUI custom node implementing the Sonic method for audio-driven portrait animation, generating talking-head videos from a single portr… | 56 | 1140 | active |
| THUMNLab/AutoGL AutoGL is an autoML framework and toolkit for machine learning on graphs, built on PyTorch with PyTorch Geometric and DGL backends. It prov… | 47 | 1140 | active |
| rohitgandikota/sliders Official implementation of Concept Sliders, LoRA adaptors that enable precise, plug-and-play control of attributes in diffusion models like… | 52 | 1139 | active |
| clovaai/deep-text-recognition-benchmark Official PyTorch implementation of a four-stage scene text recognition (OCR) framework from an ICCV 2019 paper, with training and evaluatio… | 32 | 3942 | maintenance |
| EyeTrackVR/EyeTrackVR EyeTrackVR is a free, open-source, DIY software platform that turns affordable cameras and IR LEDs mounted inside a VR headset into an eye … | 87 | 1138 | active |
| chengtan9907/OpenSTL OpenSTL is a comprehensive benchmark and modular framework for spatio-temporal predictive learning, covering video prediction methods acros… | 54 | 1137 | active |
| facebookresearch/watermark-anything Official PyTorch implementation and pretrained models for the paper 'Watermark Anything with Localized Messages', which embeds multiple loc… | 10 | 1137 | active |
| diego-vicente/som-tsp A Python implementation that solves the Traveling Salesman Problem using Self-Organizing Maps, a neural network technique adapted to find a… | 32 | 3933 | maintenance |
| xzf-thu/Mega-ASR Mega-ASR is a foundation automatic speech recognition model trained on 2.6M samples spanning 7 atomic acoustic conditions and 54 compound r… | 59 | 1136 | active |
| vec2text/vec2text A Python library for text embedding inversion: training and running models that reconstruct text sequences from their sentence embeddings. … | 57 | 1136 | active |
| horseee/LLM-Pruner LLM-Pruner is a PyTorch library implementing structural pruning of large language models based on gradient information, as published at Neu… | 29 | 1136 | active |
| IST-DASLab/marlin Marlin is a highly optimized FP16xINT4 matrix multiplication CUDA kernel for LLM inference that achieves near-ideal 4x speedups at batch si… | 26 | 1136 | active |
| Dao-AILab/quack QuACK is a collection of high-performance GPU kernels (RMSNorm, LayerNorm, softmax, cross-entropy, GEMM with epilogues) written in NVIDIA's… | 85 | 1135 | active |
| CellProfiler/CellProfiler CellProfiler is a free, open-source desktop application for quantitative analysis of biological images, letting biologists build modular im… | 68 | 1135 | active |
| Kiteretsu77/APISR APISR is a deep-learning based super-resolution tool that restores and enhances low-quality, low-resolution anime images and videos using t… | 37 | 1135 | active |
| rhymes-ai/Allegro Allegro is an open-source text-to-video generation model that produces high-quality 720p videos up to 6 seconds at 15 FPS from text prompts… | 24 | 1135 | active |
| SarahWeiii/CoACD CoACD is a C++ library (with Python bindings and a Unity package) that performs approximate convex decomposition of 3D triangle meshes usin… | 98 | 1134 | active |
| OpenGVLab/SAM-Med2D Official implementation of SAM-Med2D, a fine-tuned Segment Anything Model (SAM) for 2D medical image segmentation, trained on the SA-Med2D-… | 28 | 1134 | active |
| Janspiry/Image-Super-Resolution-via-Iterative-Refinement An unofficial PyTorch implementation of SR3 (Image Super-Resolution via Iterative Refinement), a diffusion-based model for image super-reso… | 32 | 3923 | maintenance |
| aliyun/SimAI SimAI is a large-scale network simulation toolkit from Alibaba Cloud for modeling AI training and inference workloads on GPU clusters, publ… | 66 | 1133 | active |
| huawei-noah/trustworthyAI A collection of trustworthy AI projects from Huawei Noah's Ark Lab, centered on gCastle, a causal structure learning toolchain with many gr… | 61 | 1132 | active |
| zai-org/SCAIL-2 Official implementation of SCAIL-2, an open-source model for end-to-end controlled character animation that drives character videos from re… | 58 | 1132 | active |
| FlagOpen/RoboBrain2.5 RoboBrain 2.5 is an open-source embodied AI foundation model from BAAI that combines multimodal large language model capabilities with 3D s… | 50 | 1132 | active |
| mgonzs13/yolo_ros A ROS 2 wrapper for Ultralytics YOLO models (YOLOv8 through YOLO26) providing object detection, tracking, instance segmentation, human pose… | 93 | 1131 | active |
| sooftware/conformer An unofficial PyTorch implementation of the Conformer architecture (convolution-augmented Transformer) from the INTERSPEECH 2020 paper, tar… | 72 | 1131 | active |
| noahcao/OC_SORT OC-SORT is a pure motion-model-based multi-object tracker for video, improving on SORT by fixing Kalman filter limitations to handle occlus… | 67 | 1131 | stable |
| Tencent/hpc-ops HPC-Ops is a production-grade C++/CUDA operator library for high-performance LLM inference, developed by Tencent's Hunyuan AI Infra team. I… | 59 | 1131 | active |
| auto-novel/auto-novel AutoNovel is a website application that automatically machine-translates light novels (web novels, bunko novels, and local files) using LLM… | 77 | 1130 | active |
| yohanshin/WHAM WHAM is the official PyTorch implementation of the CVPR 2024 paper 'Reconstructing World-grounded Humans with Accurate 3D Motion'. It estim… | 26 | 1130 | active |
| PriesiaMioShirakana/DragonianVoice A C++ inference library for running ONNX-based TTS, SVC (singing voice conversion), and SVS (singing voice synthesis) models, supporting ar… | 40 | 1129 | active |
| ikostrikov/pytorch-a2c-ppo-acktr-gail A PyTorch implementation of several deep reinforcement learning algorithms: A2C, PPO, ACKTR, and GAIL (imitation learning). It works with O… | 32 | 3903 | maintenance |
| Dobiasd/frugally-deep frugally-deep is a lightweight header-only C++ library for running inference (forward passes) on Keras/TensorFlow models without linking ag… | 85 | 1128 | active |
| atiilla/GeoIntel GeoIntel is a Python tool that uses Google's Gemini API to estimate where a photo was taken through AI-powered geolocation analysis. It off… | 58 | 1128 | active |
| open-mmlab/mmtracking MMTracking is OpenMMLab's PyTorch-based toolbox for video perception tasks, unifying video object detection, multiple object tracking, sing… | 23 | 3897 | maintenance |
| NovaSearch-Team/RAG-Retrieval A Python library and toolkit for unified fine-tuning, inference, and distillation of RAG retrieval models, including embedding models, ColB… | 60 | 1127 | active |
| BeingBeyond/Being-H Being-H is a family of human-centric embodied foundation models, including VLA models (Being-H0.5, Being-H0) and latent world-action models… | 63 | 1126 | active |
| brownhci/WebGazer WebGazer.js is a JavaScript eye tracking library that uses a standard webcam to predict a user's gaze location on a web page in real time. … | 65 | 3888 | maintenance |
| lkwq007/stablediffusion-infinity A web application for outpainting with Stable Diffusion on an effectively infinite canvas, built with PyScript and Gradio. It extends image… | 32 | 3887 | maintenance |
| espressif/esp-dl ESP-DL is Espressif's lightweight neural network inference framework for ESP-series chips, with a custom .espdl model format, quantization … | 72 | 1125 | active |
| OpenGVLab/VideoMamba VideoMamba is a state space model (Mamba-based) architecture for efficient video understanding, released with code and pretrained models fr… | 25 | 1125 | active |
| hijohnnylin/neuronpedia Neuronpedia is an open source AI interpretability platform for exploring, visualizing, and steering the internals of language models. It ho… | 92 | 1124 | active |
| geopavlakos/hamer HaMeR (Hand Mesh Recovery) is a transformer-based model that reconstructs 3D hand meshes from single monocular images using the MANO parame… | 56 | 1124 | active |
| OpenDriveLab/UniVLA UniVLA is an open-source framework for training cross-embodiment vision-language-action (VLA) robot policies using task-centric latent acti… | 43 | 1124 | active |
| hanruihua/ir-sim IR-SIM is an open-source, Python-based lightweight robot simulator for navigation, control, and learning. It provides a simple YAML-driven … | 100 | 1122 | active |
| cggos/imu_x_fusion A C++/ROS library implementing loosely-coupled IMU + X (GNSS, 6DoF odometry) fusion localization using ESKF, IEKF, UKF variants, and MAP es… | 38 | 1122 | active |
| THUDM/SwissArmyTransformer SwissArmyTransformer (sat) is a PyTorch library for developing custom Transformer model variants where models like BERT, GPT, T5, GLM, and … | 23 | 1121 | active |
| premieroctet/photoshot Photoshot is an open-source web application that generates custom AI avatars from user-uploaded selfies using fine-tuned text-to-image mode… | 22 | 3870 | maintenance |
| FlagAI-Open/FlagAI FlagAI is a Python toolkit for training, fine-tuning, and deploying large-scale AI models across NLP, CV, and vision-language tasks. It int… | 64 | 3869 | maintenance |