function: machine-learning
5378 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| cambrian-mllm/cambrian Cambrian-1 is a fully open family of vision-centric multimodal large language models (MLLMs) from NYU's VISIONx group, with training and ev… | 47 | 2013 | active |
| philipperemy/keras-tcn A Keras/TensorFlow implementation of Temporal Convolutional Networks (TCN) with dilated causal convolutions, usable as a drop-in layer alte… | 62 | 2012 | active |
| tatsu-lab/alpaca_eval AlpacaEval is an LLM-based automatic evaluation framework for instruction-following language models, producing win rates against a GPT-4 ba… | 36 | 2012 | active |
| uncertainty-toolbox/uncertainty-toolbox A Python library for predictive uncertainty quantification, providing metrics, visualizations, and recalibration procedures for regression … | 27 | 2012 | active |
| lyhue1991/torchkeras torchkeras is a lightweight PyTorch model training template library that brings Keras-style compile/fit/evaluate APIs to PyTorch. Its core … | 55 | 2009 | active |
| AntixK/PyTorch-VAE A collection of Variational Autoencoder (VAE) model implementations in PyTorch, including Beta-VAE, VQ-VAE, IWAE, WAE, and others, with a f… | 38 | 7665 | maintenance |
| kohya-ss/musubi-tuner Musubi Tuner is a set of Python scripts for training LoRA (Low-Rank Adaptation) adapters for video and image generation model architectures… | 85 | 2002 | active |
| WhatDreamsCost/WhatDreamsCost-ComfyUI A collection of free custom ComfyUI nodes and workflows, centered on LTX Director, a timeline-based tool for directing LTX video generation… | 57 | 2001 | active |
| bytetriper/RAE Official PyTorch implementation of 'Diffusion Transformers with Representation Autoencoders' (RAE), a two-stage image generation pipeline u… | 48 | 2001 | active |
| zai-org/GLM-130B GLM-130B is an open bilingual (English and Chinese) 130-billion-parameter dense language model pre-trained with the General Language Model … | 32 | 7651 | maintenance |
| apple/coreai-models Apple's repository of model export recipes, Python primitives, and Swift runtime utilities for building on-device AI with the Core AI frame… | 67 | 1999 | active |
| mil-tokyo/webdnn WebDNN is a framework for running deep neural network inference directly in the web browser, accepting ONNX models without Python preproces… | 65 | 1999 | active |
| DEIM DEIMv2 is a real-time object detection framework that extends the DEIM DETR family with DINOv3-pretrained and distilled backbones plus a Sp… | 62 | 1999 | active |
| jianchang512/vocal-separate A minimal local web-based tool for separating vocals from background music in audio or video files, using Spleeter 2stems/4stems/5stems mod… | 10 | 1998 | active |
| elyra-ai/elyra Elyra is a set of AI-centric extensions for JupyterLab, including a visual pipeline editor for building and executing notebook-based pipeli… | 68 | 1996 | active |
| clementchadebec/benchmark_VAE Pythae is a PyTorch library that unifies implementations of many Variational Autoencoder (VAE) variants under a common interface, enabling … | 23 | 1995 | active |
| facebookresearch/dino PyTorch implementation of DINO, a self-supervised learning method for training Vision Transformers, with pretrained model weights. It is th… | 10 | 7611 | maintenance |
| lightseekorg/tokenspeed TokenSpeed is a high-performance LLM inference engine designed for agentic workloads, aiming for TensorRT-LLM-level performance with vLLM-l… | 68 | 1989 | active |
| Morizeyao/GPT2-Chinese A Python library providing GPT-2 training and text generation code tailored for Chinese, built on HuggingFace Transformers with BERT or BPE… | 32 | 7597 | maintenance |
| allenai/scispacy scispaCy is a Python library providing full spaCy pipelines, models, and custom pipes for processing scientific, biomedical, and clinical t… | 54 | 1988 | active |
| Hyperopt Hyperopt is a Python library for distributed asynchronous hyperparameter optimization over search spaces with real-valued, discrete, and co… | 86 | 7592 | maintenance |
| Lumiwealth/lumibot Lumibot is a Python framework for building, backtesting, and deploying algorithmic trading strategies and AI-agent trading teams across sto… | 95 | 1987 | active |
| logpai/logparser Logparser is a Python machine learning toolkit and benchmark suite for automated log parsing. It extracts event templates from unstructured… | 34 | 1987 | active |
| chaidiscovery/chai-lab Chai-1 is a state-of-the-art multi-modal foundation model for biomolecular structure prediction, handling proteins, small molecules, DNA, R… | 65 | 1986 | active |
| featureform/featureform Featureform is a virtual feature store that sits atop your existing data infrastructure and orchestrates it to define, manage, and serve ML… | 36 | 1985 | active |
| xingyizhou/CenterNet CenterNet is a PyTorch implementation of the 'Objects as Points' detector, which models objects as single center points detected via keypoi… | 32 | 7573 | maintenance |
| hkchengrex/XMem XMem is a PyTorch model for semi-supervised video object segmentation that tracks objects through long videos using an Atkinson-Shiffrin-in… | 23 | 1983 | stable |
| kwuking/TimeMixer Official PyTorch implementation of TimeMixer, an ICLR 2024 model for time series forecasting using decomposable multiscale mixing. It has s… | 46 | 1981 | active |
| lucidrains/titans-pytorch An unofficial PyTorch implementation of the Titans architecture, a neural long-term memory module for transformers that learns to memorize … | 70 | 1980 | active |
| davda54/sam An unofficial PyTorch implementation of Sharpness-Aware Minimization (SAM) and its adaptive variant ASAM, provided as an optimizer wrapper … | 32 | 1980 | stable |
| patrikhuber/eos A lightweight, header-only 3D Morphable Face Model (3DMM) fitting library written in modern C++11/14, with Python bindings. It provides mod… | 31 | 1980 | active |
| google-deepmind/alphagenome A Python SDK providing programmatic access to Google DeepMind's AlphaGenome model, which predicts genomic regulatory outputs such as gene e… | 89 | 1979 | active |
| cloneofsimo/lora A Python library for applying Low-Rank Adaptation (LoRA) to quickly fine-tune text-to-image diffusion models like Stable Diffusion. It prod… | 22 | 7550 | maintenance |
| adobe-research/custom-diffusion Custom Diffusion is a research codebase for efficiently fine-tuning text-to-image diffusion models like Stable Diffusion on a few example i… | 69 | 1978 | stable |
| JIA-Lab-research/DreamOmni2 DreamOmni2 is the official PyTorch implementation of a CVPR 2026 Highlight model for multimodal instruction-based image editing and generat… | 51 | 1978 | active |
| Linzaer/Ultra-Light-Fast-Generic-Face-Detector-1MB An ultra-lightweight face detection model (~1MB FP32, ~300KB quantized) designed for edge computing devices, with slim and RFB variants tra… | 32 | 7542 | maintenance |
| PrimeIntellect-ai/prime-rl prime-rl is a Python framework for large-scale, fully asynchronous reinforcement learning training of language models, built on FSDP2 for t… | 87 | 1975 | active |
| haykgrigo3/TimeCapsuleLLM TimeCapsuleLLM is a research project training language models from scratch (and fine-tuning small base models) exclusively on text from spe… | 75 | 1975 | active |
| ZhuJHua/moodiary Moodiary is a fully open-source, cross-platform diary/journaling app built with Flutter and Rust. It supports markdown, plain text, and ric… | 69 | 1975 | active |
| tandpfun/wardrobe A self-hosted web application that detects garments in photos, extracts clean product cutouts, and generates modeled editorial previews usi… | 54 | 1975 | active |
| openlm-research/open_llama OpenLLaMA is a permissively licensed (Apache-2.0) open reproduction of Meta AI's LLaMA, releasing 3B, 7B, and 13B pretrained model weights … | 30 | 7528 | maintenance |
| showlab/Show-o Show-o is a research repository implementing a unified transformer model that combines autoregressive and discrete diffusion modeling for m… | 50 | 1973 | active |
| android/androidify An open-source Android sample app from Google that lets users create custom Android bot avatars using AI image generation via the Gemini AP… | 68 | 1972 | active |
| WassimTenachi/PhySO PhySO is a Python library for physical symbolic optimization that uses deep reinforcement learning to discover analytical physical laws fro… | 52 | 1971 | active |
| FireRedTeam/FireRedASR FireRedASR is a family of open-source industrial-grade automatic speech recognition models supporting Mandarin, Chinese dialects, and Engli… | 52 | 1971 | active |
| theamusing/perfectPixel A Python library that automatically detects the optimal grid size in AI-generated pixel art images and refines them into clean, perfectly a… | 45 | 1971 | active |
| bigcode-project/starcoder StarCoder is a 15B-parameter code language model trained on 80+ programming languages, and this repository hosts the fine-tuning and infere… | 30 | 7506 | maintenance |
| google-deepmind/tapnet Google DeepMind's official repository for Tracking Any Point (TAP), containing the TAP-Vid and TAPVid-3D benchmarks, the TAPIR and TAPNext … | 74 | 1968 | active |
| meta-recsys/generative-recommenders Meta's research library implementing HSTU and M-FALCON from the ICML'24 paper 'Actions Speak Louder than Words: Trillion-Parameter Sequenti… | 69 | 1966 | active |
| Netflix/void-model VOID (Video Object and Interaction Deletion) is a research model from Netflix that removes objects from videos along with the physical inte… | 54 | 1965 | active |
| siliconflow/onediff OneDiff is an out-of-the-box acceleration library for diffusion models, providing PyTorch compilation tools and optimized GPU kernels. It i… | 48 | 1964 | active |
| official-pikafish/Pikafish Pikafish is a free, open-source UCI xiangqi (Chinese chess) engine derived from Stockfish, using NNUE neural network evaluation to analyze … | 76 | 1963 | active |
| NVIDIA-NeMo/RL NeMo RL is NVIDIA's open-source post-training library for scaling reinforcement learning methods (GRPO, PPO, DPO, SFT, distillation) on LLM… | 81 | 1961 | active |
| OpenMotionLab/MotionGPT MotionGPT is a unified motion-language model that treats 3D human motion as a foreign language by converting motion into discrete motion to… | 32 | 1961 | active |
| 2U1/Qwen-VL-Series-Finetune An open-source Python repository providing training scripts for fine-tuning Alibaba's Qwen-VL series of vision-language models (Qwen2-VL, Q… | 67 | 1960 | active |
| open-mmlab/mmagic MMagic is OpenMMLab's toolbox for generative and multimodal AI image/video creation, built on PyTorch. It provides a large model zoo coveri… | 23 | 7457 | maintenance |
| CannyLab/tsne-cuda A CUDA-accelerated implementation of the FIt-SNE t-SNE algorithm with Python bindings, offering up to 1200x speedup over scikit-learn. It e… | 95 | 1957 | stable |
| kadirnar/whisper-plus A Python library wrapping OpenAI Whisper-family models (including distil-whisper and MLX variants) for fast speech-to-text transcription wi… | 60 | 1956 | active |
| SizheAn/PanoHead PanoHead is the official PyTorch implementation of a CVPR 2023 paper presenting a 3D-aware GAN that synthesizes geometry-aware, view-consis… | 29 | 1956 | active |
| alibaba/EasyCV EasyCV is an all-in-one PyTorch-based computer vision toolkit from Alibaba covering self-supervised learning, vision transformers, and majo… | 32 | 1954 | active |
| eriklindernoren/PyTorch-YOLOv3 A minimal PyTorch implementation of YOLOv3 supporting training, inference, and evaluation, with compatibility for YOLOv4 and YOLOv7 weights… | 32 | 7440 | maintenance |
| Fafa-DL/Awesome-Backbones A PyTorch-based framework that integrates many deep learning backbone models (CNNs and vision transformers like ResNet, EfficientNet, Swin … | 33 | 1953 | active |
| meta-pytorch/opacus Opacus is a PyTorch library for training neural networks with differential privacy via DP-SGD, requiring minimal code changes through its P… | 79 | 1952 | active |
| autodiff/autodiff autodiff is a C++17 library for automatic differentiation that computes derivatives of functions efficiently using forward mode (dual numbe… | 24 | 1952 | active |
| Yuliang-Liu/Monkey Monkey is a large multi-modal model (LMM) research project from CVPR 2024 that improves image understanding via higher input resolution and… | 65 | 1951 | active |
| haoheliu/versatile_audio_super_resolution AudioSR is a Python library and CLI tool that performs versatile audio super-resolution, upsampling any audio (music, speech, sound effects… | 45 | 1951 | active |
| LTH14/mar Official PyTorch implementation of MAR (Masked Autoregressive) image generation with DiffLoss, from the NeurIPS 2024 paper 'Autoregressive … | 54 | 1949 | stable |
| openai/guided-diffusion OpenAI's codebase for guided diffusion models from the paper 'Diffusion Models Beat GANs on Image Synthesis', including classifier conditio… | 32 | 7419 | maintenance |
| Vchitect/Latte Official PyTorch implementation of Latte, a latent diffusion transformer for video generation. It includes model definitions, pre-trained c… | 71 | 1948 | active |
| cra-ros-pkg/robot_localization robot_localization is a ROS package of nonlinear state estimation nodes (EKF, UKF, and navsat_transform) for fusing IMU, odometry, and othe… | 64 | 1947 | stable |
| kyegomez/BitNet A PyTorch implementation of the BitNet architecture from the paper 'BitNet: Scaling 1-bit Transformers for Large Language Models', providin… | 72 | 1945 | active |
| Graph Convolutional Networks (GCN) A TensorFlow implementation of Graph Convolutional Networks (GCN) for semi-supervised node classification on graphs, accompanying the ICLR … | 32 | 7400 | maintenance |
| jd-opensource/JoyAI-Echo JoyAI-Echo is a Python framework for long-horizon audio-visual generation, producing coherent multi-shot videos up to ~5 minutes with paire… | 58 | 1943 | active |
| tensorlayer/TensorLayer TensorLayer is a TensorFlow-based deep learning and reinforcement learning library offering customizable neural layers for researchers and … | 23 | 7381 | maintenance |
| PixArt-alpha/PixArt-sigma PixArt-Σ is a PyTorch implementation of a diffusion transformer model for high-resolution (up to 4K) text-to-image generation, trained with… | 25 | 1939 | active |
| starik222/BooruDatasetTagManager A desktop tag editor for managing booru-style tagged image and video datasets used to train Stable Diffusion models such as LoRAs, embeddin… | 76 | 1938 | active |
| microsoft/Magma Magma is Microsoft Research's foundation model for multimodal AI agents, released as an 8B vision-language model that understands images an… | 53 | 1937 | active |
| JuliaAI/MLJ.jl MLJ.jl is a machine learning framework for Julia providing a common interface to over 200 models, with meta-algorithms for model selection,… | 93 | 1935 | stable |
| NVlabs/RADIO Official PyTorch implementation of AM-RADIO and its successors (RADIOv2.5, C-RADIOv4), agglomerative vision foundation models distilled fro… | 64 | 1933 | active |
| PAIR-code/facets Facets is a pair of web-component visualizations (Overview and Dive) for understanding and analyzing machine learning datasets, embeddable … | 10 | 7336 | maintenance |
| Tencent-Hunyuan/HunyuanOCR HunyuanOCR-1.5 is a lightweight end-to-end OCR vision-language model from Tencent, with a unified inference environment, llama.cpp PC-side … | 59 | 1930 | active |
| tum-pbs/PhiFlow PhiFlow is an open-source Python simulation toolkit for solving partial differential equations with support for optimization and machine le… | 72 | 1929 | active |
| Audio-AGI/AudioSep AudioSep is the official implementation of the 'Separate Anything You Describe' foundation model for open-domain, language-queried audio so… | 28 | 1929 | active |
| Yuanshi9815/OminiControl OminiControl is a universal control framework for Diffusion Transformer models like FLUX, supporting subject-driven and spatial control (ed… | 62 | 1927 | active |
| RightNow-AI/picolm PicoLM is a pure C11 LLM inference engine that runs 1-billion parameter models in GGUF format on extremely constrained hardware like $10 bo… | 46 | 1921 | active |
| OpenTalker/video-retalking VideoReTalking is a Python research system from SIGGRAPH Asia 2022 that edits real-world talking-head videos to match a given audio track, … | 23 | 7280 | maintenance |
| GPflow/GPflow GPflow is a Python library for building Gaussian process models on top of TensorFlow 2 and TensorFlow Probability. It implements modern Gau… | 95 | 1916 | active |
| pymatting/pymatting PyMatting is a Python library for alpha matting that estimates an alpha matte from an input image and a hand-drawn trimap to extract foregr… | 67 | 1914 | active |
| deepseek-ai/DeepSeek-LLM DeepSeek LLM is a family of open-source large language models (7B and 67B, Base and Chat variants) trained from scratch on 2 trillion token… | 27 | 7257 | maintenance |
| FACEGOOD/FACEGOOD-Audio2Face FACEGOOD Audio2Face is an open-source deep learning framework that converts audio into facial blendshape weights for driving digital humans… | 64 | 1909 | active |
| diffgram/diffgram Diffgram is a self-hosted AI datastore for managing schemas, BLOBs, and predictions, with built-in human supervision (data labeling), data … | 62 | 1909 | active |
| sicxu/Deep3DFaceRecon_pytorch A PyTorch implementation of Deep3DFaceReconstruction, a weakly-supervised CNN method for reconstructing 3D face geometry from a single imag… | 32 | 1907 | stable |
| NVlabs/nvdiffrast Nvdiffrast is a PyTorch library from NVIDIA providing high-performance, GPU-accelerated primitive operations for rasterization-based differ… | 55 | 1905 | stable |
| visual-layer/fastdup fastdup is a free Python tool for rapidly analyzing image and video datasets to surface duplicates, outliers, broken, dark, bright, blurry,… | 67 | 1904 | active |
| Faceplugin-ltd/Open-Source-Face-Recognition-SDK An open-source face recognition SDK by Faceplugin providing face detection, landmark extraction, feature embedding generation, and face tem… | 64 | 1903 | active |
| lucidrains/byol-pytorch A PyTorch library implementing the Bootstrap Your Own Latent (BYOL) self-supervised learning method from DeepMind. It wraps any image-based… | 58 | 1903 | active |
| google-deepmind/penzai Penzai is a JAX research toolkit for building, editing, and visualizing neural networks as legible, functional pytree data structures. It i… | 10 | 1901 | active |
| TheSpaghettiDetective/obico-server Obico Server is the self-hostable backend of the Obico smart 3D printing platform, providing AI-based print failure detection, webcam strea… | 76 | 1899 | active |
| sapientinc/HRM-Text HRM-Text is a 1B-parameter text generation model based on the hierarchical recurrent HRM architecture, released with a complete pretraining… | 53 | 1899 | active |
| flexflow/flexflow-train FlexFlow Train is a deep learning framework that accelerates distributed DNN training by automatically searching for efficient parallelizat… | 67 | 1898 | active |