function: llm-training
818 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| transformerlab/transformerlab-app Transformer Lab is an open-source desktop application (built with Electron and Python) that provides a unified GUI for training, fine-tunin… | 84 | 5179 | active |
| h2oai/h2o-llmstudio H2O LLM Studio is a framework and no-code GUI for fine-tuning state-of-the-art large language models, built by H2O.ai. It supports LoRA and… | 96 | 5172 | active |
| hiyouga/EasyR1 EasyR1 is an efficient, scalable reinforcement learning training framework for large language models and vision-language models, built as a… | 65 | 5129 | active |
| lightvector/KataGo KataGo is an open-source Go (baduk) engine trained via AlphaZero-like self-play, one of the strongest Go bots available. It runs as a GTP e… | 99 | 5036 | active |
| Plachtaa/VITS-fast-fine-tuning A Python pipeline for fast fine-tuning of VITS text-to-speech models, enabling speaker adaptation in under an hour from short audio, long a… | 10 | 5012 | active |
| facebookresearch/lingua Meta Lingua is a minimal, fast LLM training and inference library built on easy-to-modify PyTorch components for research purposes. It supp… | 36 | 4766 | active |
| FluxML/Flux.jl Flux.jl is a machine learning library written entirely in Julia, providing lightweight abstractions over Julia's native GPU support and aut… | 98 | 4739 | stable |
| mindspore-ai/mindspore MindSpore is an open-source deep learning framework for training and inference across mobile, edge, and cloud scenarios. It provides automa… | 32 | 4700 | active |
| PKU-Alignment/align-anything Align-Anything is a modular Python framework for aligning any-to-any (all-modality) large models with human intentions and values using fee… | 48 | 4667 | active |
| harbor-framework/harbor Harbor is a Python framework from the creators of Terminal-Bench for evaluating and optimizing AI agents and language models in containeriz… | 85 | 4664 | active |
| RLinf/RLinf RLinf is an open-source, flexible and scalable reinforcement learning training infrastructure for embodied AI (vision-language-action model… | 74 | 4655 | active |
| deepseek-ai/Engram Official implementation of Engram, a conditional memory module from DeepSeek that modernizes N-gram embeddings for O(1) lookup as a new spa… | 43 | 4614 | active |
| TencentCloudADP/youtu-agent Youtu-Agent is a Python framework for building, running, and evaluating autonomous LLM agents, supporting ReAct-style single agents and pla… | 57 | 4604 | active |
| thirdlayerinc/autoagent AutoAgent is a Python-based meta-agent harness that autonomously builds and iterates on AI agent harnesses. It edits an agent's system prom… | 48 | 4566 | active |
| PrimeIntellect-ai/verifiers verifiers is a Python library for creating environments to train and evaluate large language models with reinforcement learning. It integra… | 88 | 4559 | active |
| mosaicml/llm-foundry LLM Foundry is a PyTorch-based codebase for training, finetuning, evaluating, and deploying large language models from 125M to 70B+ paramet… | 65 | 4441 | active |
| microsoft/FLAML FLAML is a lightweight Python library for automated machine learning (AutoML) and hyperparameter tuning. It efficiently finds quality model… | 89 | 4390 | active |
| ymcui/Chinese-LLaMA-Alpaca A project releasing Chinese-adapted LLaMA and instruction-tuned Alpaca large language models with extended Chinese vocabulary, plus pre-tra… | 56 | 18931 | maintenance |
| tloen/alpaca-lora A repository of scripts for reproducing Stanford Alpaca-style instruction tuning of LLaMA models using low-rank adaptation (LoRA) on consum… | 30 | 18906 | maintenance |
| OpenManus/OpenManus-RL OpenManus-RL is an open-source project for reinforcement-learning-based tuning of LLM agents, built on the verl training framework. It prov… | 46 | 4155 | active |
| IDEA-CCNL/Fengshenbang-LM Fengshenbang-LM is an open-source suite of Chinese large language models and a PyTorch training framework from IDEA Research's CCNL lab, ai… | 71 | 4123 | active |
| StarsfieldAI/R1-V R1-V is an open-source research codebase for training vision-language models with reinforcement learning (RLVR/GRPO), demonstrating strong … | 21 | 4063 | active |
| FedML-AI/FedML FedML (TensorOpera) is a unified Python library for large-scale distributed training, model serving, and federated learning across GPU clou… | 45 | 4062 | active |
| marcelroed/gigatoken Gigatoken is a Rust-based tokenizer library for language modeling that tokenizes text at GB/s throughput, claiming ~1000x speedups over Hug… | 60 | 4061 | active |
| thinking-machines-lab/tinker-cookbook Tinker Cookbook is a Python library of realistic examples and abstractions for post-training (fine-tuning) language models via the Tinker A… | 85 | 4058 | active |
| tensorflow/tensor2tensor Tensor2Tensor (T2T) is a Python library of deep learning models and datasets built on TensorFlow, developed by the Google Brain team to mak… | 10 | 17464 | maintenance |
| huggingface/smollm Hugging Face's repository for the SmolLM and SmolVLM families of compact, fully open language and vision-language models, including trainin… | 59 | 3884 | active |
| hkust-nlp/simpleRL-reason A research codebase from HKUST-NLP implementing a simple reinforcement learning recipe (rule-based rewards on GSM8K/Math data) to train LLM… | 47 | 3874 | active |
| vllm-project/llm-compressor LLM Compressor is a Python library for applying quantization and pruning algorithms to large language models, producing compressed-tensors … | 90 | 3726 | active |
| Stability-AI/StableLM StableLM is Stability AI's repository of open-weight decoder-only transformer language models, including the 3B-parameter StableLM-3B-4E1T … | 30 | 15684 | maintenance |
| NExT-GPT/NExT-GPT NExT-GPT is an end-to-end any-to-any multimodal large language model that accepts and generates arbitrary combinations of text, image, vide… | 37 | 3638 | active |
| google-research/big_vision Google Research's official Jax/Flax codebase for training large-scale vision models such as Vision Transformer, SigLIP, MLP-Mixer, and LiT … | 42 | 3528 | active |
| pathwaycom/bdh BDH (Dragon Hatchling) is a biologically inspired large language model architecture that bridges deep learning and neuroscience, implemente… | 54 | 3519 | active |
| EverMind-AI/MSA MSA (Memory Sparse Attention) is a Python framework for end-to-end trainable sparse latent-memory attention that scales LLM context to 100M… | 53 | 3515 | active |
| NVIDIA/TransformerEngine Transformer Engine is an NVIDIA library for accelerating Transformer model training and inference on NVIDIA GPUs using low-precision format… | 99 | 3504 | active |
| aiming-lab/MetaClaw MetaClaw is a continual meta-learning framework that lets an LLM agent evolve from real conversations, combining skill synthesis from failu… | 68 | 3497 | active |
| sentient-agi/OML-1.0-Fingerprinting A Python library for embedding secret cryptographic fingerprints into LLMs via fine-tuning, so model owners can prove ownership and detect … | 13 | 3497 | active |
| huggingface/optimum Optimum is a Hugging Face library that extends Transformers, Diffusers, timm, and Sentence Transformers with hardware-specific optimization… | 98 | 3469 | active |
| NVlabs/Eagle Eagle is NVIDIA's family of frontier vision-language models (Eagle, Eagle 2, Eagle 2.5) built with data-centric training strategies, plus L… | 64 | 3462 | active |
| PaddlePaddle/PARL PARL is a flexible, high-performance reinforcement learning framework built on PaddlePaddle, providing Model/Algorithm/Agent abstractions a… | 41 | 3453 | active |
| aqlaboratory/openfold OpenFold is a faithful, trainable PyTorch reproduction of DeepMind's AlphaFold 2 for protein structure prediction. It is memory-efficient a… | 48 | 3420 | active |
| NovaSky-AI/SkyThought SkyThought is the open-source repository behind Sky-T1, a family of reasoning language models trained for under $450, including training sc… | 25 | 3399 | active |
| alibaba/ROLL ROLL is an open-source reinforcement learning library from Alibaba for training large language models at scale, supporting algorithms like … | 78 | 3374 | active |
| microsoft/LoRA loralib is the official PyTorch implementation of LoRA (Low-Rank Adaptation), which fine-tunes large language models by injecting trainable… | 23 | 13767 | maintenance |
| cocktailpeanut/fluxgym FluxGym is a simple web UI for training FLUX LoRA models with low VRAM support (12GB/16GB/20GB). It combines the AI-Toolkit Gradio frontend… | 65 | 3246 | active |
| determined-ai/determined Determined is an open-source deep learning platform that combines distributed training, hyperparameter tuning, experiment tracking, and GPU… | 39 | 3236 | active |
| NVIDIA/physicsnemo NVIDIA PhysicsNeMo is an open-source Python deep-learning framework for building, training, fine-tuning, and inferring physics AI models us… | 89 | 3198 | active |
| Nerogar/OneTrainer OneTrainer is a GUI and CLI application for fine-tuning diffusion image models, supporting full fine-tuning, LoRA, and embeddings across ma… | 74 | 3184 | active |
| tekaratzas/RustGPT A transformer-based large language model implemented entirely in pure Rust with no external ML frameworks, using only ndarray for matrix op… | 38 | 3157 | active |
| MakazhanAlpamys/Soup Soup is a Python CLI that fine-tunes and post-trains LLMs from a single YAML config, supporting 23 methods (SFT, DPO, ORPO, SimPO, KTO, etc… | 77 | 3102 | active |
| ridgerchu/matmulfreellm A Python implementation of MatMul-Free LM, a language model architecture that eliminates matrix multiplication operations using ternary wei… | 49 | 3089 | active |
| deepseek-ai/DualPipe DualPipe is a Python library implementing a bidirectional pipeline parallelism algorithm that overlaps forward and backward computation wit… | 48 | 2998 | active |
| OpenMOSS/MOSS MOSS is an open-source tool-augmented conversational large language model from Fudan University, released with base models, SFT models, plu… | 68 | 12230 | maintenance |
| patrick-kidger/equinox Equinox is a Python library providing neural networks and scientific computing utilities for JAX, using PyTorch-like class-based syntax whe… | 92 | 2957 | stable |
| pytorch/ao TorchAO is a PyTorch-native library for model optimization through quantization and sparsity. It supports quantizing weights, gradients, op… | 89 | 2957 | active |
| karpathy/char-rnn char-rnn is a Torch/Lua implementation of multi-layer recurrent neural networks (RNN, LSTM, GRU) for character-level language modeling. It … | 32 | 12095 | maintenance |
| bghira/SimpleTuner SimpleTuner is a Python fine-tuning toolkit for image, video, and audio diffusion models built on Hugging Face Diffusers. It provides a web… | 92 | 2912 | active |
| zjunlp/EasyEdit EasyEdit is an easy-to-use knowledge editing framework for large language models, supporting methods that modify model parameters or steer … | 71 | 2908 | active |
| eric-mitchell/direct-preference-optimization A reference implementation of Direct Preference Optimization (DPO) for training language models from human preference data, built on Huggin… | 29 | 2907 | stable |
| learnables/learn2learn learn2learn is a PyTorch library for meta-learning research, providing utilities for few-shot task creation, high-level wrappers for algori… | 48 | 2893 | active |
| adapter-hub/adapters Adapters is a Python add-on library for HuggingFace Transformers that integrates 10+ parameter-efficient fine-tuning methods (bottleneck ad… | 80 | 2826 | active |
| physical-superintelligence-lab/Psi0 Psi-Zero (Ψ₀) is an open vision-language-action (VLA) foundation model for dexterous humanoid loco-manipulation, combining a Qwen3-VL backb… | 59 | 2802 | active |
| KellerJordan/Muon Muon is a PyTorch optimizer for the hidden layers of neural networks, based on orthogonalized momentum updates via Newton-Schulz iteration.… | 59 | 2801 | active |
| huggingface/nanotron Nanotron is a minimalistic Python library from Hugging Face for pretraining large language models with 3D parallelism (data, tensor, and pi… | 56 | 2800 | active |
| huggingface/setfit SetFit is a Python library for efficient, prompt-free few-shot fine-tuning of Sentence Transformers for text classification. It achieves hi… | 64 | 2784 | active |
| mll-lab-nu/RAGEN RAGEN is a Python framework for training reasoning LLM agents with multi-turn reinforcement learning using the StarPO algorithm. It also pr… | 65 | 2778 | active |
| artidoro/qlora QLoRA is the official implementation of the QLoRA paper, an efficient finetuning approach that backpropagates through a frozen 4-bit quanti… | 29 | 10998 | maintenance |
| roboflow/maestro maestro is a Python library from Roboflow that streamlines fine-tuning of multimodal vision-language models such as Florence-2, PaliGemma 2… | 62 | 2694 | active |
| qualcomm/aimet AIMET (AI Model Efficiency Toolkit) is a Python library from Qualcomm providing advanced quantization and compression techniques for traine… | 99 | 2688 | active |
| stochasticai/xTuring xTuring is a Python library for fine-tuning, evaluating, and running open-source large language models such as LLaMA, GPT-J, GPT-2, Qwen, a… | 52 | 2674 | active |
| ZHZisZZ/dllm dLLM is a Python library that unifies training, inference, and evaluation of diffusion language models such as LLaDA and Dream. It builds o… | 59 | 2672 | active |
| facebookresearch/ParlAI ParlAI is a Python framework from Facebook AI Research for sharing, training, and evaluating dialogue models across many openly available d… | 10 | 10621 | maintenance |
| plexe-ai/plexe Plexe is a Python library and CLI that builds machine learning models from natural language descriptions using a multi-agent AI workflow. Y… | 61 | 2614 | active |
| NVlabs/LongLive LongLive is an NVIDIA research framework providing parallel training and inference infrastructure for real-time long video generation, usin… | 60 | 2563 | active |
| ymcui/Chinese-BERT-wwm A collection of Chinese pre-trained language models (BERT-wwm, BERT-wwm-ext, RoBERTa-wwm-ext, RBT3, etc.) built with Whole Word Masking, re… | 67 | 10223 | maintenance |
| lamini-ai/lamini The official Python client and SDK for Lamini's hosted API for building generative AI applications. It lets developers call Lamini's LLM in… | 37 | 2534 | active |
| learning-at-home/hivemind Hivemind is a PyTorch library for decentralized deep learning across the Internet, enabling training of large models on hundreds of volunte… | 59 | 2515 | active |
| AMAP-ML/SkillClaw SkillClaw is a Python-based framework that lets AI agent skills evolve collectively from real interactions across sessions, agents, devices… | 58 | 2514 | active |
| KohakuBlueleaf/LyCORIS LyCORIS is a Python library implementing parameter-efficient fine-tuning algorithms (LoRA/LoCon, LoHa, LoKr, IA3, DyLoRA, and more) for Sta… | 73 | 2508 | active |
| yifan123/flow_grpo Flow-GRPO is the official PyTorch implementation of a NeurIPS 2025 paper that trains flow matching models (e.g., SD3.5, FLUX.1, Qwen-Image,… | 56 | 2498 | active |
| google-parfait/tensorflow-federated TensorFlow Federated (TFF) is an open-source Python framework for machine learning and other computations on decentralized data. It provide… | 79 | 2448 | active |
| Unakar/Logic-RL Logic-RL is a research framework that reproduces R1-Zero-style LLM reasoning via rule-based reinforcement learning (GRPO) on logic puzzles,… | 26 | 2448 | active |
| marin-community/marin Marin is an open-source Python framework and research program for training foundation models, covering the full pipeline from data curation… | 78 | 2428 | active |
| tflearn/tflearn TFLearn is a modular deep learning library providing a higher-level, Keras-like API on top of TensorFlow for building and training neural n… | 23 | 9576 | maintenance |
| google/tunix Tunix is a lightweight JAX-based library for post-training large language models, supporting supervised fine-tuning, preference optimizatio… | 83 | 2415 | active |
| AI-Hypercomputer/maxtext MaxText is a high-performance, scalable open-source LLM training library written in pure Python/JAX, targeting Google Cloud TPUs and GPUs. … | 97 | 2408 | active |
| microsoft/Olive Olive is Microsoft's AI model optimization toolkit for the ONNX Runtime, automating finetuning, conversion, quantization, and compression o… | 91 | 2382 | active |
| apple/axlearn AXLearn is a Python deep learning library built on JAX and XLA for developing and training large-scale models, with an object-oriented conf… | 71 | 2372 | active |
| alibaba/EasyRec EasyRec is a TensorFlow-based framework from Alibaba for building large-scale deep learning recommendation models covering matching, rankin… | 65 | 2356 | active |
| google-deepmind/optax Optax is a gradient processing and optimization library for JAX, offering composable building blocks like optimizers and loss functions. It… | 82 | 2325 | stable |
| facebookresearch/schedule_free A PyTorch library implementing schedule-free optimizers (SGD, AdamW, RAdam variants) that remove the need for learning rate schedules or sp… | 67 | 2323 | active |
| XiaomiMiMo/MiMo Xiaomi's MiMo is a 7B-parameter reasoning language model trained from pretraining through posttraining with reinforcement learning, release… | 30 | 2299 | active |
| togethercomputer/OpenChatKit OpenChatKit is an open-source toolkit from Together, LAION, and Ontocord.ai for building and fine-tuning chat language models, including in… | 30 | 8984 | maintenance |
| huggingface/picotron Picotron is a minimalist, hackable distributed training framework for pre-training Llama-like large language models using 4D parallelism (d… | 40 | 2289 | active |
| Liuziyu77/Visual-RFT Official research code for Visual-RFT and Visual-ARFT, applying GRPO-based reinforcement fine-tuning with rule-based verifiable rewards to … | 42 | 2271 | active |
| aixcoder-plugin/aiXcoder-7B Official repository for aiXcoder-7B, an open-weights 7B-parameter code large language model trained on 1.2T tokens for code completion, gen… | 29 | 2271 | active |
| OpenDCAI/DataFlex DataFlex is a data-centric training framework built on LLaMA-Factory that dynamically schedules LLM training data during optimization. It p… | 67 | 2270 | active |
| 666DZY666/micronet micronet is a Python library for deep neural network model compression and deployment built on PyTorch. It provides quantization (QAT, PTQ,… | 41 | 2266 | active |
| radixark/miles Miles is an open-source, enterprise-grade reinforcement learning framework for large-scale LLM and VLM post-training, forked from and co-ev… | 72 | 2263 | active |
| aws/sagemaker-python-sdk The SageMaker Python SDK is an open-source Python library for training and deploying machine learning models on Amazon SageMaker. It suppor… | 95 | 2261 | active |