Ross ROSS = Recommend OSS · open-source software intelligence for agents

BytedTsinghua-SIA/DAPO

An Open-source RL System from ByteDance Seed and Tsinghua AIR observed · 2026-08-28

github.com/BytedTsinghua-SIA/DAPO · Python observed · 2026-08-28

Health v2 · maintenance only

29/100

  • Activity 21
  • Release rhythm 35
  • Longevity 38

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 534
  • days_rel: n/a
  • days_push: 479
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1861 stars · 85 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

DAPO is an open-source reinforcement learning system for large-scale LLM training, released by ByteDance Seed and Tsinghua AIR. It implements the Decoupled Clip and Dynamic Sampling Policy Optimization algorithm on top of the verl framework, with code, datasets, and model weights achieving 50% on AIME 2024.

Use cases

  • train llm with reinforcement learning
  • reproduce dapo rl training on qwen models
  • run rlhf-style policy optimization for math reasoning
  • evaluate a 32b model on aime 2024
  • research scalable rl for large language models

When to choose

  • you want to reproduce or extend state-of-the-art LLM RL training
  • you need the DAPO algorithm, dataset, and checkpoints together
  • you already use the verl framework and want proven RL recipes

When to avoid

  • you need a production-ready training platform with support guarantees
  • you lack multi-GPU infrastructure for large-scale RL
  • you need a permissively licensed codebase, since no license is specified

Facets

library · maturity active

llm-training reinforcement-learning machine-learning benchmarking large-language-models reinforcement-learning deep-learning machine-learning python rlhf policy-optimization verl math-reasoning aime research-code gpu linux docker

1 source

Member repositories

RepositoryRoleHealth v2
BytedTsinghua-SIA/DAPOmain29

For agents

markdown · JSON · MCP: product_card(name="BytedTsinghua-SIA/DAPO")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem