Ross ROSS = Recommend OSS · open-source software intelligence for agents

TsinghuaC3I/Awesome-RL-for-LRMs resource

A Survey of Reinforcement Learning for Large Reasoning Models observed · 2026-08-28

github.com/TsinghuaC3I/Awesome-RL-for-LRMs · homepage · TeX · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

57/100

  • Activity 98
  • Release rhythm 15
  • Longevity 37
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 531
  • days_rel: 357
  • days_push: 13
  • n_releases_24m: 1

Full methodology

Adoption not part of the score

2482 stars · 133 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A curated awesome-list accompanying the survey paper 'A Survey of Reinforcement Learning for Large Reasoning Models' from Tsinghua University. It catalogs research papers on applying reinforcement learning to LLMs and Large Reasoning Models (LRMs), organized by the survey's category structure.

Use cases

  • find papers on RL for LLM reasoning
  • survey reinforcement learning methods for large reasoning models
  • research RL training after DeepSeek-R1
  • find resources on RLHF and reasoning model training
  • track recent research on LRM scalability and RL algorithms
  • get started with academic literature on reasoning models

When to choose

  • you need a curated, categorized reading list of RL-for-reasoning research
  • you are writing a literature review on LLM reasoning and RL
  • you want to follow the state of the art since DeepSeek-R1

When to avoid

  • you need runnable training code or an RL framework rather than a paper list
  • you want beginner tutorials rather than research papers

Facets

learning-resource · maturity active

reinforcement-learning llm-training machine-learning reinforcement-learning large-language-models artificial-intelligence tutorials awesome-lists cross-platform awesome-list survey reasoning large-reasoning-models deepseek-r1 paper-list rlhf

6 sources

Member repositories

RepositoryRoleHealth v2
TsinghuaC3I/Awesome-RL-for-LRMsmain57

For agents

markdown · JSON · MCP: product_card(name="TsinghuaC3I/Awesome-RL-for-LRMs")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem