opendilab/awesome-RLHF resource
A curated list of reinforcement learning with human feedback resources (continually updated) observed · 2026-08-28
Health v2 · maintenance only
68/100
- Activity 83
- Release rhythm 35
- Longevity 92
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 1297
- days_rel: n/a
- days_push: 105
- n_releases_24m: 0
Adoption not part of the score
4422 stars · 258 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
A curated, continually updated list of research papers, codebases, datasets, blogs, and books about Reinforcement Learning with Human Feedback (RLHF). It tracks the frontier of RLHF for both large language models and applications like video games.
Use cases
- find papers on RLHF for LLM alignment
- learn how reinforcement learning with human feedback works
- find open-source RLHF codebases and datasets
- track the latest RLHF research in 2024-2026
- study reward modeling and human preference learning
- get reading material for aligning language models with human values
When to choose
- you need a starting point to survey the RLHF literature
- you want curated links to papers, code, and datasets in one place
- you are researching LLM alignment or preference-based RL
When to avoid
- you need runnable RLHF training code rather than a resource list
- you need a maintained software library with an API
- you need non-RLHF reinforcement learning resources
Facets
learning-resource · maturity active
reinforcement-learning llm-training machine-learning reinforcement-learning large-language-models machine-learning artificial-intelligence tutorials cross-platform awesome-list rlhf curated-list research-papers human-feedback alignment
1 source
- readme: https://github.com/opendilab/awesome-RLHF · fetched 2026-08-28 · 76d0ad05abfd
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| opendilab/awesome-RLHF | main | 68 |
For agents
markdown · JSON · MCP: product_card(name="opendilab/awesome-RLHF")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem