Ross ROSS = Recommend OSS · open-source software intelligence for agents

opendilab/awesome-RLHF resource

A curated list of reinforcement learning with human feedback resources (continually updated) observed · 2026-08-28

github.com/opendilab/awesome-RLHF · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

68/100

  • Activity 83
  • Release rhythm 35
  • Longevity 92

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1297
  • days_rel: n/a
  • days_push: 105
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

4422 stars · 258 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

A curated, continually updated list of research papers, codebases, datasets, blogs, and books about Reinforcement Learning with Human Feedback (RLHF). It tracks the frontier of RLHF for both large language models and applications like video games.

Use cases

  • find papers on RLHF for LLM alignment
  • learn how reinforcement learning with human feedback works
  • find open-source RLHF codebases and datasets
  • track the latest RLHF research in 2024-2026
  • study reward modeling and human preference learning
  • get reading material for aligning language models with human values

When to choose

  • you need a starting point to survey the RLHF literature
  • you want curated links to papers, code, and datasets in one place
  • you are researching LLM alignment or preference-based RL

When to avoid

  • you need runnable RLHF training code rather than a resource list
  • you need a maintained software library with an API
  • you need non-RLHF reinforcement learning resources

Facets

learning-resource · maturity active

reinforcement-learning llm-training machine-learning reinforcement-learning large-language-models machine-learning artificial-intelligence tutorials cross-platform awesome-list rlhf curated-list research-papers human-feedback alignment

1 source

Member repositories

RepositoryRoleHealth v2
opendilab/awesome-RLHFmain68

For agents

markdown · JSON · MCP: product_card(name="opendilab/awesome-RLHF")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem