#79 · Primary category: AI Tool Directories & Curated Lists

awesome-RLHF

deep-learning deep-reinforcement-learning human-feedback large-language-models reinforcement-learning rlhf

A curated list of reinforcement learning with human feedback resources (continually updated)

Project last updated:05/20/26

GitHub Stars

4.4K

Forks

259

Contributors

37

License

Apache-2.0

Why we included this project

For anyone trying to understand how language models get aligned to human preferences, this curated index of RLHF research is a strong starting point. It organizes papers by year, from the early foundational work through 2026, so you can follow how methods like PPO and DPO developed without hunting through conference proceedings. The maintainers also collect codebases, datasets, blogs, and books, which helps when you want to move from reading about a technique to actually running it. The overview section explains the core idea of RLHF in plain language, so even someone new to reinforcement learning can get oriented. Researchers, grad students, and engineers scoping out the alignment space will find it saves real time.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category