atlas

Research map / Reinforcement learning

RL from feedback: research map

737 accepted papers on RL from feedback in Reinforcement learning, from ICML, NeurIPS, ICLR, CVPR and AAAI (2016–2026), grouped into 2 clusters and 7 approaches. The busiest year so far is 2026.

Within Reinforcement learning, its share grew from 5.8% in 2023–24 to 26.8% in 2025–26 (91 → 620 papers at ICML, NeurIPS, CVPR and AAAI, the venues with data for all four years).

2016: 1162017: 3172018: 0182019: 3192020: 7202021: 5212022: 7222023: 17232024: 74242025: 149252026: 47126

Explore RL from feedback in the interactive map

Working on something in this topic? Describe your idea in scime atlas to see which approach it falls under, the closest papers by meaning and how crowded the spot has become.

Approaches and key papers

reasoning · grpo · llm · 506 papers

Approaches in this cluster:

Most cited and most cited since 2024:

human · rlhf · preferences · 231 papers

Approaches in this cluster:

Most cited and most cited since 2024:

Related topics in Reinforcement learning