Awesome Reinforcement Learning
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Yaodong Yang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yaodong Yang
117
papers ·
100
citations ·
32
h-index
Peking University · Beijing Academy of Artificial Intelligence · Beijing Haidian Hospital · Shaanxi University of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Mitigating Reward Over-Optimization in RLHF via Behavior-Supported Regularization
2025 · 10 citations
Attacking Cooperative Multi-Agent Reinforcement Learning by Adversarial Minority Influence
2023 · 3 citations
Factorized Q-Learning for Large-Scale Multi-Agent Systems
2018 · 1 citations
Policy Improvement Reinforcement Learning
2026
Heterogeneous Agent Collaborative Reinforcement Learning
2026
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
2026
Constrained Language Model Policy Optimization via Risk-aware Stepwise Alignment
2025
Social World Model-Augmented Mechanism Design Policy Learning
2025
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies
2025
Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
2025
J1: Exploring Simple Test-Time Scaling for LLM-as-a-Judge
2025
When Slower Isn't Truer: Inverse Scaling Law of Truthfulness in Multimodal Reasoning
2025
Risk-aware Direct Preference Optimization under Nested Risk Measure
2025
Differentiable Information Enhanced Model-Based Reinforcement Learning
2025
Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
2025
Top co-authors
Jun Wang
· 10
Jiaming Ji
· 6
Song-Chun Zhu
· 4
Boyuan Chen
· 3
Chengdong Ma
· 3
Chi-Min Chan
· 3
Jiahao Li
· 3
Jiayi Zhou
· 3
Juntao Dai
· 3
Kaile Wang
· 3
Le Cong Dinh
· 3
Sirui Han
· 3
Topics
cs.AI
cs.LG
cs.GT
cs.CL
cs.MA
Multi-Agent
Value-Based
cs.CY
cs.RO
Game AI