Awesome Reinforcement Learning
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Hao Zhang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Hao Zhang
287
papers ·
500
citations ·
73
h-index
Tianjin University of Technology · Simon Fraser University · Shanghai University of Electric Power · China Tobacco · Michigan State University · University of Edinburgh
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Arctic-Text2SQL-R1: Simple Rewards, Strong Reasoning in Text-to-SQL
2025 · 2 citations
HALO: Learning Human-Robot Collaboration via Heterogeneous-Agent Lyapunov Policy Optimization
2026
Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections
2026
Towards Interpretable and Inference-Optimal COT Reasoning with Sparse Autoencoder-Guided Generation
2025
Automated Vehicles Should be Connected with Natural Language
2025
Directly Learning Stock Trading Strategies Through Profit Guided Loss Functions
2025
Kimi K2: Open Agentic Intelligence
2025
ReasonMed: A 370K Multi-Agent Generated Dataset for Advancing Medical Reasoning
2025
lmgame-Bench: How Good are LLMs at Playing Games?
2025
Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition
2025
Technical Note: Game-Theoretic Interactions of Different Orders
2020
Top co-authors
Ion Stoica
· 4
Lanxiang Hu
· 3
Tajana Rosing
· 3
Eric P. Xing
· 2
Haojian Jin
· 2
Yonghao Zhuang
· 2
Abhilash Shankarampeta
· 1
Adam Mahdi
· 1
Alex Ororbia
· 1
and Zhen Kan
· 1
Angang Du
· 1
Ang Li
· 1
Topics
cs.AI
cs.CL
cs.LG
cs.MA
Multi-Agent
Policy Gradient
Safe RL
cs.CV
cs.RO
cs.CE