Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhanke Zhou — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Zhanke Zhou
17
papers ·
127
citations ·
5
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
2025 · 1 citations
The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning
2026
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
2026
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
2026
AlphaApollo: A System for Deep Agentic Reasoning
2025
Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models
2025
Top co-authors
Bo Han
· 6
Jiangchao Yao
· 4
Tongliang Liu
· 3
Xiao Feng
· 3
Xuan Li
· 3
Brando Miranda
· 2
Sanmi Koyejo
· 2
Xiangyu Lu
· 2
Jianing Zhu
· 1
Linrui Xu
· 1
Lu Zhang
· 1
Masashi Sugiyama
· 1
Topics
cs.LG
cs.AI
Model-Based RL
cs.CL
Multi-Agent
Game AI
RLHF & Alignment
Value-Based