Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Wenhao Zhan — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Wenhao Zhan
15
papers ·
21
citations ·
20
h-index
Ningbo University · Universitat Autònoma de Barcelona · Sun Yat-sen University · Vall d'Hebron Institut de Recerca · The First Affiliated Hospital, Sun Yat-sen University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Policy Mirror Descent for Regularized Reinforcement Learning: A Generalized Framework with Linear Convergence
2021 · 8 citations
Offline Reinforcement Learning with Realizability and Single-policy Concentrability
2022 · 6 citations
Reward-agnostic Fine-tuning: Provable Statistical Benefits of Hybrid Reinforcement Learning
2023 · 3 citations
REBEL: Reinforcement Learning via Regressing Relative Rewards
2024 · 2 citations
PAC Reinforcement Learning for Predictive State Representations
2022 · 1 citations
Dataset Reset Policy Optimization for RLHF
2024 · 1 citations
Accelerating RL for LLM Reasoning with Optimal Advantage Regression
2025
Decentralized Optimistic Hyperpolicy Mirror Descent: Provably No-Regret Learning in Markov Games
2022
Provable Offline Preference-Based Reinforcement Learning
2023
Provable Reward-Agnostic Preference-Based Reinforcement Learning
2023
Provably Efficient CVaR RL in Low-rank MDPs
2023
Exploiting Structure in Offline Multi-Agent RL: The Benefits of Low Interaction Rank
2024
Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF
2024
Top co-authors
Jason D. Lee
· 13
Wen Sun
· 8
Kiant\'e Brantley
· 4
Jonathan D. Chang
· 3
Masatoshi Uehara
· 3
Zhaolin Gao
· 3
Baihe Huang
· 2
Gokul Swamy
· 2
Yuejie Chi
· 2
Yuxin Chen
· 2
Dipendra Misra
· 1
Farzan Farnia
· 1
Topics
Offline RL
Policy Gradient
Model-Based RL
RLHF & Alignment
Exploration
Multi-Agent
Safe RL
Game AI
Value-Based