Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zihan Zhang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Zihan Zhang
51
papers ·
140
citations ·
5
h-index
South Ural State University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Almost Optimal Model-Free Reinforcement Learning via Reference-Advantage Decomposition
2020 · 45 citations
Regret Minimization for Reinforcement Learning by Evaluating the Optimal Bias Function
2019 · 23 citations
Model-Free Reinforcement Learning: from Clipped Pseudo-Regret to Sample Complexity
2020 · 13 citations
Nearly Minimax Optimal Reward-free Reinforcement Learning
2020 · 11 citations
Is Reinforcement Learning More Difficult Than Bandits? A Near-optimal Algorithm Escaping the Curse of Horizon
2020 · 4 citations
Horizon-Free Reinforcement Learning in Polynomial Time: the Power of Stationary Policies
2022 · 2 citations
Sharper Model-free Reinforcement Learning for Average-reward Markov Decision Processes
2023 · 2 citations
Near-Optimal Regret Bounds for Multi-batch Reinforcement Learning
2022 · 1 citations
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
2024 · 1 citations
Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence
2026
Tighter Regret Bounds for Contextual Action-Set Reinforcement Learning
2026
Frozen Policy Iteration: Computationally Efficient RL under Linear $Q^π$ Realizability for Deterministic Dynamics
2026
Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO
2025
Sharp Variance-Dependent Bounds in Reinforcement Learning: Best of Both Worlds in Stochastic and Deterministic Environments
2023
Settling the Sample Complexity of Online Reinforcement Learning
2023
Top co-authors
Simon S. Du
· 8
Xiangyang Ji
· 7
Runlong Zhou
· 3
Yuan Zhou
· 3
Jason D. Lee
· 2
Maryam Fazel
· 2
Yuxin Chen
· 2
Yuhang Jiang
· 1
Zijun Chen
· 1
Topics
Model-Based RL
Value-Based
Exploration
cs.LG
stat.ML
RLHF & Alignment
Policy Gradient
Offline RL
Multi-Agent
Safe RL