Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Andrea Zanette — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Andrea Zanette
21
papers ·
166
citations ·
11
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Tighter Problem-Dependent Regret Bounds in Reinforcement Learning without Domain Knowledge using Value Function Bounds
2019 · 66 citations
Learning Near Optimal Policies with Low Inherent Bellman Error
2020 · 38 citations
Exponential Lower Bounds for Batch Reinforcement Learning: Batch RL can be Exponentially Harder than Online RL
2020 · 18 citations
Frequentist Regret Bounds for Randomized Least-Squares Value Iteration
2019 · 17 citations
Provable Benefits of Actor-Critic Methods for Offline Reinforcement Learning
2021 · 14 citations
Cautiously Optimistic Policy Optimization and Exploration with Linear Function Approximation
2021 · 7 citations
Problem Dependent Reinforcement Learning Bounds Which Can Identify Bandit Structure in MDPs
2019 · 4 citations
ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
2024 · 2 citations
Expanding the Capabilities of Reinforcement Learning via Text Feedback
2026
Maximum Likelihood Reinforcement Learning
2026
Shrinking the Variance: Shrinkage Baselines for Reinforcement Learning with Verifiable Rewards
2025
SPEED-RL: Faster Training of Reasoning Models via Online Curriculum Learning
2025
Can Large Reasoning Models Self-Train?
2025
Bellman Residual Orthogonalization for Offline Reinforcement Learning
2022
Stabilizing Q-learning with Linear Architectures for Provably Efficient Learning
2022
Top co-authors
Emma Brunskill
· 5
Fahim Tajwar
· 3
Martin J. Wainwright
· 3
Alessandro Lazaric
· 2
Daman Arora
· 2
Ruiqi Zhang
· 2
Yuda Song
· 2
Aarti Singh
· 1
Alekh Agarwal
· 1
Aviral Kumar
· 1
Ching-An Cheng
· 1
Daman Arora
· 1
Topics
Value-Based
Policy Gradient
Exploration
Model-Based RL
Offline RL
RLHF & Alignment
Meta-RL
Multi-Agent