Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Quanquan Gu โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Quanquan Gu
60
papers ยท
355
citations ยท
53
h-index
University of California, Los Angeles ยท Zhejiang Shuren University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
A Finite Time Analysis of Two Time-Scale Actor Critic Methods
2020 ยท 37 citations
Sample Efficient Policy Gradient Methods with Recursive Variance Reduction
2019 ยท 34 citations
An Improved Convergence Analysis of Stochastic Variance-Reduced Policy Gradient
2019 ยท 33 citations
Provably Efficient Reinforcement Learning for Discounted MDPs with Feature Mapping
2020 ยท 30 citations
Nearly Minimax Optimal Reinforcement Learning for Linear Mixture Markov Decision Processes
2020 ยท 23 citations
Logarithmic Regret for Reinforcement Learning with Linear Function Approximation
2020 ยท 21 citations
A Finite-Time Analysis of Q-Learning with Neural Network Function Approximation
2019 ยท 18 citations
Nearly Minimax Optimal Reinforcement Learning for Discounted MDPs
2020 ยท 9 citations
Reward-Free Model-Based Reinforcement Learning with Linear Function Approximation
2021 ยท 6 citations
Nearly Minimax Optimal Regret for Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation
2021 ยท 4 citations
Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
2021 ยท 4 citations
Nearly Minimax Optimal Reinforcement Learning for Linear Markov Decision Processes
2022 ยท 4 citations
Almost Optimal Algorithms for Two-player Zero-Sum Linear Mixture Markov Games
2021 ยท 3 citations
Computationally Efficient Horizon-Free Reinforcement Learning for Linear Mixture MDPs
2022 ยท 3 citations
Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization
2024 ยท 2 citations
Top co-authors
Dongruo Zhou
ยท 17
Jiafan He
ยท 11
Heyang Zhao
ยท 7
Jiafan He
ยท 7
Weitong Zhang
ยท 6
Chenlu Ye
ยท 5
Pan Xu
ยท 4
Tianhao Wang
ยท 3
Yifei Min
ยท 3
Felicia Gao
ยท 2
Junkai Zhang
ยท 2
Kaixuan Ji
ยท 2
Topics
Model-Based RL
Exploration
Value-Based
Offline RL
Policy Gradient
RLHF & Alignment
Multi-Agent
Safe RL
Game AI