Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Jason D. Lee โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Jason D. Lee
36
papers ยท
330
citations ยท
41
h-index
Princeton University ยท Services Australia ยท University of California, Berkeley
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
On the Theory of Policy Gradient Methods: Optimality, Approximation, and Distribution Shift
2019 ยท 111 citations
Agnostic Q-learning with Function Approximation in Deterministic Systems: Tight Bounds on Approximation Error and Sample Complexity
2020 ยท 23 citations
Bilinear Classes: A Structural Framework for Provable Generalization in RL
2021 ยท 21 citations
Neural Temporal-Difference and Q-Learning Provably Converge to Global Optima
2019 ยท 15 citations
Policy Mirror Descent for Regularized Reinforcement Learning: A Generalized Framework with Linear Convergence
2021 ยท 8 citations
MUSBO: Model-based Uncertainty Regularized and Sample Efficient Batch Optimization for Deployment Constrained Reinforcement Learning
2021 ยท 6 citations
Offline Reinforcement Learning with Realizability and Single-policy Concentrability
2022 ยท 6 citations
Provably Efficient Reinforcement Learning in Partially Observable Dynamical Systems
2022 ยท 3 citations
Reward-agnostic Fine-tuning: Provable Statistical Benefits of Hybrid Reinforcement Learning
2023 ยท 3 citations
Provably Efficient Policy Optimization for Two-Player Zero-Sum Markov Games
2021 ยท 2 citations
Local Optimization Achieves Global Optimality in Multi-Agent Reinforcement Learning
2023 ยท 2 citations
REBEL: Reinforcement Learning via Regressing Relative Rewards
2024 ยท 2 citations
PAC Reinforcement Learning for Predictive State Representations
2022 ยท 1 citations
Dataset Reset Policy Optimization for RLHF
2024 ยท 1 citations
Accelerating RL for LLM Reasoning with Optimal Advantage Regression
2025
Top co-authors
Wenhao Zhan
ยท 10
Wen Sun
ยท 6
Masatoshi Uehara
ยท 5
Simon S. Du
ยท 5
Kiant\'e Brantley
ยท 4
Zhuoran Yang
ยท 4
Baihe Huang
ยท 3
Gaurav Mahajan
ยท 3
Jonathan D. Chang
ยท 3
Nathan Kallus
ยท 3
Ruosong Wang
ยท 3
Sham M. Kakade
ยท 3
Topics
Model-Based RL
Policy Gradient
Value-Based
Offline RL
Exploration
RLHF & Alignment
Multi-Agent
Game AI
Safe RL
math.OC