Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Bo Dai โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Bo Dai
25
papers ยท
633
citations ยท
29
h-index
Southwest University of Science and Technology ยท China Mobile (China) ยท Chang'an University ยท Fudan University Shanghai Cancer Center ยท State Key Laboratory of Mobile Networks and Mobile Multimedia Technology
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
SBEED: Convergent Reinforcement Learning with Nonlinear Function Approximation
2017 ยท 120 citations
DualDICE: Behavior-Agnostic Estimation of Discounted Stationary Distribution Corrections
2019 ยท 90 citations
GenDICE: Generalized Offline Estimation of Stationary Values
2020 ยท 50 citations
Reinforcement Learning via Fenchel-Rockafellar Duality
2020 ยท 25 citations
Learning from Conditional Distributions via Dual Embeddings
2016 ยท 20 citations
Learning Sparse Rewarded Tasks from Sub-Optimal Demonstrations
2020 ยท 11 citations
Offline Policy Selection under Uncertainty
2020 ยท 10 citations
Neural Stochastic Dual Dynamic Programming
2021 ยท 9 citations
Off-Policy Imitation Learning from Observations
2021 ยท 7 citations
Zeroth-Order Supervised Policy Improvement
2020 ยท 2 citations
SAFER: Data-Efficient and Safe Reinforcement Learning via Skill Acquisition
2022 ยท 2 citations
The Curse Of Passive Data Collection In Batch Reinforcement Learning
2021 ยท 1 citations
Spectral Decomposition Representation for Reinforcement Learning
2022 ยท 1 citations
Latent Variable Representation for Reinforcement Learning
2022 ยท 1 citations
Rethinking the Global Convergence of Softmax Policy Gradient with Linear Function Approximation
2025
Top co-authors
Dale Schuurmans
ยท 8
Chenjun Xiao
ยท 3
Ofir Nachum
ยท 3
Tongzheng Ren
ยท 3
Csaba Szepesvari
ยท 2
Kaixiang Lin
ยท 2
Niao He
ยท 2
Tianjun Zhang
ยท 2
Yinlam Chow
ยท 2
Zhuangdi Zhu
ยท 2
Albert Shaw
ยท 1
Alekh Agarwal
ยท 1
Topics
Offline RL
Model-Based RL
Value-Based
Exploration
Policy Gradient
Meta-RL
Safe RL