Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Lihong Li โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Lihong Li
36
papers ยท
1295
citations ยท
33
h-index
Hebei University of Engineering ยท North China University of Science and Technology ยท TED University ยท Hangzhou Dianzi University ยท Shanxi Datong University ยท Shenyang Jianzhu University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Composite Task-Completion Dialogue Policy Learning via Hierarchical Deep Reinforcement Learning
2017 ยท 153 citations
A User Simulator for Task-Completion Dialogues
2016 ยท 140 citations
SBEED: Convergent Reinforcement Learning with Nonlinear Function Approximation
2017 ยท 120 citations
Breaking the Curse of Horizon: Infinite-Horizon Off-Policy Estimation
2018 ยท 112 citations
BBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systems
2016 ยท 98 citations
DualDICE: Behavior-Agnostic Estimation of Discounted Stationary Distribution Corrections
2019 ยท 90 citations
AlgaeDICE: Policy Gradient from Arbitrary Experience
2019 ยท 82 citations
Stochastic Variance Reduction Methods for Policy Evaluation
2017 ยท 69 citations
Combating Reinforcement Learning's Sisyphean Curse with Intrinsic Fear
2016 ยท 50 citations
GenDICE: Generalized Offline Estimation of Stationary Values
2020 ยท 50 citations
Subgoal Discovery for Hierarchical Dialogue Policy Learning
2018 ยท 39 citations
Towards End-to-End Reinforcement Learning of Dialogue Agents for Information Access
2016 ยท 36 citations
Understanding Domain Randomization for Sim-to-real Transfer
2021 ยท 33 citations
A Kernel Loss for Solving the Bellman Equation
2019 ยท 26 citations
Scalable Bilinear $ฯ$ Learning Using State and Action Features
2018 ยท 22 citations
Top co-authors
Jianfeng Gao
ยท 7
Bo Dai
ยท 6
Xiujun Li
ยท 5
Bing Yin
ยท 4
Dale Schuurmans
ยท 4
Changlong Yu
ยท 3
Hyokun Yun
ยท 3
Jianshu Chen
ยท 3
Ofir Nachum
ยท 3
Qiang Liu
ยท 3
Yinlam Chow
ยท 3
Zachary C. Lipton
ยท 3
Topics
Value-Based
Offline RL
Model-Based RL
Policy Gradient
cs.LG
Exploration
cs.CL
Safe RL
cs.AI
Multi-Agent