Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tengyang Xie — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Tengyang Xie
20
papers ·
184
citations ·
10
h-index
University of Wisconsin–Madison
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Towards Optimal Off-Policy Evaluation for Reinforcement Learning with Marginalized Importance Sampling
2019 · 53 citations
Batch Value-function Approximation with Only Realizability
2020 · 30 citations
Q* Approximation Schemes for Batch Reinforcement Learning: A Theoretical Comparison
2020 · 16 citations
Policy Finetuning: Bridging Sample-Efficient Offline and Online Reinforcement Learning
2021 · 16 citations
Bellman-consistent Pessimism for Offline Reinforcement Learning
2021 · 16 citations
Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency
2021 · 14 citations
Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
2024 · 10 citations
A Variant of the Wang-Foster-Kakade Lower Bound for the Discounted Setting
2020 · 9 citations
Adversarially Trained Actor Critic for Offline Reinforcement Learning
2022 · 7 citations
Adversarial Model for Offline Reinforcement Learning
2023 · 4 citations
Privacy Preserving Off-Policy Evaluation
2019 · 2 citations
Interaction-Grounded Learning with Action-inclusive Feedback
2022 · 2 citations
Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
2024 · 2 citations
Towards Principled Representation Learning From Videos For Reinforcement Learning
2024 · 1 citations
Interaction-Grounded Learning
2021 · 1 citations
Top co-authors
Nan Jiang
· 6
Ching-An Cheng
· 5
Nan Jiang
· 4
Dylan J. Foster
· 3
John Langford
· 3
Paul Mineiro
· 3
Akanksha Saran
· 2
Alekh Agarwal
· 2
Ida Momennejad
· 2
Mohak Bhardwaj
· 2
Nan Jiang
· 2
Philip Amortila
· 2
Topics
Offline RL
Model-Based RL
Value-Based
Exploration
Safe RL
RLHF & Alignment
Meta-RL
Policy Gradient
Game AI
cs.CV