Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Haipeng Luo — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Haipeng Luo
20
papers ·
74
citations ·
23
h-index
University of Southern California · California Southern University · International Union of Railways
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Model-free Reinforcement Learning in Infinite-horizon Average-reward Markov Decision Processes
2019 · 19 citations
Non-stationary Reinforcement Learning without Prior Knowledge: An Optimal Black-box Approach
2021 · 17 citations
Model selection for contextual bandits
2019 · 10 citations
Online Learning for Stochastic Shortest Path Model via Posterior Sampling
2021 · 6 citations
Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation
2020 · 5 citations
A Model-free Learning Algorithm for Infinite-horizon Average-reward MDPs with Near-optimal Regret
2020 · 3 citations
Learning Infinite-Horizon Average-Reward Markov Decision Processes with Constraints
2022 · 3 citations
Policy Optimization in Adversarial MDPs: Improved Exploration via Dilated Bonuses
2021 · 2 citations
Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
2022 · 1 citations
No-Regret Online Reinforcement Learning with Adversarial Losses and Transitions
2023 · 1 citations
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena
2024 · 1 citations
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
2025
Reinforcement Learning from Adversarial Preferences in Tabular MDPs
2025
Policy Optimization for Stochastic Shortest Path
2022
Near-Optimal Goal-Oriented Reinforcement Learning in Non-Stationary Environments
2022
Top co-authors
Chen-Yu Wei
· 7
Mehdi Jafarnia-Jahromi
· 4
Rahul Jain
· 4
Aviv Rosenberg
· 3
Liyu Chen
· 2
Rahul Jain
· 2
Tiancheng Jin
· 2
Akhil Agnihotri
· 1
Akshay Krishnamurthy
· 1
Asaf Cassel
· 1
Can Xu
· 1
Can Xu
· 1
Topics
Model-Based RL
Exploration
Policy Gradient
Value-Based
Multi-Agent
Safe RL
Game AI
RLHF & Alignment
math.OC
math.ST