Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yishay Mansour — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yishay Mansour
20
papers ·
433
citations ·
0
h-index
Google (United States) · Tel Aviv University · Google (Israel)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Online Convex Optimization in Adversarial Markov Decision Processes
2019 · 45 citations
Guarantees for Epsilon-Greedy Reinforcement Learning with Function Approximation
2022 · 26 citations
Near-optimal Regret Bounds for Stochastic Shortest Path
2020 · 18 citations
Unknown mixing times in apprenticeship and reinforcement learning
2019 · 3 citations
Reinforcement Learning with Feedback Graphs
2020 · 2 citations
Planning in Hierarchical Reinforcement Learning: Guarantees for Using Local Policies
2019 · 1 citations
Agnostic Reinforcement Learning with Low-Rank MDPs and Rich Observations
2021 · 1 citations
Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
2022 · 1 citations
There is no Accuracy-Interpretability Tradeoff in Reinforcement Learning for Mazes
2022 · 1 citations
Improved Regret for Efficient Online Reinforcement Learning with Linear Function Approximation
2023 · 1 citations
Near-optimal Regret Using Policy Optimization in Online MDPs with Aggregate Bandit Feedback
2025
Hierarchical Reinforcement Learning: Approximating Optimal Discounted TSP Using Local Policies
2018
Learning Adversarial Markov Decision Processes with Delayed Feedback
2020
Online Markov Decision Processes with Aggregate Bandit Feedback
2021
Cooperative Online Learning in Stochastic and Adversarial MDPs
2022
Top co-authors
Aviv Rosenberg
· 5
Haim Kaplan
· 5
Alon Cohen
· 4
Tal Lancewicki
· 4
Ayush Sekhari
· 3
Christoph Dann
· 3
Karthik Sridharan
· 3
Mehryar Mohri
· 3
Tom Zahavy
· 3
Michal Moshkovitz
· 2
Tomer Koren
· 2
Asaf Cassel
· 1
Topics
Model-Based RL
Exploration
Policy Gradient
Value-Based
Safe RL
Offline RL
Meta-RL
Game AI
Multi-Agent