Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Dipendra Misra — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Dipendra Misra
18
papers ·
113
citations ·
17
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Lipschitz Continuity in Model-based Reinforcement Learning
2018 · 37 citations
Combating the Compounding-Error Problem with a Multi-step Model
2019 · 27 citations
Kinematic State Abstraction and Provably Efficient Rich-Observation Reinforcement Learning
2019 · 13 citations
Equivalence Between Wasserstein and Value-Aware Loss for Model-based Reinforcement Learning
2018 · 9 citations
Towards a Simple Approach to Multi-step Model-based Reinforcement Learning
2018 · 6 citations
Learning to Generate Better Than Your LLM
2023 · 4 citations
Interactive Learning from Activity Description
2021 · 3 citations
Towards Principled Representation Learning From Videos For Reinforcement Learning
2024 · 1 citations
Provable RL with Exogenous Distractors via Multistep Inverse Dynamics
2021 · 1 citations
Provable Safe Reinforcement Learning with Binary Feedback
2022 · 1 citations
Survival Instinct in Offline Reinforcement Learning
2023 · 1 citations
Policy Improvement using Language Feedback Models
2024 · 1 citations
Dataset Reset Policy Optimization for RLHF
2024 · 1 citations
Provably Sample-Efficient RL with Side Information about Latent Dynamics
2022
Sample-Efficient Reinforcement Learning in the Presence of Exogenous Information
2022
Top co-authors
John Langford
· 6
Kavosh Asadi
· 4
Akshay Krishnamurthy
· 3
Alex Lamb
· 3
Michael L. Littman
· 3
Yonathan Efroni
· 3
Evan Cater
· 2
Jonathan D. Chang
· 2
Miro Dud\'ik
· 2
Akanksha Saran
· 1
Alekh Agarwal
· 1
Andrew Bennett
· 1
Topics
Model-Based RL
Exploration
Meta-RL
Offline RL
RLHF & Alignment
Value-Based
Multi-Agent
Safe RL
Policy Gradient
cs.CV