Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yarin Gal — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yarin Gal
34
papers ·
690
citations ·
55
h-index
Apple (Israel) · Apple (United Kingdom) · Science Oxford · Apple (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Concrete Dropout
2017 · 145 citations
Learning Invariant Representations for Reinforcement Learning without Reconstruction
2020 · 97 citations
VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning
2019 · 65 citations
Invariant Causal Prediction for Block MDPs
2020 · 37 citations
Generalizing from a few environments in safety-critical reinforcement learning
2019 · 9 citations
PsiPhi-Learning: Reinforcement Learning with Demonstrations using Successor Features and Inverse Temporal Difference Learning
2021 · 5 citations
Adversarial recovery of agent rewards from latent spaces of the limit order book
2019 · 3 citations
Outcome-Driven Reinforcement Learning via Variational Inference
2021 · 3 citations
On Pathologies in KL-Regularized Reinforcement Learning from Expert Demonstrations
2022 · 2 citations
ReLU to the Rescue: Improve Your On-Policy Actor-Critic with Positive Advantages
2023 · 2 citations
Learning Dynamics and Generalization in Reinforcement Learning
2022 · 1 citations
Can Active Sampling Reduce Causal Confusion in Offline Reinforcement Learning?
2023 · 1 citations
Iterative Deployment Improves Planning Skills in LLMs
2025
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
2025
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
2025
Top co-authors
Angelos Filos
· 4
Clare Lyle
· 3
Sergey Levine
· 3
Tim G. J. Rudner
· 3
Amy Zhang
· 2
Gunshi Gupta
· 2
Luckeciano C. Melo
· 2
Marta Kwiatkowska
· 2
Rowan McAllister
· 2
Adrien Gaidon
· 1
Alessandro Abate
· 1
Alex Kendall
· 1
Topics
Exploration
Model-Based RL
Safe RL
Meta-RL
Offline RL
Value-Based
Policy Gradient
Multi-Agent
RLHF & Alignment
cs.LG