Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yoshua Bengio — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yoshua Bengio
114
papers ·
12127
citations ·
186
h-index
Centre Universitaire de Mila · Université de Montréal
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
An Actor-Critic Algorithm for Sequence Prediction
2016 · 224 citations
A Deep Reinforcement Learning Chatbot
2017 · 200 citations
Combined Reinforcement Learning via Abstract Representations
2018 · 97 citations
Revisiting Fundamentals of Experience Replay
2020 · 83 citations
Hyperbolic Discounting and Learning over Multiple Horizons
2019 · 55 citations
InfoBot: Transfer and Exploration via the Information Bottleneck
2019 · 47 citations
Learning To Navigate The Synthetically Accessible Chemical Space Using Reinforcement Learning
2020 · 43 citations
Recall Traces: Backtracking Models for Efficient Reinforcement Learning
2018 · 25 citations
Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training
2025 · 24 citations
Reinforcement Learning with Competitive Ensembles of Information-Constrained Primitives
2019 · 23 citations
Universal Successor Representations for Transfer Reinforcement Learning
2018 · 19 citations
A Consciousness-Inspired Planning Agent for Model-Based Reinforcement Learning
2021 · 18 citations
Learning Dynamics Model in Reinforcement Learning by Incorporating the Long Term Future
2019 · 17 citations
A Deep Reinforcement Learning Chatbot (Short Version)
2018 · 14 citations
The effects of negative adaptation in Model-Agnostic Meta-Learning
2018 · 14 citations
Top co-authors
Anirudh Goyal
· 13
Moksh Jain
· 7
Aaron Courville
· 5
Dinghuai Zhang
· 5
Doina Precup
· 5
Hugo Larochelle
· 5
Minsu Kim
· 5
Sergey Levine
· 5
Siddarth Venkatraman
· 5
Emmanuel Bengio
· 4
Glen Berseth
· 4
Ling Pan
· 4
Topics
Model-Based RL
Exploration
Meta-RL
Value-Based
Policy Gradient
Offline RL
Safe RL
Multi-Agent
RLHF & Alignment
cs.LG