Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bogdan Mazoure — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Bogdan Mazoure
21
papers ·
72
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Reinforcement and InfoMax Learning
2020 · 16 citations
Leveraging exploration in off-policy algorithms via normalizing flows
2019 · 12 citations
GAN Q-learning
2018 · 8 citations
Attraction-Repulsion Actor-Critic for Continuous Control Reinforcement Learning
2019 · 8 citations
Cross-Trajectory Representation Learning for Zero-Shot Generalization in RL
2021 · 7 citations
Large Language Models as Generalizable Policies for Embodied Tasks
2023 · 7 citations
Improving Long-Term Metrics in Recommendation Systems using Short-Horizon Reinforcement Learning
2021 · 2 citations
Efficient Planning under Partial Observability with Unnormalized Q Functions and Spectral Learning
2019 · 1 citations
Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces
2020 · 1 citations
Improving Zero-shot Generalization in Offline Reinforcement Learning using Generalized Similarity Functions
2021 · 1 citations
GRACE: A Language Model Framework for Explainable Inverse Reinforcement Learning
2025
Contrastive Value Learning: Implicit Models for Simple Offline RL
2022
Accelerating exploration and representation learning with offline pre-training
2023
Value function estimation using conditional diffusion models for control
2023
On the benefits of pixel-based hierarchical policies for task generalization
2024
Top co-authors
Thang Doan
· 5
Alexander Toshev
· 4
Devon Hjelm
· 4
Doina Precup
· 4
Joelle Pineau
· 3
R Devon Hjelm
· 3
Walter Talbott
· 3
Audrey Durand
· 2
Guillaume Rabusseau
· 2
Jonathan Tompson
· 2
Josh Susskind
· 2
Ofir Nachum
· 2
Topics
Model-Based RL
Value-Based
Exploration
Policy Gradient
Meta-RL
Offline RL
RLHF & Alignment
Multi-Agent
Game AI
Safe RL