Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Marcus Hutter — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Marcus Hutter
18
papers ·
62
citations ·
33
h-index
Australian National University · Google DeepMind (United Kingdom) · Google (United Kingdom)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Avoiding Wireheading with Value Reinforcement Learning
2016 · 28 citations
Thompson Sampling is Asymptotically Optimal in General Environments
2016 · 18 citations
Counterfactual Credit Assignment in Model-Free Reinforcement Learning
2020 · 8 citations
Pessimism About Unknown Unknowns Inspires Conservatism
2020 · 4 citations
Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
2019 · 3 citations
Distributional Bellman Operators over Mean Embeddings
2023 · 1 citations
Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning
2025
Death and Suicide in Universal Artificial Intelligence
2016
Generalised Discount Functions applied to a Monte-Carlo AImu Implementation
2017
Curiosity Killed or Incapacitated the Cat and the Asymptotically Optimal Agent
2020
Exact Reduction of Huge Action Spaces in General Reinforcement Learning
2020
Reinforcement Learning with Information-Theoretic Actuation
2021
Reducing Planning Complexity of General Reinforcement Learning with Non-Markovian Abstractions
2021
Atari-5: Distilling the Arcade Learning Environment down to Five Games
2022
Universal Agent Mixtures and the Geometry of Intelligence
2023
Top co-authors
Michael K. Cohen
· 3
Tom Everitt
· 3
Elliot Catt
· 2
Jan Leike
· 2
Matthew Aitchison
· 2
Sultan Javed Majeed
· 2
Alaa Saade
· 1
Alexander Meulemans
· 1
Angelika Steger
· 1
Anian Ruoss
· 1
Anna Harutyunyan
· 1
Arthur Gretton
· 1
Topics
Value-Based
Model-Based RL
Safe RL
Game AI
Multi-Agent
Exploration
RLHF & Alignment
Meta-RL
Policy Gradient