Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
J\"urgen Schmidhuber — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
J\"urgen Schmidhuber
39
papers ·
414
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving Generalization in Meta Reinforcement Learning using Learned Objectives
2019 · 59 citations
Training Agents using Upside-Down Reinforcement Learning
2019 · 20 citations
Directly Forecasting Belief for Reinforcement Learning with Delays
2025 · 4 citations
General Policy Evaluation and Improvement by Learning to Identify Few But Crucial States
2022 · 3 citations
The Benefits of Model-Based Generalization in Reinforcement Learning
2022 · 3 citations
On the Convergence and Stability of Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning, and Online Decision Transformers
2025 · 1 citations
Upside-Down Reinforcement Learning Can Diverge in Stochastic Environments With Episodic Resets
2022 · 1 citations
Goal-Conditioned Generators of Deep Policies
2022 · 1 citations
Exploring through Random Curiosity with General Value Functions
2022 · 1 citations
Guiding Online Reinforcement Learning with Action-Free Offline Pretraining
2023 · 1 citations
Upside Down Reinforcement Learning with Policy Generators
2025
Reward-Weighted Regression Converges to a Global Optimum
2021
All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL
2022
Learning Relative Return Policies With Upside-Down Reinforcement Learning
2022
Unsupervised Learning of Temporal Abstractions with Slot-based Transformers
2022
Top co-authors
Francesco Faccio
· 9
Dylan R. Ashley
· 7
Louis Kirsch
· 5
Aditya Ramesh
· 4
Rupesh Kumar Srivastava
· 4
Miroslav \v{S}trupl
· 3
Sjoerd van Steenkiste
· 3
Vincent Herrmann
· 3
Chao Huang
· 2
Kai Arulkumaran
· 2
Kenny Young
· 2
Qingyuan Wu
· 2
Topics
Value-Based
Model-Based RL
Offline RL
Meta-RL
Policy Gradient
Exploration
Game AI
Multi-Agent
RLHF & Alignment
Safe RL