Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Markus Wulfmeier — most-cited papers & profile · Large Language Models
← authors
·
overview
Markus Wulfmeier
42
papers ·
349
citations ·
16
h-index
Google DeepMind (United Kingdom) · Google (United Kingdom) · Goldsmiths University of London
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Reverse Curriculum Generation for Reinforcement Learning
2017 · 140 citations
Compositional Transfer in Hierarchical Reinforcement Learning
2019 · 27 citations
Continuous-Discrete Reinforcement Learning for Hybrid Control in Robotics
2020 · 27 citations
Imitate and Repurpose: Learning Reusable Robot Movement Skills From Human and Animal Behaviors
2022 · 20 citations
Mutual Alignment Transfer Learning
2017 · 19 citations
Is Bang-Bang Control All You Need? Solving Continuous Control with Bernoulli Policies
2021 · 15 citations
TACO: Learning Task Decomposition Via Temporal Alignment For Control
2018 · 12 citations
Towards General and Autonomous Learning of Core Skills: A Case Study in Locomotion
2020 · 10 citations
Towards General and Autonomous Learning of Core Skills: A Case Study in Locomotion
2020 · 10 citations
The Challenges of Exploration for Offline Reinforcement Learning
2022 · 8 citations
Attention-Privileged Reinforcement Learning
2019 · 6 citations
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
2023 · 6 citations
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
2023 · 6 citations
Incorporating Human Domain Knowledge into Large Scale Cost Function Learning
2016 · 3 citations
Simple Sensor Intentions for Exploration
2020 · 3 citations
Top co-authors
Jordi Grau-Moya
· 1
Jörg Bornschein
· 1
Razvan Pascanu
· 1
Thomas Schmied
· 1
Topics
Exploration
Control
Value-Based
Meta-RL
Sim-to-Real
Offline RL
Model-Based RL
Manipulation
Policy Gradient
Multi-Agent