Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Gao Huang โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Gao Huang
94
papers ยท
1409
citations ยท
62
h-index
Wuhan College ยท Wuhan Research Institute of Materials Protection ยท 81th Hospital of PLA ยท Tsinghua University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory
2023 ยท 24 citations
Believe What You See: Implicit Constraint Approach for Offline Multi-Agent Reinforcement Learning
2021 ยท 18 citations
Regularized Anderson Acceleration for Off-Policy Deep Reinforcement Learning
2019 ยท 12 citations
Hundreds Guide Millions: Adaptive Offline Reinforcement Learning with Expert Guidance
2023 ยท 9 citations
STORM: Efficient Stochastic Transformer based World Models for Reinforcement Learning
2023 ยท 8 citations
On the Power of Multitask Representation Learning in Linear MDP
2021 ยท 5 citations
Boosting Offline Reinforcement Learning via Data Rebalancing
2022 ยท 3 citations
Decoupled Prioritized Resampling for Offline RL
2023 ยท 3 citations
A Mixture of Surprises for Unsupervised Reinforcement Learning
2022 ยท 2 citations
Train Once, Get a Family: State-Adaptive Balances for Offline-to-Online Reinforcement Learning
2023 ยท 2 citations
Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
2025 ยท 1 citations
Absolute Zero: Reinforced Self-play Reasoning with Zero Data
2025 ยท 1 citations
Value-Consistent Representation Learning for Data-Efficient Reinforcement Learning
2022 ยท 1 citations
Augmenting Unsupervised Reinforcement Learning with Self-Reference
2023 ยท 1 citations
DyMoDreamer: World Modeling with Dynamic Modulation
2025
Top co-authors
Shenzhi Wang
ยท 5
Shiji Song
ยท 5
Yang Yue
ยท 5
Andrew Zhao
ยท 3
Bingyi Kang
ยท 3
Matthieu Gaetan Lin
ยท 3
Qisen Yang
ยท 3
Rui Lu
ยท 3
Shuicheng Yan
ยท 3
Jian Sun
ยท 2
Matthieu Lin
ยท 2
Shiji Song
ยท 2
Topics
Model-Based RL
Offline RL
Value-Based
Exploration
Game AI
Meta-RL
RLHF & Alignment
Policy Gradient
Safe RL
Multi-Agent