Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Huaxiu Yao — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Huaxiu Yao
43
papers ·
1170
citations ·
25
h-index
University of North Carolina at Chapel Hill
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
TempoVLA: Learning Speed-Controllable Vision-Language-Action Policies
2026
Provable and Practical In-Context Policy Optimization for Self-Improvement
2026
MetaClaw: Just Talk -- An Agent That Meta-Learns and Evolves in the Wild
2026
SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
2026
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
2026
Scaling Agent Learning via Experience Synthesis
2025
Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning
2025
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning
2025
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards
2025
MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning
2025
Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors
2025
It Takes Two: On the Seamlessness between Reward and Policy Model in RLHF
2024
Top co-authors
Peng Xia
· 6
Jiaqi Liu
· 5
Siwei Han
· 4
Kaide Zeng
· 3
Yiyang Zhou
· 3
Cihang Xie
· 2
Fang Wu
· 2
Haonian Ji
· 2
Jianwen Chen
· 2
Kaiwen Xiong
· 2
Mingyu Ding
· 2
Xiangru Tang
· 2
Topics
cs.AI
cs.LG
Model-Based RL
Multi-Agent
RLHF & Alignment
cs.RO
cs.CL
cs.CV
Value-Based
Meta-RL