Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuxiao Qu — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yuxiao Qu
12
papers ·
304
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning
2025 · 121 citations
Recursive Introspection: Teaching Language Model Agents How to Self-Improve
2024 · 8 citations
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
2023 · 4 citations
Introspective X Training: Feedback Conditioning Improves Scaling Across all LLM Training Stages
2026
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
2026
POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
2026
Top co-authors
Aviral Kumar
· 4
Amrith Setlur
· 3
Ruslan Salakhutdinov
· 3
Virginia Smith
· 2
Brandon Cui
· 1
David Acuna
· 1
Edward Emanuel Beeching
· 1
Eric Xing
· 1
Feng Yao
· 1
Hyunwoo Kim
· 1
Jaehun Jung
· 1
Lewis Tunstall
· 1
Topics
cs.LG
cs.AI
Meta-RL
RLHF & Alignment
Model-Based RL
Exploration
cs.CL
Offline RL
Multi-Agent