Awesome Reinforcement Learning
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Reading Packs
News
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Reading Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Ganqu Cui — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Ganqu Cui
52
papers ·
497
citations ·
13
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
2025 · 83 citations
RLPR: Extrapolating RLVR to General Domains without Verifiers
2025 · 69 citations
NFT: Bridging Supervised Learning and Reinforcement Learning in Math Reasoning
2025 · 28 citations
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
2026
TEMPO: Scaling Test-time Training for Large Reasoning Models
2026
Teaching Large Reasoning Models Effective Reflection
2026
From f(x) and g(x) to f(g(x)): LLMs Learn New Skills in RL by Composing Old Ones
2025
Intern-S1: A Scientific Multimodal Foundation Model
2025
Noise Contrastive Alignment of Language Models with Explicit Rewards
2024
Advancing LLM Reasoning Generalists with Preference Trees
2024
Top co-authors
Ning Ding
· 6
Lifan Yuan
· 5
Bowen Zhou
· 4
Hanbin Wang
· 3
Maosong Sun
· 3
Yu Cheng
· 3
Zhiyuan Liu
· 3
Hao Peng
· 2
Huayu Chen
· 2
Jun Zhu
· 2
Xiaoye Qu
· 2
Yuchen Fan
· 2
Topics
cs.CL
cs.AI
cs.LG
cs.RO
Model-Based RL
Offline RL
RLHF & Alignment