Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ganqu Cui — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Ganqu Cui
18
papers ·
9
citations ·
13
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
A Survey of Reinforcement Learning for Large Reasoning Models
2025 · 2 citations
Process Reinforcement through Implicit Rewards
2025 · 1 citations
Free Process Rewards without Process Labels
2024 · 1 citations
How Far Can Unsupervised RLVR Scale LLM Training?
2026
P1: Mastering Physics Olympiads with Reinforcement Learning
2025
FlowRL: Matching Reward Distributions for LLM Reasoning
2025
Wisdom of the Crowd: Reinforcement Learning from Coevolutionary Collective Feedback
2025
MiniCPM4: Ultra-Efficient LLMs on End Devices
2025
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
2025
Learning to Reason under Off-Policy Guidance
2025
TTRL: Test-Time Reinforcement Learning
2025
Top co-authors
Bowen Zhou
· 6
Ning Ding
· 6
Yuxin Zuo
· 5
Lifan Yuan
· 4
Xingtai Lv
· 4
Youbang Sun
· 4
Bingxiang He
· 3
Ermo Hua
· 3
Haozhan Li
· 3
Huayu Chen
· 3
Kaiyan Zhang
· 3
Li Sheng
· 3
Topics
RLHF & Alignment
Model-Based RL
Policy Gradient
Exploration
cs.CL
Offline RL
Multi-Agent
cs.AR
cs.DC
Meta-RL