Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhenwen Liang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Zhenwen Liang
30
papers ·
213
citations ·
0
h-index
Shanghai University of Medicine and Health Sciences
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models
2025 · 26 citations
Reasoning or Memorization? Direction-Aware Diversity Exploration in LLM Reinforcement Learning
2026
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
2026
Save the Good Prefix: Precise Error Penalization via Process-Supervised RL to Enhance LLM Reasoning
2026
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
2026
Can LLMs Guide Their Own Exploration? Gradient-Guided Reinforcement Learning for LLM Reasoning
2025
Stable and Efficient Single-Rollout RL for Multimodal Reasoning
2025
Dual-Uncertainty Guided Policy Learning for Multimodal Reasoning
2025
Evolving Language Models without Labels: Majority Drives Selection, Novelty Promotes Variation
2025
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
2025
Top co-authors
Haitao Mi
· 9
Dong Yu
· 8
Dian Yu
· 7
Kishan Panaganti
· 5
Wenhao Yu
· 5
Haolin Liu
· 4
Linfeng Song
· 4
Rui Liu
· 4
Yujun Zhou
· 4
Pratap Tokekar
· 2
Runpeng Dai
· 2
Sidi Lu
· 2
Topics
cs.AI
cs.LG
cs.CL
Value-Based
cs.CV
Policy Gradient
Exploration
Model-Based RL
RLHF & Alignment