Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yongbin Li — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yongbin Li
75
papers ·
218
citations ·
22
h-index
Northwest University · First Hospital of Xi'an
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fine-Tuning Language Models with Reward Learning on Policy
2024 · 4 citations
Entropy Pacing Policy Optimization for Multi-Task Agentic Reinforcement Learning
2026
RollArt: Scaling Agentic RL Training via Disaggregated Infrastructure
2025
STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning
2026
ESPO: Early-Stopping Proximal Policy Optimization
2026
Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions
2026
Understanding Generalization in Role-Playing Models via Information Theory
2025
CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
2025
RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization
2025
A Simple "Motivation" Can Enhance Reinforcement Finetuning of Large Reasoning Models
2025
Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective Ambiguity
2025
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence
2025
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
2024
Top co-authors
Fei Huang
· 6
Binhua Li
· 4
Ting-En Lin
· 4
Dacheng Tao
· 3
Hao Lang
· 3
Junjie Zhang
· 3
Shunyu Liu
· 3
Ge Li
· 2
Guozheng Ma
· 2
Rongyu Cao
· 2
Tieyun Qian
· 2
Xiang Huang
· 2
Topics
Model-Based RL
RLHF & Alignment
cs.LG
cs.AI
Exploration
cs.CL
Multi-Agent
Value-Based
Policy Gradient
Game AI