Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chi Zhang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Chi Zhang
202
papers ·
4897
citations ·
14
h-index
University of South China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
2025 · 1767 citations
Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks
2023 · 11 citations
Buffer Pool Aware Query Scheduling via Deep Reinforcement Learning
2020 · 10 citations
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations
2021 · 7 citations
Constrained Reinforcement Learning for Short Video Recommendation
2022 · 5 citations
FAPO: Flawed-Aware Policy Optimization for Efficient and Reliable Reasoning
2025 · 3 citations
Multi-task Offline Reinforcement Learning for Online Advertising in Recommender Systems
2025 · 3 citations
Learning Unmanned Aerial Vehicle Control for Autonomous Target Following
2017 · 3 citations
Dynamic Dispatching for Large-Scale Heterogeneous Fleet via Multi-agent Deep Reinforcement Learning
2020 · 3 citations
BRAC+: Improved Behavior Regularized Actor Critic for Offline Reinforcement Learning
2021 · 3 citations
Maximum Entropy Model Rollouts: Fast Model Based Policy Optimization without Compounding Errors
2020 · 2 citations
Two-Stage Constrained Actor-Critic for Short Video Recommendation
2023 · 2 citations
VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks
2025 · 1 citations
DianJin-R1: Evaluating and Enhancing Financial Reasoning in Large Language Models
2025 · 1 citations
Learning Practical Communication Strategies in Cooperative Multi-Agent Reinforcement Learning
2022 · 1 citations
Top co-authors
Haibin Lin
· 5
Xin Liu
· 4
Bole Ma
· 3
Gaohong Liu
· 3
Jiaze Chen
· 3
LingJun Liu
· 3
Lin Yan
· 3
Mingxuan Wang
· 3
Mofan Zhang
· 3
Qiying Yu
· 3
Ruofei Zhu
· 3
Xiaochen Zuo
· 3
Topics
Policy Gradient
Value-Based
Model-Based RL
Safe RL
cs.LG
RLHF & Alignment
Offline RL
Multi-Agent
cs.AI
Meta-RL