Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yiming Zhang — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yiming Zhang
60
papers ·
1721
citations ·
28
h-index
Guangdong University of Technology · Xiamen University · Peking University · People's Liberation Army No. 150 Hospital · Xi'an Jiaotong University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
First Order Constrained Optimization in Policy Space
2020 · 52 citations
On-Policy Deep Reinforcement Learning for the Average-Reward Criterion
2021 · 12 citations
Supervised Policy Update for Deep Reinforcement Learning
2018 · 3 citations
Aggressive Q-Learning with Ensembles: Achieving Both High Sample Efficiency and High Asymptotic Performance
2021 · 3 citations
Efficient Entropy for Policy Gradient with Multidimensional Action Space
2018 · 1 citations
The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs
2026
Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning
2025
Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?
2025
Towards Large-Scale In-Context Reinforcement Learning by Meta-Training in Randomized Worlds
2025
Multi-Agent Reinforcement Learning for Multi-Cell Spectrum and Power Allocation
2023
Top co-authors
Quan Vuong
· 2
Zihan Chen
· 2
Bo Yu
· 1
Cho-Jui Hsieh
· 1
Dahua Lin
· 1
Fan Wang
· 1
Fan Zheng
· 1
Haifeng Wang
· 1
Haiteng Zhao
· 1
Hengguang Zhou
· 1
Junhao Shen
· 1
Kai Chen
· 1
Topics
Policy Gradient
cs.AI
Exploration
cs.LG
Model-Based RL
Safe RL
Multi-Agent
Meta-RL
Value-Based