Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Mingli Song — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Mingli Song
40
papers ·
179
citations ·
50
h-index
Beijing Institute of Fashion Technology · Zhejiang University of Science and Technology · Zhejiang University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Is Centralized Training with Decentralized Execution Framework Centralized Enough for MARL?
2023 · 12 citations
Ask-AC: An Initiative Advisor-in-the-Loop Actor-Critic Framework
2022 · 2 citations
Bi-level Mean Field: Dynamic Grouping for Large-Scale MARL
2025 · 1 citations
Interaction Pattern Disentangling for Multi-Agent Reinforcement Learning
2022 · 1 citations
Mixture-of-Schedulers: An Adaptive Scheduling Agent as a Learned Router for Expert Policies
2025
Breaking the Exploration Bottleneck: Rubric-Scaffolded Reinforcement Learning for General LLM Reasoning
2025
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
2025
SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data
2025
Contrastive Identity-Aware Learning for Multi-Agent Value Decomposition
2022
Curricular Subgoals for Inverse Reinforcement Learning
2023
Agent-Aware Training for Agent-Agnostic Action Advising in Deep Reinforcement Learning
2023
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
2024
Transmission Interface Power Flow Adjustment: A Deep Reinforcement Learning Approach based on Multi-task Attribution Map
2024
Top co-authors
Shunyu Liu
· 12
Kaixuan Chen
· 6
Yihe Zhou
· 6
Tongya Zheng
· 5
Jie Song
· 4
Kongcheng Zhang
· 3
Yunpeng Qing
· 3
Zunlei Feng
· 3
Dacheng Tao
· 2
Jingyuan Cong
· 2
Kaixuan Chen
· 2
Na Yu
· 2
Topics
Model-Based RL
Multi-Agent
Meta-RL
Offline RL
Exploration
RLHF & Alignment
Policy Gradient
Value-Based
Game AI