Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Muning Wen — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Muning Wen
21
papers ·
256
citations ·
8
h-index
Shanghai Jiao Tong University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning
2021 · 83 citations
Multi-Agent Reinforcement Learning is a Sequence Modeling Problem
2022 · 79 citations
MALib: A Parallel Framework for Population-based Multi-agent Reinforcement Learning
2021 · 25 citations
Settling the Variance of Multi-Agent Policy Gradients
2021 · 24 citations
Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks
2021 · 11 citations
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
2024 · 2 citations
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning
2025 · 1 citations
Reinforcing Language Agents via Policy Optimization with Action Decomposition
2024 · 1 citations
PARL-MT: Learning to Call Functions in Multi-Turn Conversation with Progress Awareness
2025
MARFT: Multi-Agent Reinforcement Fine-Tuning
2025
Learning Humanoid Standing-up Control across Diverse Postures
2025
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning
2025
Entropy-Regularized Token-Level Policy Optimization for Language Agent Reinforcement
2024
Autonomous Goal Detection and Cessation in Reinforcement Learning: A Case Study on Source Term Estimation
2024
Top co-authors
Weinan Zhang
· 6
Jun Wang
· 5
Yaodong Yang
· 5
Jakub Grudzien Kuba
· 3
Ying Wen
· 3
Junwei Liao
· 2
Shangding Gu
· 2
Yiwei Shi
· 2
Ziyu Wan
· 2
Adam Wierman
· 1
Anjie Liu
· 1
Bo Xu
· 1
Topics
Multi-Agent
Model-Based RL
Policy Gradient
Game AI
RLHF & Alignment
Exploration
Safe RL
Meta-RL
Offline RL
Value-Based