Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yang Yu — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Yang Yu
64
papers ·
38
citations ·
12
h-index
Education University of Hong Kong
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Heterogeneous Multi-agent Zero-Shot Coordination by Coevolution
2022 · 6 citations
NeoRL-2: Near Real-World Benchmarks for Offline Reinforcement Learning with Extended Realistic Scenarios
2025 · 5 citations
Multi-agent In-context Coordination via Decentralized Memory Retrieval
2025 · 1 citations
Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement Learning
2024 · 1 citations
Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
2026
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
2026
RLVR Training of LLMs Does Not Improve Thinking Ability for General QA: Evaluation Method and a Simple Solution
2026
Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention
2025
Pass@k Metric for RLVR: A Diagnostic Tool of Exploration, But Not an Objective
2025
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
2025
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference
2025
Multi-Agent Policy Transfer via Task Relationship Modeling
2022
A Note on Target Q-learning For Solving Finite MDPs with A Generative Oracle
2022
Transferable Reward Learning by Dynamics-Agnostic Discriminator Ensemble
2022
Learning Physically Realizable Skills for Online Packing of General 3D Shapes
2022
Top co-authors
Feng Chen
· 3
Lei Yuan
· 3
Zongzhang Zhang
· 3
Chao Qian
· 2
Cong Guan
· 2
Jing-Cheng Pang
· 2
Lanqing Li
· 2
Rong-Jun Qin
· 2
Tian Xu
· 2
Yi-Chen Li
· 2
Ziniu Li
· 2
Aaron Blakeman
· 1
Topics
Model-Based RL
Policy Gradient
cs.LG
RLHF & Alignment
Offline RL
Multi-Agent
Exploration
cs.AI
cs.MA
cs.RO