Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiao Liu — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Xiao Liu
111
papers ·
1669
citations ·
0
h-index
Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
2025 · 206 citations
Multi-Agent Reinforcement Learning in NOMA-aided UAV Networks for Cellular Offloading
2020 · 2 citations
MobileRL: Online Agentic Reinforcement Learning for Mobile GUI Agents
2025 · 1 citations
Learning Explicit Credit Assignment for Cooperative Multi-Agent Reinforcement Learning via Polarization Policy Gradient
2022 · 1 citations
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions
2024 · 1 citations
LongCat-Flash-Thinking-2601 Technical Report
2026
CFLight: Enhancing Safety with Traffic Signal Control through Counterfactual Learning
2025
AgentRL: Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework
2025
ComputerRL: Scaling End-to-End Online Reinforcement Learning for Computer Use Agents
2025
Keeping Minimal Experience to Achieve Efficient Interpretable Policy Distillation
2022
BCRLSP: An Offline Reinforcement Learning Framework for Sequential Targeted Promotion
2022
Fidelity-Induced Interpretable Policy Extraction for Reinforcement Learning
2023
Safe Hybrid-Action Reinforcement Learning-Based Decision and Control for Discretionary Lane Change
2024
Autonomous vehicle decision and control through reinforcement learning with traffic flow randomization
2024
Does RLHF Scale? Exploring the Impacts From Data, Model, and Method
2024
Top co-authors
Yuxiao Dong
· 5
Hanchen Zhang
· 3
Jie Tang
· 3
Zhenyu Hou
· 3
Aohan Zeng
· 2
Hanyu Lai
· 2
Hongning Wang
· 2
Minlie Huang
· 2
Yang Gao
· 2
Yifan Xu
· 2
Yilin Niu
· 2
Yuan Lin
· 2
Topics
Policy Gradient
Multi-Agent
Model-Based RL
Value-Based
cs.AI
Safe RL
Game AI
RLHF & Alignment
cs.LG
Offline RL