Awesome Reinforcement Learning
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiangyu Zhao — most-cited papers & profile · Reinforcement Learning
← authors
·
overview
Xiangyu Zhao
108
papers ·
1424
citations ·
43
h-index
City University of Hong Kong · China Aerospace Science and Technology Corporation · China Southern Power Grid (China) · Nanjing University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Reinforcement Learning for List-wise Recommendations
2018 · 109 citations
Multi-Task Recommendations with Reinforcement Learning
2023 · 57 citations
Deep reinforcement learning for search, recommendation, and online advertising: a survey
2018 · 44 citations
Toward Simulating Environments in Reinforcement Learning Based Recommendations
2019 · 17 citations
DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender Systems
2019 · 9 citations
Multi-task Offline Reinforcement Learning for Online Advertising in Recommender Systems
2025 · 3 citations
Reasoning through Exploration: A Reinforcement Learning Framework for Robust Function Calling
2025 · 1 citations
Data-Efficient Reinforcement Learning for Malaria Control
2021 · 1 citations
Building a 3-Player Mahjong AI using Deep Reinforcement Learning
2022 · 1 citations
A General Neural Causal Model for Interactive Recommendation
2023 · 1 citations
DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent
2026
Fairness Begins with State: Purifying Latent Preferences for Hierarchical Reinforcement Learning in Interactive Recommendation
2026
Efficient Reasoning via Reward Model
2025
Intern-S1: A Scientific Multimodal Foundation Model
2025
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
2025
Top co-authors
Dawei Yin
· 4
Jiliang Tang
· 4
Jingtong Gao
· 3
Long Xia
· 3
Peng Jiang
· 3
Qingpeng Cai
· 3
Dong Wang
· 2
Jialin Liu
· 2
Jun Li
· 2
Kun Gai
· 2
Lixin Zou
· 2
Maolin Wang
· 2
Topics
Model-Based RL
Policy Gradient
Offline RL
Exploration
Value-Based
Multi-Agent
RLHF & Alignment
cs.AI
cs.LG
Meta-RL