Awesome Reinforcement Learning
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Qingpeng Cai โ most-cited papers & profile ยท Reinforcement Learning
โ authors
ยท
overview
Qingpeng Cai
36
papers ยท
440
citations ยท
18
h-index
Wenzhou Medical University ยท Kuaishou (China)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Multi-Task Recommendations with Reinforcement Learning
2023 ยท 57 citations
Softmax Deep Double Deterministic Policy Gradients
2020 ยท 45 citations
A Deep Reinforcement Learning Framework for Rebalancing Dockless Bike Sharing Systems
2018 ยท 23 citations
ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual Actor
2022 ยท 6 citations
Constrained Reinforcement Learning for Short Video Recommendation
2022 ยท 5 citations
Reinforcement Learning with Dynamic Boltzmann Softmax Updates
2019 ยท 4 citations
Reinforcing User Retention in a Billion Scale Short Video Recommender System
2023 ยท 4 citations
Reinforcement Mechanism Design for e-commerce
2017 ยท 3 citations
Multi-Path Policy Optimization
2019 ยท 3 citations
PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement
2022 ยท 3 citations
Policy Optimization with Model-based Explorations
2018 ยท 2 citations
Generator and Critic: A Deep Reinforcement Learning Approach for Slate Re-ranking in E-commerce
2020 ยท 2 citations
Two-Stage Constrained Actor-Critic for Short Video Recommendation
2023 ยท 2 citations
State Regularized Policy Optimization on Data with Dynamics Shift
2023 ยท 2 citations
Deterministic Policy Gradients With General State Transitions
2018 ยท 1 citations
Top co-authors
Peng Jiang
ยท 12
Ling Pan
ยท 9
Kun Gai
ยท 8
Dong Zheng
ยท 7
Bo An
ยท 6
Shuchang Liu
ยท 5
Longbo Huang
ยท 4
Zhenghai Xue
ยท 4
Jingtong Gao
ยท 3
Xiangyu Zhao
ยท 3
Bin Yang
ยท 2
Chi Zhang
ยท 2
Topics
Policy Gradient
Value-Based
Exploration
Model-Based RL
Offline RL
Safe RL
cs.AI
Multi-Agent
Game AI
cs.CL