Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Peng Jiang — most-cited papers & profile · Multimodal
← authors
·
overview
Peng Jiang
39
papers ·
388
citations ·
6
h-index
University of Iowa · State Grid Corporation of China (China) · Dongguan People’s Hospital · Kuaishou (China) · University of Science and Technology Beijing
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
M3oE: Multi-Domain Multi-Task Mixture-of Experts Recommendation Framework
2024 · 68 citations
Multi-Task Recommendations with Reinforcement Learning
2023 · 57 citations
Multi-Task Recommendations with Reinforcement Learning
2023 · 57 citations
Exploration and Regularization of the Latent Action Space in Recommendation
2023 · 54 citations
Generative Flow Network for Listwise Recommendation
2023 · 43 citations
Towards Robust Recommendation via Decision Boundary-aware Graph Contrastive Learning
2024 · 21 citations
DAS: Dual-Aligned Semantic IDs Empowered Industrial Recommender System
2025 · 16 citations
Modeling User Retention through Generative Flow Networks
2024 · 13 citations
Modeling User Fatigue for Sequential Recommendation
2024 · 12 citations
Future Impact Decomposition in Request-level Recommendations
2024 · 11 citations
TrackRec: Iterative Alternating Feedback with Chain-of-Thought via Preference Alignment for Recommendation
2025 · 8 citations
ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual Actor
2022 · 6 citations
Constrained Reinforcement Learning for Short Video Recommendation
2022 · 5 citations
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer
2025 · 4 citations
Reinforcing User Retention in a Billion Scale Short Video Recommender System
2023 · 4 citations
Topics
Collaborative Filtering
Ranking & CTR
Value-Based
Policy Gradient
Offline RL
Model-Based RL
Sequential & Session
Safe RL
Exploration
LLM-based