Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xun Guo — most-cited papers & profile · Multimodal
← authors
·
overview
Xun Guo
13
papers ·
11
citations ·
15
h-index
City University of Hong Kong · Jilin University · Nanjing University of Science and Technology · Wuhan University · Institute of Disaster Prevention · First Hospital of Jilin University · China Earthquake Administration
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
StableVideo: Text-driven Consistency-aware Diffusion Video Editing
2023 · 5 citations
MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
2023 · 4 citations
Semantic-aligned Fusion Transformer for One-shot Object Detection
2022 · 2 citations
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
2026
SciForma: Structure-Faithful Generation of Scientific Diagrams
2026
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model
2026
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
2026
VideoWorld: Exploring Knowledge Learning from Unlabeled Videos
2025
Top co-authors
Bin Li
· 1
Haoqing Wang
· 1
Jiahao Li
· 1
Jinghao Guo
· 1
Jinglu Wang
· 1
Kaichen Zhang
· 1
Peng Zhang
· 1
Senqiao Yang
· 1
Shicheng Zheng
· 1
Wenxuan Xie
· 1
Xiang An
· 1
Xiao Li
· 1
Topics
Model Architecture
Efficiency
Training Techniques
cs.CE
Evaluation
Benchmarks
Code Agents
Uncategorized
Reinforcement Learning
Object Detection