Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chengyue Wu — most-cited papers & profile · Multimodal
← authors
·
overview
Chengyue Wu
8
papers ·
25
citations ·
17
h-index
The University of Texas MD Anderson Cancer Center · The University of Texas at Austin
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
2024 · 24 citations
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
2025
Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition
2025
$π$-Tuning: Transferring Multimodal Foundation Models with Optimal Multi-task Interpolation
2023
JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation
2024
Top co-authors
Chong Ruan
· 2
Xiaokang Chen
· 2
Xingchao Liu
· 2
Yiyang Ma
· 2
Aixin Liu
· 1
Aoxue Li
· 1
Bingxuan Wang
· 1
Chi Chen
· 1
Damai Dai
· 1
Da Peng
· 1
Haowei Zhang
· 1
Haowei Zhang
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language