Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Delong Chen — most-cited papers & profile · Multimodal
← authors
·
overview
Delong Chen
8
papers ·
16
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
VL-JEPA: Joint Embedding Predictive Architecture For Vision-language
2025 · 1 citations
Planning With Reasoning Using Vision Language World Model
2025 · 1 citations
Chain-of-talkers (cotalk): Fast Human Annotation Of Dense Image Captions
2025
Visual Instruction Tuning with Polite Flamingo
2023
Few-shot Adaptation of Multi-modal Foundation Models: A Survey
2024
Top co-authors
Allen Bolourchi
· 2
Pascale Fung
· 2
Théo Moutakanni
· 2
Willy Chung
· 2
Baoyuan Wang
· 1
Chuanyi Zhang
· 1
Fan Liu
· 1
Fan Liu
· 1
Jianfeng Liu
· 1
J. M. Yu
· 1
Mustafa Shukor
· 1
Tejaswi Kasarla
· 1
Topics
Vision-Language Models
Video-Language
Embodied & Agents
Visual QA & Reasoning
Image-Text Retrieval
Instruction Tuning
Benchmarks