Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Long Zhao — most-cited papers & profile · Multimodal
← authors
·
overview
Long Zhao
15
papers ·
130
citations ·
4
h-index
China Tobacco
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Distilling Vision-Language Models on Millions of Videos
2024 · 10 citations
More Than Just Attention: Improving Cross-Modal Attentions with Contrastive Constraints for Image-Text Matching
2021 · 1 citations
VULCAN: Tool-augmented Multi Agents For Iterative 3D Object Arrangement
2025
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
2025
Top co-authors
Dimitris N. Metaxas
· 2
Ting Liu
· 2
Yuxiao Chen
· 2
Boqing Gong
· 1
Chun-Te Chu
· 1
Di Liu
· 1
Florian Schroff
· 1
Haizhou Shi
· 1
Hao Wang
· 1
Hartwig Adam
· 1
Hui Miao
· 1
Jialin Wu
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Embodied & Agents
Image-Text Retrieval
Instruction Tuning