Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shao-Yen Tseng — most-cited papers & profile · Multimodal
← authors
·
overview
Shao-Yen Tseng
7
papers ·
8
citations ·
10
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
KD-VLP: Improving End-to-End Vision-and-Language Pretraining with Object Knowledge Distillation
2021 · 4 citations
VL-InterpreT: An Interactive Visualization Tool for Interpreting Vision-Language Transformers
2022 · 3 citations
MuMUR : Multilingual Multimodal Universal Retrieval
2022
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
2024
Top co-authors
Vasudev Lal
· 4
Estelle Aflalo
· 3
Chenfei Wu
· 2
Gabriela Ben Melech Stan
· 2
Nan Duan
· 2
Yongfei Liu
· 2
Avinash Madasu
· 1
Gedas Bertasius
· 1
Man Luo
· 1
Meng Du
· 1
Sayak Paul
· 1
Shachar Rosenman
· 1
Topics
Vision-Language Models
Video-Language
Image-Text Retrieval
Visual QA & Reasoning
Benchmarks