Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zi-Yi Dou — most-cited papers & profile · Multimodal
← authors
·
overview
Zi-Yi Dou
9
papers ·
131
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Coarse-to-Fine Vision-Language Pre-training with Fusion in the Backbone
2022 · 67 citations
MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models
2024 · 1 citations
Vapr -- Vision-language Preference Alignment For Reasoning
2025
ACQUIRED: A Dataset for Answering Counterfactual Questions In Real-Life Videos
2023
Top co-authors
Nanyun Peng
· 4
Aishwarya Kamath
· 1
Ce Liu
· 1
Fabrice Harel-Canada
· 1
Jia-Chen Gu
· 1
Jianfeng Gao
· 1
Jianfeng Wang
· 1
Kai-Wei Chang
· 1
Lijuan Wang
· 1
Linjie Li
· 1
Marjorie Freedman
· 1
Mohsen Fayyaz
· 1
Topics
Visual QA & Reasoning
Vision-Language Models
Video-Language
Benchmarks
Image-Text Retrieval