Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Qi Chen — most-cited papers & profile · Multimodal
← authors
·
overview
Qi Chen
50
papers ·
307
citations ·
15
h-index
University of Science and Technology of China · Microsoft Research (United Kingdom)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training
2024 · 3 citations
Beyond the Global Scores: Fine-Grained Token Grounding as a Robust Detector of LVLM Hallucinations
2026
3D-DRES: Detailed 3D Referring Expression Segmentation
2026
Overthinking Causes Hallucination: Tracing Confounder Propagation in Vision Language Models
2026
Long-horizon Visual Imitation Learning Via Plan And Code Reflection
2025
Localizing Before Answering: A Hallucination Evaluation Benchmark for Grounded Medical Multimodal LLMs
2025
Attention-driven GUI Grounding: Leveraging Pretrained Multimodal Large Language Models without Fine-Tuning
2024
Top co-authors
Minh Khoi Ho
· 3
Phi Le Nguyen
· 3
Anton van den Hengel
· 2
Johan W. Verjans
· 2
Tuan Dung Nguyen
· 2
Vu Minh Hieu Phan
· 2
Yutong Xie
· 2
Abin Shoby
· 1
and Vu Minh Hieu Phan
· 1
Biao Wu
· 1
Changli Wu
· 1
Cheng Chang
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language
Embodied & Agents
Image-Text Retrieval