Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Heng Tao Shen — most-cited papers & profile · Multimodal
← authors
·
overview
Heng Tao Shen
24
papers ·
207
citations ·
94
h-index
Tongji University · University of Electronic Science and Technology of China · Hubei Provincial Center for Disease Control and Prevention · Yibin University · Hangzhou Medical College
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Universal Weighting Metric Learning for Cross-Modal Matching
2020 · 107 citations
Support-set based Multi-modal Representation Enhancement for Video Captioning
2022 · 3 citations
Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
2024 · 2 citations
MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct
2024 · 1 citations
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
2025
LISA: A Layer-wise Integration and Suppression Approach for Hallucination Mitigation in Multimodal Large Language Models
2025
Top co-authors
Lianli Gao
· 3
Pengpeng Zeng
· 3
Jingkuan Song
· 2
Yang Yang
· 2
Beitao Chen
· 1
Fei Huang
· 1
Haonan Zhang
· 1
Haoxi Zeng
· 1
Haoxuan Li
· 1
Hui Xu
· 1
Jie Shao
· 1
Jingkuan Song
· 1
Topics
Vision-Language Models
Video-Language
Image-Text Retrieval
Benchmarks
cs.CV
Instruction Tuning
Visual QA & Reasoning