Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhengzhong Tu — most-cited papers & profile · Multimodal
← authors
·
overview
Zhengzhong Tu
22
papers ·
912
citations ·
16
h-index
Texas A&M University System · Mitchell Institute · Texas A&M University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving
2024 · 2 citations
mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation
2025 · 1 citations
T2T-VICL: Unlocking the Boundaries of Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs
2025
Edge-based Multimodal Sensor Data Fusion With Vision Language Models (vlms) For Real-time Autonomous Vehicle Accident Avoidance
2025
MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning
2025
Re-Align: Aligning Vision Language Models via Retrieval-Augmented Direct Preference Optimization
2025
Top co-authors
Chan-Wei Hu
· 2
Huaxiu Yao
· 2
Chengxuan Qian
· 1
Chenxi Liu
· 1
Chia-Ju Chen
· 1
Fengze Yang
· 1
Heng Huang
· 1
Hongyuan Hua
· 1
Huixin Zhang
· 1
James Tompkin
· 1
Jiacheng Zhu
· 1
Jielin Qiu
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Video-Language
Benchmarks
Image-Text Retrieval