Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Huaxiu Yao — most-cited papers & profile · Multimodal
← authors
·
overview
Huaxiu Yao
27
papers ·
1097
citations ·
25
h-index
University of North Carolina at Chapel Hill
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding
2024 · 3 citations
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
2024 · 3 citations
AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving
2024 · 2 citations
Calibrated Self-Rewarding Vision Language Models
2024 · 1 citations
From Eduvisbench To Eduvisagent: A Benchmark And Multi-agent Framework For Reasoning-driven Pedagogical Visualization
2025
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning
2025
Re-Align: Aligning Vision Language Models via Retrieval-Augmented Direct Preference Optimization
2025
MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?
2024
MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models
2024
Top co-authors
Zhaorun Chen
· 5
Chenhang Cui
· 3
Mingyu Ding
· 2
Siwei Han
· 2
Xiyao Wang
· 2
Yiyang Zhou
· 2
Yiyang Zhou
· 2
Zhengzhong Tu
· 2
Zhuokai Zhao
· 2
Bo Li
· 1
Canyu Chen
· 1
Cao Xiao
· 1
Topics
Vision-Language Models
Benchmarks
Video-Language
Visual QA & Reasoning
Embodied & Agents
Image-Text Retrieval
Instruction Tuning