Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Weiming Hu — most-cited papers & profile · Multimodal
← authors
·
overview
Weiming Hu
26
papers ·
175
citations ·
67
h-index
Chinese Academy of Sciences · Institute of Automation
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning
2022 · 11 citations
MIBench: Evaluating Multimodal Large Language Models over Multiple Images
2024 · 1 citations
Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval
2026
Autoprune: Each Complexity Deserves A Pruning Policy
2025
MMPhysVideo: Scaling Physical Plausibility in Video Generation via Joint Multimodal Modeling
2026
Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift
2026
OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models
2026
MMhops-R1: Multimodal Multi-hop Reasoning
2025
Semantics-enhanced Cross-modal Masked Image Modeling for Vision-Language Pre-training
2024
Top co-authors
Bing Li
· 5
Chunfeng Yuan
· 5
Fei Huang
· 2
Haiyang Xu
· 2
Haowei Liu
· 2
Ji Zhang
· 2
Ming Yan
· 2
Yaya Shi
· 2
Yuxin Chen
· 2
Zhipeng Zhang
· 2
Ziqi Zhang
· 2
Zongyang Ma
· 2
Topics
Vision-Language Models
Benchmarks
Video-Language
Visual QA & Reasoning
Image-Text Retrieval
Embodied & Agents