Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ziwei Liu — most-cited papers & profile · Multimodal
← authors
·
overview
Ziwei Liu
45
papers ·
932
citations ·
82
h-index
Nanyang Technological University · Harbin Institute of Technology · Sichuan University · State Key Laboratory of Robotics and Systems
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MMBench: Is Your Multi-modal Model an All-around Player?
2023 · 33 citations
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
2023 · 32 citations
MIMIC-IT: Multi-Modal In-Context Instruction Tuning
2023 · 28 citations
Large Language Models are Visual Reasoning Coordinators
2023 · 14 citations
Learning without Forgetting for Vision-Language Models
2023 · 5 citations
Detecting and Grounding Multi-Modal Media Manipulation
2023 · 3 citations
Unsolvable Problem Detection: Robust Understanding Evaluation for Large Multimodal Models
2024 · 1 citations
Streamline Without Sacrifice -- Squeeze Out Computation Redundancy In LMM
2025
Collaborative Multi-Modal Coding for High-Quality 3D Generation
2025
Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning
2025
ShotBench: Expert-Level Cinematic Understanding in Vision-Language Models
2025
EgoLM: Multi-Modal Language Model of Egocentric Motions
2024
Top co-authors
Jingkang Yang
· 4
Bo Li
· 3
Yuanhan Zhang
· 3
Chunyuan Li
· 2
Liangyu Chen
· 2
Atsuyuki Miyai
· 1
Conghui He
· 1
Conghui He
· 1
Dahua Lin
· 1
Da-Wei Zhou
· 1
De-Chuan Zhan
· 1
Dian Zheng
· 1
Topics
Vision-Language Models
Video-Language
Visual QA & Reasoning
Benchmarks
Instruction Tuning
cs.CG
cs.GR
cs.MM
Embodied & Agents
Image-Text Retrieval