Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhongyi Zhou — most-cited papers & profile · Multimodal
← authors
·
overview
Zhongyi Zhou
4
papers ·
4
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
2025 · 2 citations
See2Refine: Vision-Language Feedback Improves LLM-Based eHMI Action Designers
2026
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
2025
Top co-authors
Yichen Zhu
· 2
Chaomin Shen
· 1
Chaomin Shen
· 1
Ding Xia
· 1
Dongyuan Li
· 1
Fan Gao
· 1
Feifei Feng
· 1
Junjie Wen
· 1
Junjie Wen
· 1
Mark Colley
· 1
Minjie Zhu
· 1
Ning Liu
· 1
Topics
Vision-Language Models
Video-Language
Embodied & Agents
Visual QA & Reasoning