Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tao Kong — most-cited papers & profile · Multimodal
← authors
·
overview
Tao Kong
14
papers ·
714
citations ·
29
h-index
Binzhou University · Hefei Urban Planning & Design Institute · Geological Exploration Institute of Shandong Zhengyuan
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
What Matters in Training a GPT4-Style Language Model with Multimodal Inputs?
2023 · 1 citations
Bridgevla: Input-output Alignment For Efficient 3D Manipulation Learning With Vision-language Models
2025
What Matters in Building Vision-Language-Action Models for Generalist Robots
2024
Top co-authors
Hanbo Zhang
· 2
Peiyan Li
· 2
Bingyi Kang
· 1
Di Guo
· 1
Dong Wang
· 1
Guoqiang Wei
· 1
Hongtao Wu
· 1
Huaping Liu
· 1
Jiangnan Xia
· 1
Jiani Zheng
· 1
Jirong Liu
· 1
Liang Wang
· 1
Topics
Vision-Language Models
Video-Language
Embodied & Agents
Benchmarks
Instruction Tuning