Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yinan Zheng — most-cited papers & profile · Multimodal
← authors
·
overview
Yinan Zheng
13
papers ·
15
citations ·
42
h-index
Northwestern University · Shanxi University of Traditional Chinese Medicine · Guangdong Province Women and Children Hospital
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining
2026
X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
2025
Physiagent: An Embodied Agent Framework In Physical World
2025
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning
2024
Top co-authors
Xianyuan Zhan
· 4
Jianxiong Li
· 3
Jinliang Zheng
· 3
Dongxiu Liu
· 2
Haoyi Niu
· 2
Jingjing Liu
· 2
Zhihao Wang
· 2
D G Liu
· 1
Hang Su
· 1
Hao Wang
· 1
Jiangmiao Pang
· 1
Jiayin Zou
· 1
Topics
Vision-Language Models
Embodied & Agents
Video-Language
Benchmarks