Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zheng-Jun Zha — most-cited papers & profile · Multimodal
← authors
·
overview
Zheng-Jun Zha
18
papers ·
172
citations ·
74
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Stacked Convolutional Deep Encoding Network for Video-Text Retrieval
2020 · 12 citations
Visionselector: End-to-end Learnable Visual Token Compression For Efficient Multimodal Llms
2025 · 1 citations
WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query Tokens
2025
Top co-authors
Dacheng Yin
· 1
Dong Li
· 1
Fengyun Rao
· 1
Jian Yang
· 1
Jing Lyu
· 1
Kecheng Zheng
· 1
Kunlin Liu
· 1
Rui Zhao
· 1
Wei Zhai
· 1
Wenrui Yan
· 1
Xiaoxuan He
· 1
Xin Lü
· 1
Topics
Vision-Language Models
cs.CV
Image-Text Retrieval
Video-Language