Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hang Xu — most-cited papers & profile · Multimodal
← authors
·
overview
Hang Xu
13
papers ·
39
citations ·
39
h-index
Huawei Technologies (China) · Wenzhou Medical University · Taizhou Central Hospital · Zhejiang Taizhou Hospital · Taizhou University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Wukong: A 100 Million Large-scale Chinese Cross-modal Pre-training Benchmark
2022 · 29 citations
CLIP$^2$: Contrastive Language-Image-Point Pretraining from Real-World Point Cloud Data
2023 · 3 citations
Open-world Semantic Segmentation via Contrasting and Clustering Vision-Language Embedding
2022 · 1 citations
UniGS: Unified Language-Image-3D Pretraining with Gaussian Splatting
2025
Visual-Language Navigation Pretraining via Prompt-based Environmental Self-exploration
2022
NLIP: Noise-robust Language-Image Pre-training
2022
Top co-authors
Xiaodan Liang
· 6
Chunjing Xu
· 3
Jianhua Han
· 3
Runhui Huang
· 2
Xiwen Liang
· 2
Yihan Zeng
· 2
Chaoqiang Ye
· 1
Chenhan Jiang
· 1
Dit-Yan Yeung
· 1
Fengda Zhu
· 1
Guansong Lu
· 1
Haoyuan Li
· 1
Topics
Vision-Language Models
Image-Text Retrieval
Video-Language
Benchmarks
Embodied & Agents
Instruction Tuning