Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xuelong Li — most-cited papers & profile · Multimodal
← authors
·
overview
Xuelong Li
126
papers ·
183
citations ·
0
h-index
Anhui University · Hefei University of Technology · Northwestern Polytechnical University · China Telecom (China) · China Telecom
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Open-Vocabulary Octree-Graph for 3D Scene Understanding
2024 · 9 citations
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
2025
Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives
2025
VLMQ: Token Saliency-Driven Post-Training Quantization for Vision-language Models
2025
HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models
2025
Decoupling The Image Perception And Multimodal Reasoning For Reasoning Segmentation With Digital Twin Representations
2025
From Captions to Rewards (CAREVL): Leveraging Large Language Model Experts for Enhanced Reward Modeling in Large Vision-Language Models
2025
Openfly: A comprehensive platform for aerial vision-language navigation
2025
Vision-to-Language Tasks Based on Attributes and Attention Mechanism
2019
Top co-authors
Chenhui Li
· 2
Jiawei Shao
· 2
Aihong Yuan and Xiaoqiang Lu
· 1
Bin Zhao
· 1
Bin Zhao
· 1
Chaoya Jiang
· 1
Chi Zhang
· 1
Chi Zhang
· 1
Dell Zhang
· 1
Dong Wang
· 1
Dong Wang
· 1
Haibin Huang
· 1
Topics
Vision-Language Models
Benchmarks
Video-Language
Visual QA & Reasoning
Embodied & Agents
Image-Text Retrieval