Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuxuan Liang — most-cited papers & profile · Multimodal
← authors
·
overview
Yuxuan Liang
39
papers ·
1332
citations ·
39
h-index
Guangdong University of Technology · Beijing University of Posts and Telecommunications · Hong Kong University of Science and Technology · Guangzhou University · University of Hong Kong · South China University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Pyramid Token Pruning for High-Resolution Large Vision-Language Models via Region, Token, and Instruction-Guided Importance
2025
HERO: Rethinking Visual Token Early Dropping In High-resolution Large Vision-language Models
2025
Recognition Through Reasoning: Reinforcing Image Geo-localization With Large Vision-language Models
2025
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
2025
Top co-authors
Xiaolei Chen
· 2
Xu Li
· 2
Bin Li
· 1
Haotian Chen
· 1
Huan Li
· 1
Ling Li
· 1
Ming Jin
· 1
Qingsong Wen
· 1
Siru Zhong
· 1
Weilin Ruan
· 1
Xiangyang Xue
· 1
Yao Zhou
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Instruction Tuning
Visual QA & Reasoning