Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiaqi Liu — most-cited papers & profile · Multimodal
← authors
·
overview
Jiaqi Liu
67
papers ·
45
citations ·
20
h-index
China Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Towards Faithful Reasoning in Remote Sensing: A Perceptually-Grounded GeoSpatial Chain-of-Thought for Vision-Language Models
2025 · 7 citations
Lead: The LLM Enhanced Planning System Converged With End-to-end Autonomous Driving
2025 · 2 citations
Skymoe: A Vision-language Foundation Model For Enhancing Geospatial Interpretation With Mixture Of Experts
2025 · 1 citations
Learning Action Priors for Cross-embodiment Robot Manipulation
2026
AA: A Multi-view Multimodal Dataset for Screen-based Gaze Estimation
2026
Mixture Of Horizons In Action Chunking
2025
U-VLM: Hierarchical Vision Language Modeling for Report Generation
2026
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
2026
SimpleOCR: Rendering Visualized Questions to Teach MLLMs to Read
2026
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning
2025
BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset
2025
Geodit: A Diffusion-based Vision-language Model For Geospatial Understanding
2025
Top co-authors
Dong Jing
· 3
Mingyu Ding
· 3
Haoran Liu
· 2
Huaxiu Yao
· 2
Peng Xia
· 2
Siwei Han
· 2
Yiyang Zhou
· 2
Bo Yang
· 1
Chang Liu
· 1
Chengkai Xu
· 1
Gang Wang
· 1
Guanyu Li
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Embodied & Agents
Visual QA & Reasoning