Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiangyang Ji — most-cited papers & profile · Multimodal
← authors
·
overview
Xiangyang Ji
73
papers ·
238
citations ·
48
h-index
Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Video-Text as Game Players: Hierarchical Banzhaf Interaction for Cross-Modal Representation Learning
2023 · 3 citations
CCMB: A Large-scale Chinese Cross-modal Benchmark
2022 · 1 citations
Look Clearly Before Answering: Mitigating Hallucinations in LVLMs via Saliency-Driven Perceptual Realignment
2026
DynFly: Dynamic-Aware Continuous Trajectory Generation for UAV Vision-Language Navigation in Urban Environments
2026
EventFlash: Towards Efficient MLLMs for Event-Based Vision
2026
EventGPT: Event Stream Understanding with Multimodal Large Language Models
2024
Top co-authors
Jianing Li
· 2
Ming Li
· 2
Wen Jiang
· 2
Bin Xu
· 1
Chang Liu
· 1
Fei Richard Yu
· 1
Jie Chen
· 1
Jinfa Huang
· 1
Liang Zhang
· 1
Lin Yao
· 1
Li Wang
· 1
Li Yuan
· 1
Topics
Vision-Language Models
Benchmarks
Video-Language
Visual QA & Reasoning
Instruction Tuning
Image-Text Retrieval
Embodied & Agents