Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Linke Ouyang — most-cited papers & profile · Multimodal
← authors
·
overview
Linke Ouyang
3
papers ·
42
citations ·
7
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition
2023 · 31 citations
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
2024 · 11 citations
Docr-inspector: Fine-grained And Automated Evaluation Of Document Parsing With VLM
2025
Top co-authors
Bin Wang
· 2
Conghui He
· 2
Dahua Lin
· 2
Hang Yan
· 2
Haodong Duan
· 2
Pan Zhang
· 2
Songyang Zhang
· 2
Wei Li
· 2
Xiaoyi Dong
· 2
Xingcheng Zhang
· 2
Xinyue Zhang
· 2
Yuhang Cao
· 2
Topics
Benchmarks
Vision-Language Models
Image-Text Retrieval
Video-Language