Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuying Ge — most-cited papers & profile · Multimodal
← authors
·
overview
Yuying Ge
16
papers ·
42
citations ·
11
h-index
Tencent (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
JourneyDB: A Benchmark for Generative Image Understanding
2023 · 11 citations
SEED-Bench-2: Benchmarking Multimodal Large Language Models
2023 · 6 citations
VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation
2023 · 4 citations
SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension
2024 · 3 citations
SEED-Story: Multimodal Long Story Generation with Large Language Model
2024 · 2 citations
ViT-Lens: Towards Omni-modal Representations
2023 · 1 citations
SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
2024 · 1 citations
TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation
2025
Exploring the Effect of Reinforcement Learning on Video Understanding: Insights from SEED-Bench-R1
2025
Retrieving-to-Answer: Zero-Shot Video Question Answering with Frozen Large Language Models
2023
Top co-authors
Ying Shan
· 8
Yixiao Ge
· 8
Bohao Li
· 2
Hongsheng Li
· 2
Jinguo Zhu
· 2
Junting Pan
· 2
Kun Yi
· 2
Renrui Zhang
· 2
Ruimao Zhang
· 2
Yu Qiao
· 2
Aojun Zhou
· 1
Chen Li
· 1
Topics
Vision-Language Models
Benchmarks
Visual QA & Reasoning
Video-Language
Instruction Tuning
Image-Text Retrieval
Audio-Visual
Embodied & Agents