Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yangyang Guo — most-cited papers & profile · Multimodal
← authors
·
overview
Yangyang Guo
11
papers ·
35
citations ·
16
h-index
Anhui University · Northwest University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Answer Questions with Right Image Regions: A Visual Attention Regularization Approach
2021 · 5 citations
ELIP: Efficient Discriminative Language-Image Pre-training with Fewer Vision Tokens
2023 · 1 citations
On Modality Bias Recognition and Reduction
2022
A Unified End-to-End Retriever-Reader Framework for Knowledge-based VQA
2022
Do Vision-Language Transformers Exhibit Visual Commonsense? An Empirical Study of VCR
2024
Top co-authors
Liqiang Nie
· 5
Mohan Kankanhalli
· 4
Yibing Liu
· 2
Yongkang Wong
· 2
Zhiyong Cheng
· 2
Alberto Del Bimbo
· 1
Haoyu Zhang
· 1
Harry Cheng
· 1
Jianhua Yin
· 1
Kejie Wang
· 1
Weifeng Liu
· 1
Xiaolin Chen
· 1
Topics
Visual QA & Reasoning
Benchmarks
Vision-Language Models
Video-Language
Image-Text Retrieval