Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Mohan Kankanhalli — most-cited papers & profile · Multimodal
← authors
·
overview
Mohan Kankanhalli
26
papers ·
268
citations ·
65
h-index
National University of Singapore
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
What Makes for Good Visual Tokenizers for Large Language Models?
2023 · 5 citations
ELIP: Efficient Discriminative Language-Image Pre-training with Fewer Vision Tokens
2023 · 1 citations
On Modality Bias Recognition and Reduction
2022
A Unified End-to-End Retriever-Reader Framework for Knowledge-based VQA
2022
UNK-VQA: A Dataset and a Probe into the Abstention Ability of Multi-modal Large Models
2023
Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models
2023
Do Vision-Language Transformers Exhibit Visual Commonsense? An Empirical Study of VCR
2024
Top co-authors
Liqiang Nie
· 6
Yangyang Guo
· 4
Yongkang Wong
· 2
Zhiyong Cheng
· 2
Alberto Del Bimbo
· 1
Fangkai Jiao
· 1
Fan Liu
· 1
Guangzhi Wang
· 1
Haoyu Zhang
· 1
Harry Cheng
· 1
Kejie Wang
· 1
Xiaohan Ding
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language
Image-Text Retrieval
cs.CV