Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sarah Erfani — most-cited papers & profile · Multimodal
← authors
·
overview
Sarah Erfani
14
papers ·
54
citations ·
27
h-index
The University of Melbourne
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
2024 · 3 citations
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
2026
Investigating The Functional Roles Of Attention Heads In Vision Language Models: Evidence For Reasoning Modules
2025
Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis
2025
Top co-authors
James Bailey
· 3
Jey Han Lau
· 2
Krista A. Ehinger
· 2
Xueqi Ma
· 2
Yanbei Jiang
· 2
Christopher Leckie
· 1
Feng Liu
· 1
Hanxun Huang
· 1
Haopeng Li
· 1
Jinhao Li
· 1
Kaiyuan Cui
· 1
Lei Feng
· 1
Topics
Vision-Language Models
Video-Language
Visual QA & Reasoning
Benchmarks
cs.CV
cs.LG