Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Aishwarya Agrawal — most-cited papers & profile · Multimodal
← authors
·
overview
Aishwarya Agrawal
12
papers ·
19
citations ·
15
h-index
Georgia Institute of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting
2022 · 7 citations
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
2023 · 5 citations
MoqaGPT : Zero-Shot Multi-modal Open-domain Question Answering with Large Language Model
2023 · 4 citations
Reassessing Evaluation Practices in Visual Question Answering: A Case Study on Out-of-Distribution Generalization
2022 · 3 citations
Discovering Failure Modes in Vision-Language Models using RL
2026
Learning What Matters: Prioritized Concept Learning Via Relative Error-driven Sample Selection
2025
Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Compositional Understanding
2023
Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison
2024
Benchmarking Vision Language Models for Cultural Understanding
2024
Top co-authors
Le Zhang
· 3
Rabiul Awal
· 3
Aida Nematzadeh
· 2
Kanishk Jain
· 2
Oscar Mañas
· 2
Shravan Nayak
· 2
Anita Gergely
· 1
Elnaz Davoodi
· 1
Emanuele Bugliarello
· 1
Fengran Mo
· 1
Ivana Kaji\'c
· 1
Jian-Yun Nie
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Video-Language
Embodied & Agents
Instruction Tuning
Image-Text Retrieval