Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Usman Naseem — most-cited papers & profile · Multimodal
← authors
·
overview
Usman Naseem
20
papers ·
12
citations ·
31
h-index
Macquarie University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Mrgagents: A Multi-agent Framework For Improved Medical Report Generation With Med-lvlms
2025 · 4 citations
LLM2CLIP: Powerful Language Model Unlocks Richer Cross-Modality Representation
2024 · 2 citations
Medcfvqa: A Causal Approach To Mitigate Modality Preference Bias In Medical Visual Question Answering
2025 · 2 citations
Agentic Moderation: Multi-agent Design For Safer Vision-language Models
2025
Dual-bench: Measuring Over-refusal And Robustness In Vision-language Models
2025
Alleviating Textual Reliance In Medical Language-guided Segmentation Via Prototype-driven Semantic Approximation
2025
Intersectional Fairness In Vision-language Models For Medical Image Disease Classification
2025
Top co-authors
Shuchang Ye
· 3
Mingyuan Meng
· 2
Adam G. Dunn
· 1
Aoqi Wu
· 1
Chong Luo
· 1
Chunyu Wang
· 1
Dongdong Chen
· 1
Juan Ren
· 1
Kaixuan Ren
· 1
Liang Hu
· 1
Lili Qiu
· 1
Mark Dras
· 1
Topics
Vision-Language Models
Video-Language
Embodied & Agents
Image-Text Retrieval
Benchmarks
Audio-Visual
Visual QA & Reasoning