Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sushant Gautam — most-cited papers & profile · Multimodal
← authors
·
overview
Sushant Gautam
2
papers ·
4
citations ·
7
h-index
Agriculture and Forestry University · Simula Metropolitan Center for Digital Engineering
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Kvasir-vqa-x1: A Multimodal Dataset For Medical Reasoning And Robust Medvqa In Gastrointestinal Endoscopy
2025 · 3 citations
Point, Detect, Count: Multi-task Medical Image Understanding With Instruction-tuned Vision-language Models
2025 · 1 citations
Top co-authors
Michael A. Riegler
· 2
Pål Halvorsen
· 2
Topics
Visual QA & Reasoning
Benchmarks
Vision-Language Models
Instruction Tuning
Video-Language