Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yonatan Bisk — most-cited papers & profile · Multimodal
← authors
·
overview
Yonatan Bisk
16
papers ·
143
citations ·
32
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Retrospectives on the Embodied AI Workshop
2022 · 19 citations
TACo: Token-aware Cascade Contrastive Learning for Video-Text Alignment
2021 · 4 citations
Worst of Both Worlds: Biases Compound in Pre-trained Vision-and-Language Models
2021 · 2 citations
Casper: Inferring Diverse Intents For Assistive Teleoperation With Vision Language Models
2025 · 1 citations
MAEA: Multimodal Attribution for Embodied AI
2023 · 1 citations
REM: Evaluating LLM Embodied Spatial Reasoning Through Multi-frame Trajectories
2025
MM-SeR: Multimodal Self-Refinement for Lightweight Image Captioning
2025
On Advances in Text Generation from Images Beyond Captioning: A Case Study in Self-Rationalization
2022
Top co-authors
Akshita Bhagia
· 1
Alan W. Black
· 1
Alexander Toshev
· 1
Ali Farhadi
· 1
Ana Marasović
· 1
Andrew Szot
· 1
Angel X. Chang
· 1
Aniruddha Kembhavi
· 1
Anthony Francis
· 1
Ben Talbot
· 1
Changan Chen
· 1
Chengshu Li
· 1
Topics
Vision-Language Models
Embodied & Agents
Visual QA & Reasoning
Video-Language
Benchmarks
Image-Text Retrieval