Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ron Litman — most-cited papers & profile · Multimodal
← authors
·
overview
Ron Litman
5
papers ·
12
citations ·
11
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Multimodal Semi-Supervised Learning for Text Recognition
2022 · 10 citations
LaTr: Layout-Aware Transformer for Scene-Text VQA
2021 · 2 citations
Towards Models that Can See and Read
2023
Question Aware Vision Transformer for Multimodal Reasoning
2024
TAP-VL: Text Layout-Aware Pre-training for Enriched Vision-Language Models
2024
Top co-authors
Aviad Aberdam
· 4
Roy Ganz
· 4
Oren Nuriel
· 3
Shai Mazor
· 3
Yair Kittenplon
· 3
Elad Ben Avraham
· 2
Ali Furkan Biten
· 1
Jonathan Fhima
· 1
R. Manmatha
· 1
Srikar Appalaraju
· 1
Yusheng Xie
· 1
Topics
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Image-Text Retrieval
Video-Language