Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yong Man Ro — most-cited papers & profile · Multimodal
← authors
·
overview
Yong Man Ro
23
papers ·
103
citations ·
42
h-index
World Vision · Korea Advanced Institute of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Unified Reinforcement and Imitation Learning for Vision-Language Models
2025
ReFoCUS: Reinforcement-guided Frame Optimization for Contextual Understanding
2025
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models
2025
What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-modal Models
2024
Top co-authors
Byung-Kwan Lee
· 2
Ryo Hachiuma
· 2
Yu-Chiang Frank Wang
· 2
Yueh-Hua Wu
· 2
Hosu Lee
· 1
Hyunjun Kim
· 1
Junho Kim
· 1
Junho Kim
· 1
Yeon Ju Kim
· 1
Topics
Vision-Language Models
Video-Language
Benchmarks
Visual QA & Reasoning