Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yong Man Ro — most-cited papers & profile · Multimodal
← authors
·
overview
Yong Man Ro
3
papers ·
15
citations ·
42
h-index
World Vision · Korea Advanced Institute of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Zero-AVSR: Zero-Shot Audio-Visual Speech Recognition with LLMs by Learning Language-Agnostic Speech Representations
2025 · 9 citations
Mitigating Adversarial Vulnerability Through Causal Parameter Estimation By Adversarial Double Machine Learning
2023 · 6 citations
TroL: Traversal of Layers for Large Language and Vision Models
2024
Topics
Speech Recognition
Speech Translation
Speech Enhancement
Multimodal Audio
Adversarial ML
Vulnerability Detection
Vision-Language
Model Architecture
Efficiency
Training Techniques