Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Youngjoon Jang — most-cited papers & profile · Multimodal
← authors
·
overview
Youngjoon Jang
10
papers ·
10
citations ·
3
h-index
University of Oxford
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
FreGrad: Lightweight and Fast Frequency-aware Diffusion Vocoder
2024 · 9 citations
Seeing Through the Conversation: Audio-Visual Speech Separation based on Diffusion Model
2023 · 1 citations
LP-CFM: Perceptual Invariance-Aware Conditional Flow Matching for Speech Modeling
2025
EDNet: A Versatile Speech Enhancement Framework with Gating Mamba Mechanism and Phase Shift-Invariant Training
2025
Fork-Merge Decoding: Enhancing Multimodal Understanding in Audio-Visual Large Language Models
2025
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text
2024
VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis
2024
Topics
Audio Generation
Multimodal Audio
eess.AS
Speech Enhancement
Text-to-Speech
cs.SD
Code
Vision-Language
Model Architecture
Training Techniques