Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Joon Son Chung — most-cited papers & profile · Large Language Models
← authors
·
overview
Joon Son Chung
15
papers ·
2311
citations ·
30
h-index
Korea Advanced Institute of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
VoxCeleb2: Deep Speaker Recognition
2018 · 2259 citations
The Conversation: Deep Audio-Visual Speech Enhancement
2018 · 46 citations
Disentangled Speech Embeddings using Cross-modal Self-supervision
2020 · 5 citations
Hearing And Seeing Through CLIP: A Framework For Self-supervised Sound Source Localization
2025 · 1 citations
Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions
2026
Video Diffusion Models Excel at Tracking Similar-Looking Objects Without Supervision
2025
SCORE: Scaling audio generation using Standardized COmposite REwards
2025
MAGE: A Coarse-to-Fine Speech Enhancer with Masked Generative Model
2025
SPADE: Structured Pruning and Adaptive Distillation for Efficient LLM-TTS
2025
InfiniteAudio: Infinite-Length Audio Generation with Consistency
2025
EDNet: A Versatile Speech Enhancement Framework with Gating Mamba Mechanism and Phase Shift-Invariant Training
2025
Seeing voices and hearing voices: learning discriminative embeddings using cross-modal self-supervision
2020
FaceFilter: Audio-visual speech separation using still images
2020
Self-Supervised Learning of Audio-Visual Objects from Video
2020
Adapting Speaker Embeddings for Speaker Diarisation
2021
Topics
Multimodal Audio
Speech Recognition
Audio Generation
Speech Enhancement
Speaker Analysis
Tracking
Video Understanding
Music Generation
Vision-Language Models
Audio-Visual