Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Seong-Whan Lee — most-cited papers & profile · Speech Audio
← authors
·
overview
Seong-Whan Lee
36
papers ·
60
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fre-GAN: Adversarial Frequency-consistent Audio Synthesis
2021 · 7 citations
FLowHigh: Towards Efficient and High-Quality Audio Super-Resolution with Single-Step Flow Matching
2025 · 5 citations
HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
2023 · 3 citations
TranSentence: Speech-to-speech Translation via Language-agnostic Sentence-level Speech Encoding without Language-parallel Data
2024 · 3 citations
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
2025 · 2 citations
DDDM-VC: Decoupled Denoising Diffusion Models with Disentangled Representation and Prior Mixup for Verified Robust Voice Conversion
2023 · 1 citations
Diff-HierVC: Diffusion-based Hierarchical Voice Conversion with Robust Pitch Generation and Masked Prior for Zero-shot Speaker Adaptation
2023 · 1 citations
VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion
2025
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
2025
EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vector
2024
DurFlex-EVC: Duration-Flexible Emotional Voice Conversion Leveraging Discrete Representations without Text Alignment
2024
Audio Dequantization for High Fidelity Audio Generation in Flow-based Neural Vocoder
2020
Multi-SpectroGAN: High-Diversity and High-Fidelity Spectrogram Generation with Adversarial Style Combination for Speech Synthesis
2020
Reinforce-Aligner: Reinforcement Alignment Search for Robust End-to-End Text-to-Speech
2021
HierVST: Hierarchical Adaptive Zero-shot Voice Style Transfer
2023
Top co-authors
Sang-Hoon Lee
· 9
Seung-Bin Kim
· 7
Hyung-Seok Oh
· 6
Deok-Hyeon Cho
· 4
Ha-Yeong Choi
· 4
Sang-Hoon Lee
· 3
Ha-Yeong Choi
· 2
Hyeong-Rae Noh
· 2
Hyun-Wook Yoon
· 2
Byung-Kwan Ko
· 1
Dong-Min Byun
· 1
Hyunseung Chung
· 1
Topics
Audio Generation
Text-to-Speech
Voice Cloning
Speech Enhancement
Music Generation
Speech Recognition
Speech Translation
Speaker Analysis