Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Junhyeok Lee — most-cited papers & profile · Speech Audio
← authors
·
overview
Junhyeok Lee
20
papers ·
25
citations ·
0
h-index
Johns Hopkins University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech
2025 · 15 citations
Assem-VC: Realistic Voice Conversion by Assembling Modern Speech Synthesis Techniques
2021 · 3 citations
Super Monotonic Alignment Search
2024 · 1 citations
PITS: Variational Pitch Inference without Fundamental Frequency for End-to-End Pitch-controllable TTS
2023 · 1 citations
Reconstruct! Don't Encode: Self-Supervised Representation Reconstruction Loss for High-Intelligibility and Low-Latency Streaming Neural Audio Codec
2026
MaskVCT: Masked Voice Codec Transformer for Zero-Shot Voice Conversion With Increased Controllability via Multiple Guidances
2025
Improving Test-Time Performance of RVQ-based Neural Codecs
2025
Controllable and Interpretable Singing Voice Decomposition via Assem-VC
2021
Talking Face Generation with Multilingual TTS
2022
NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates
2022
SANE-TTS: Stable And Natural End-to-End Multilingual Text-to-Speech
2022
PhaseAug: A Differentiable Augmentation for Speech Synthesis to Simulate One-to-Many Mapping
2022
VIFS: An End-to-End Variational Inference for Foley Sound Synthesis
2023
JenGAN: Stacked Shifted Filters in GAN-Based Speech Synthesis
2024
DualSpeech: Enhancing Speaker-Fidelity and Text-Intelligibility Through Dual Classifier-Free Guidance
2024
Top co-authors
Shrikanth Narayanan
· 2
Dading Chong
· 1
Dongchao Yang
· 1
Hyoung-Kyu Song
· 1
Zengyi Qin
· 1
Topics
Audio Generation
Text-to-Speech
Speech Enhancement
Voice Cloning
eess.AS
cs.AI
Music Generation
Speech Recognition
Speech Translation