Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Seung-Bin Kim — most-cited papers & profile · Speech Audio
← authors
·
overview
Seung-Bin Kim
14
papers ·
23
citations ·
7
h-index
Korea University · Myongji University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improved RawNet with Feature Map Scaling for Text-independent Speaker Verification using Raw Waveforms
2020 · 7 citations
FLowHigh: Towards Efficient and High-Quality Audio Super-Resolution with Single-Step Flow Matching
2025 · 5 citations
HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
2023 · 3 citations
TranSentence: Speech-to-speech Translation via Language-agnostic Sentence-level Speech Encoding without Language-parallel Data
2024 · 3 citations
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
2025 · 2 citations
DiEmo-TTS: Disentangled Emotion Representations via Self-Supervised Distillation for Cross-Speaker Emotion Transfer in Text-to-Speech
2025 · 1 citations
EmoSphere-SER: Enhancing Speech Emotion Recognition Through Spherical Representation with Auxiliary Classification
2025 · 1 citations
Segment Aggregation for short utterances speaker verification using raw waveforms
2020 · 1 citations
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment
2026
Toward Complex-Valued Neural Networks for Waveform Generation
2026
Affectron: Emotional Speech Synthesis with Affective and Contextually Aligned Nonverbal Vocalizations
2026
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
2025
EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vector
2024
EmoSphere-TTS: Emotional Style and Intensity Modeling via Spherical Emotion Vector for Controllable Emotional Text-to-Speech
2024
Top co-authors
Seong-Whan Lee
· 11
Hyung-Seok Oh
· 7
Deok-Hyeon Cho
· 6
Sang-Hoon Lee
· 3
Hye-jin Shim
· 2
Jee-weon Jung
· 2
Ju-ho Kim
· 2
Jun-Hak Yun
· 2
and Ha-Jin Yu
· 1
and Seong-Whan Lee
· 1
Deok–Hyeon Cho
· 1
Ha-Jin Yu
· 1
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Speaker Analysis
Voice Cloning
Audio Understanding
Speech Translation
Multimodal Audio
Music Generation