Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kwangyoun Kim — most-cited papers & profile · Speech Audio
← authors
·
overview
Kwangyoun Kim
16
papers ·
65
citations ·
12
h-index
NetApp (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
end-to-end training of a large vocabulary end-to-end speech recognition system
2019 · 35 citations
Small energy masking for improved neural network training for end-to-end speech recognition
2020 · 7 citations
E-Branchformer: Branchformer with Enhanced merging for speech recognition
2022 · 6 citations
Multi-mode Transformer Transducer with Stochastic Future Context
2021 · 5 citations
Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
2021 · 5 citations
Wav2Seq: Pre-training Speech-to-Text Encoder-Decoder Models Using Pseudo Languages
2022 · 5 citations
Structured Pruning of Self-Supervised Pre-trained Models for Speech Recognition and Understanding
2023 · 1 citations
Improving ASR Contextual Biasing with Guided Attention
2024 · 1 citations
Attention based on-device streaming speech recognition with large speech corpus
2020
SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition
2021
Context-aware Fine-tuning of Self-supervised Speech Models
2022
A Comparative Study on E-Branchformer vs Conformer in Speech Recognition, Translation, and Understanding Tasks
2023
Generative Context-aware Fine-tuning of Self-supervised Speech Models
2023
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
2024
Sample-Efficient Diffusion for Text-To-Speech Synthesis
2024
Top co-authors
Shinji Watanabe
· 10
Felix Wu
· 8
Prashant Sridhar
· 8
Suwon Shon
· 5
Chanwoo Kim
· 3
Jing Pan
· 3
Karen Livescu
· 3
Kilian Q. Weinberger
· 3
Kyu Han
· 3
Yifan Peng
· 3
Dhananjaya Gowda
· 2
Jiyang Tang
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio
Audio Generation