Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Gustav Eje Henter — most-cited papers & profile · Speech Audio
← authors
·
overview
Gustav Eje Henter
13
papers ·
92
citations ·
24
h-index
KTH Royal Institute of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis
2018 · 51 citations
Transformation of low-quality device-recorded speech to high-quality speech using improved SEGAN model
2019 · 30 citations
A Comparative Study of Self-Supervised Speech Representations in Read and Spontaneous TTS
2023 · 5 citations
Normalizing Flow based Hidden Markov Models for Classification of Speech Phones with Explainability
2021 · 4 citations
OverFlow: Putting flows on top of neural transducers for better TTS
2022 · 1 citations
Autovocoder: Fast Waveform Generation from a Learned Speech Representation using Differentiable Digital Signal Processing
2022 · 1 citations
VoXtream2: Full-stream TTS with dynamic speaking rate control
2026
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
2026
A processing framework to access large quantities of whispered speech found in ASMR
2023
Speaker-independent neural formant synthesis
2023
Diff-TTSG: Denoising probabilistic integrated speech and gesture synthesis
2023
On the Use of Self-Supervised Speech Representations in Spontaneous Speech Synthesis
2023
Unified speech and gesture synthesis using flow matching
2023
Top co-authors
\'Eva Sz\'ekely
· 6
Jonas Beskow
· 3
Shivam Mehta
· 3
Siyang Wang
· 3
Jaime Lorenzo-Trueba
· 2
Joakim Gustafson
· 2
Junichi Yamagishi
· 2
Simon Alexanderson
· 2
Xin Wang
· 2
Zofia Malisz
· 2
Ambika Kirkland
· 1
Antoine Honor\'e
· 1
Topics
Text-to-Speech
Audio Generation
Speech Enhancement
Speech Recognition
Audio Understanding
Multimodal Audio
Voice Cloning
Speaker Analysis
Music Generation