Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Timo Gerkmann — most-cited papers & profile · Speech Audio
← authors
·
overview
Timo Gerkmann
36
papers ·
46
citations ·
31
h-index
Universität Hamburg · Signal Processing (United States) · Hamburg University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Efficient Transformer-based Speech Enhancement Using Long Frames and STFT Magnitudes
2022 · 20 citations
Speech Signal Improvement Using Causal Generative Diffusion Models
2023 · 5 citations
Speech Enhancement and Dereverberation with Diffusion-based Generative Models
2022 · 4 citations
Speech Enhancement with Score-Based Generative Models in the Complex STFT Domain
2022 · 3 citations
Analysing Diffusion-based Generative Approaches versus Discriminative Approaches for Speech Restoration
2022 · 3 citations
Audio-Visual Speech Enhancement with Score-Based Generative Models
2023 · 3 citations
Robustness of Speech Separation Models for Similar-pitch Speakers
2024 · 3 citations
Normalized Features for Improving the Generalization of DNN Based Speech Enhancement
2017 · 1 citations
Phase-Aware Deep Speech Enhancement: It's All About The Frame Length
2022 · 1 citations
Uncertainty Estimation in Deep Speech Enhancement Using Complex Gaussian Mixture Models
2022 · 1 citations
In-the-wild Speech Emotion Conversion Using Disentangled Self-Supervised Representations and Neural Vocoder-based Resynthesis
2023 · 1 citations
Real-Time Streamable Generative Speech Restoration with Flow Matching
2025
Diffusion Buffer: Online Diffusion-based Speech Enhancement with Sub-Second Latency
2025
LipDiffuser: Lip-to-Speech Generation with Conditional Diffusion Models
2025
FlowDec: A flow-based full-band general audio codec with high perceptual quality
2025
Top co-authors
Simon Welker
· 14
Julius Richter
· 12
Bunlong Lay
· 10
Tal Peer
· 9
Jean-Marie Lemercier
· 8
Danilo de Oliveira
· 5
Kristina Tesch
· 4
Alexander Richard
· 2
Nale Lehmann‐Willenbrock
· 2
Navin Raj Prabhu
· 2
Yi-Chiao Wu
· 2
Alina Mannanova
· 1
Topics
Speech Enhancement
Audio Understanding
Speech Recognition
Audio Generation
Speech Translation
Speaker Analysis
Music Generation
Text-to-Speech
Multimodal Audio