Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sefik Emre Eskimez — most-cited papers & profile · Speech Audio
← authors
·
overview
Sefik Emre Eskimez
25
papers ·
123
citations ·
19
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation
2022 · 32 citations
E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
2024 · 25 citations
Human Listening and Live Captioning: Multi-Task Training for Speech Enhancement
2021 · 24 citations
Improving Readability for Automatic Speech Recognition Transcription
2020 · 19 citations
Federated Transfer Learning with Dynamic Gradient Aggregation
2020 · 8 citations
SpeechX: Neural Codec Language Model as a Versatile Speech Transformer
2023 · 3 citations
Real-Time Joint Personalized Speech Enhancement and Acoustic Echo Cancellation
2022 · 2 citations
Neural Speech Extraction with Human Feedback
2025 · 1 citations
Personalized Speech Enhancement: New Models and Comprehensive Evaluation
2021 · 1 citations
Dynamic Gradient Aggregation for Federated Domain Adaptation
2021
All-neural beamformer for continuous speech separation
2021
One model to enhance them all: array geometry agnostic multi-channel personalized speech enhancement
2021
Separating Long-Form Speech with Group-Wise Permutation Invariant Training
2021
Sequence-level self-learning with multiple hypotheses
2021
Leveraging Real Conversational Data for Multi-Channel Continuous Speech Separation
2022
Top co-authors
Naoyuki Kanda
· 10
Takuya Yoshioka
· 10
Jinyu Li
· 7
Manthan Thakker
· 6
Zhuo Chen
· 6
Huaming Wang
· 5
Xiaofei Wang
· 5
Min Tang
· 4
Sheng Zhao
· 4
Zirun Zhu
· 4
Canrun Li
· 3
Chung-Hsien Tsai
· 3
Topics
Speech Enhancement
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Audio Generation
Speaker Analysis
Multimodal Audio
math.IT