Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Manthan Thakker — most-cited papers & profile · Speech Audio
← authors
·
overview
Manthan Thakker
7
papers ·
60
citations ·
4
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation
2022 · 32 citations
E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
2024 · 25 citations
SpeechX: Neural Codec Language Model as a Versatile Speech Transformer
2023 · 3 citations
CoVoMix2: Advancing Zero-Shot Dialogue Generation with Fully Non-Autoregressive Flow Matching
2025
Total-Duration-Aware Duration Modeling for Text-to-Speech Systems
2024
An Investigation of Noise Robustness for Flow-Matching-Based Zero-Shot TTS
2024
Laugh Now Cry Later: Controlling Time-Varying Emotional States of Flow-Matching-Based Zero-Shot Text-to-Speech
2024
Top co-authors
Sefik Emre Eskimez
· 6
Naoyuki Kanda
· 5
Sheng Zhao
· 5
Jinyu Li
· 4
Canrun Li
· 3
Chung-Hsien Tsai
· 3
Hemin Yang
· 3
Min Tang
· 3
Zhen Xiao
· 3
Zirun Zhu
· 3
Haibin Wu
· 2
Takuya Yoshioka
· 2
Topics
Text-to-Speech
Audio Generation
Speech Enhancement
Speech Recognition
Speech Translation
Music Generation
Speaker Analysis