Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Matthew Wiesner — most-cited papers & profile · Speech Audio
← authors
·
overview
Matthew Wiesner
20
papers ·
154
citations ·
1
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
ESPnet: End-to-End Speech Processing Toolkit
2018 · 74 citations
Multi-Modal Data Augmentation for End-to-End ASR
2018 · 55 citations
The CHiME-7 DASR Challenge: Distant Meeting Transcription with Multiple Devices in Diverse Scenarios
2023 · 9 citations
Automatic Speech Recognition and Topic Identification for Almost-Zero-Resource Languages
2018 · 6 citations
Multilingual sequence-to-sequence speech recognition: architecture, transfer learning, and language modeling
2018 · 5 citations
Topic Identification for Speech without ASR
2017 · 3 citations
Recent Trends in Distant Conversational Speech Recognition: A Review of CHiME-7 and 8 DASR Challenges
2025 · 1 citations
GenVC: Self-Supervised Zero-Shot Voice Conversion
2025 · 1 citations
WST: Weakly Supervised Transducer for Automatic Speech Recognition
2025
CS-FLEURS: A Massively Multilingual and Code-Switched Speech Dataset
2025
Whisper-UT: A Unified Translation Framework for Speech and Text
2025
Scalable Controllable Accented TTS
2025
HENT-SRT: Hierarchical Efficient Neural Transducer with Self-Distillation for Joint Speech Recognition and Translation
2025
Pretraining by Backtranslation for End-to-end ASR in Low-Resource Settings
2018
Towards Zero-Shot Code-Switched Speech Recognition
2022
Top co-authors
Sanjeev Khudanpur
· 15
Shinji Watanabe
· 9
Kevin Duh
· 4
Leibny Paola Garcia
· 4
Adithya Renduchintala
· 3
Amir Hussein
· 3
Chunxi Liu
· 3
Cihan Xiao
· 3
Daniel Povey
· 3
Desh Raj
· 3
Dongji Gao
· 3
Henry Li Xinyuan
· 3
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Speaker Analysis
Audio Understanding
Audio Generation
Multimodal Audio
Speech Enhancement
Voice Cloning
Music Generation