Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shinji Watanabe — most-cited papers & profile · Speech Audio
← authors
·
overview
Shinji Watanabe
302
papers ·
1579
citations ·
74
h-index
Carnegie Mellon University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
2021 · 201 citations
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
2020 · 97 citations
ESPnet: End-to-End Speech Processing Toolkit
2018 · 74 citations
Multi-Modal Data Augmentation for End-to-End ASR
2018 · 55 citations
Listen and Fill in the Missing Letters: Non-Autoregressive Transformer for Speech Recognition
2019 · 53 citations
SUPERB: Speech processing Universal PERformance Benchmark
2021 · 51 citations
Multichannel End-to-end Speech Recognition
2017 · 46 citations
Neural Speaker Diarization with Speaker-Wise Chain Rule
2020 · 41 citations
Recent Developments on ESPnet Toolkit Boosted by Conformer
2020 · 40 citations
Branchformer: Parallel MLP-Attention Architectures to Capture Local and Global Context for Speech Recognition and Understanding
2022 · 40 citations
Searchable Hidden Intermediates for End-to-End Models of Decomposable Sequence Tasks
2021 · 34 citations
Towards Online End-to-end Transformer Automatic Speech Recognition
2019 · 31 citations
ESPnet2-TTS: Extending the Edge of TTS Research
2021 · 28 citations
ESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and Understanding
2022 · 28 citations
The Hitachi-JHU DIHARD III System: Competitive End-to-End Neural Diarization and X-Vector Clustering Systems Combined by DOVER-Lap
2021 · 27 citations
Top co-authors
Xuankai Chang
· 43
Siddhant Arora
· 36
Yifan Peng
· 30
Jiatong Shi
· 26
Jiatong Shi
· 26
Wangyou Zhang
· 24
Emiru Tsunoo
· 23
Yosuke Kashiwagi
· 23
Samuele Cornell
· 22
Jinchuan Tian
· 21
Hung-yi Lee
· 20
Jee-weon Jung
· 20
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Speech Enhancement
Speaker Analysis
Audio Generation
Music Generation
Multimodal Audio
Voice Cloning