Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Naoyuki Kanda — most-cited papers & profile · Speech Audio
← authors
·
overview
Naoyuki Kanda
56
papers ·
329
citations ·
31
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
2020 · 97 citations
Serialized Output Training for End-to-End Overlapped Speech Recognition
2020 · 61 citations
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition
2021 · 43 citations
E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
2024 · 25 citations
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition
2021 · 14 citations
Guided Source Separation Meets a Strong ASR Backend: Hitachi/Paderborn University Joint Investigation for Dinner Party ASR
2019 · 12 citations
End-to-End Neural Speaker Diarization with Self-attention
2019 · 12 citations
End-to-End Neural Speaker Diarization with Permutation-Free Objectives
2019 · 10 citations
Speech-language Pre-training for End-to-end Spoken Language Understanding
2021 · 9 citations
Simultaneous Speech Recognition and Speaker Diarization for Monaural Dialogue Recordings with Target-Speaker Acoustic Models
2019 · 6 citations
Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers
2020 · 5 citations
Auxiliary Interference Speaker Loss for Target-Speaker Speech Recognition
2019 · 3 citations
Investigation of End-To-End Speaker-Attributed ASR for Continuous Multi-Talker Recordings
2020 · 3 citations
Large-Scale Pre-Training of End-to-End Multi-Talker ASR for Meeting Transcription with Single Distant Microphone
2021 · 3 citations
Transcribe-to-Diarize: Neural Speaker Diarization for Unlimited Number of Speakers using End-to-End Speaker-Attributed ASR
2021 · 3 citations
Top co-authors
Takuya Yoshioka
· 31
Jinyu Li
· 22
Yashesh Gaur
· 17
Zhong Meng
· 17
Zhuo Chen
· 13
Sefik Emre Eskimez
· 10
Xiaofei Wang
· 9
Jian Wu
· 8
Xiong Xiao
· 8
Xiaofei Wang
· 7
Xiaofei Wang
· 7
Dongmei Wang
· 6
Topics
Speech Recognition
Speech Translation
Audio Understanding
Speaker Analysis
Speech Enhancement
Text-to-Speech
Audio Generation
Multimodal Audio
math.IT