Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Masato Mimura — most-cited papers & profile · Speech Audio
← authors
·
overview
Masato Mimura
17
papers ·
47
citations ·
19
h-index
NTT (Japan) · NTT Medical Center
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Distilling the Knowledge of BERT for Sequence-to-Sequence ASR
2020 · 40 citations
Distilling the Knowledge of BERT for CTC-based ASR
2022 · 6 citations
Speech Corpus of Ainu Folklore and End-to-end Speech Recognition for Ainu Language
2020 · 1 citations
Chunkwise Aligners for Streaming Speech Recognition
2026
Decoder-only Conformer with Modality-aware Sparse Mixtures of Experts for ASR
2026
Microphone Array Geometry Independent Multi-Talker Distant ASR: NTT System for the DASR Task of the CHiME-8 Challenge
2025
Improving OOV Detection and Resolution with External Language Models in Acoustic-to-Word ASR
2019
Generative Adversarial Training Data Adaptation for Very Low-resource Automatic Speech Recognition
2020
End-to-end Music-mixed Speech Recognition
2020
ASR Rescoring and Confidence Estimation with ELECTRA
2021
Non-autoregressive Error Correction for CTC-based ASR with Phone-conditioned Masked LM
2022
Time-domain Speech Enhancement Assisted by Multi-resolution Frequency Encoder and Decoder
2023
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
2024
Sentence-wise Speech Summarization: Task, Datasets, and End-to-End Modeling with LM Knowledge Distillation
2024
NTT Multi-Speaker ASR System for the DASR Task of CHiME-8 Challenge
2024
Top co-authors
Tatsuya Kawahara
· 9
Shinsuke Sakai
· 7
Takafumi Moriya
· 6
Takanori Ashihara
· 6
Hirofumi Inaguma
· 5
Marc Delcroix
· 5
Hayato Futami
· 4
Kohei Matsuura
· 4
Atsunori Ogawa
· 3
Hiroshi Sato
· 3
Kohei Matsuura
· 3
Sei Ueno
· 3
Topics
Speech Recognition
Speech Translation
Speech Enhancement
Speaker Analysis
Audio Understanding
Multimodal Audio
Audio Generation
Music Generation
Text-to-Speech