Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hirofumi Inaguma — most-cited papers & profile · Speech Audio
← authors
·
overview
Hirofumi Inaguma
35
papers ·
211
citations ·
21
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Seamless: Multilingual Expressive and Streaming Speech Translation
2023 · 41 citations
Distilling the Knowledge of BERT for Sequence-to-Sequence ASR
2020 · 40 citations
Recent Developments on ESPnet Toolkit Boosted by Conformer
2020 · 40 citations
SeamlessM4T: Massively Multilingual & Multimodal Machine Translation
2023 · 13 citations
ESPnet-ST: All-in-One Speech Translation Toolkit
2020 · 9 citations
A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation
2021 · 7 citations
Improved Mask-CTC for Non-Autoregressive End-to-End ASR
2020 · 6 citations
The 2020 ESPnet update: new features, broadened applications, performance improvements, and future plans
2020 · 6 citations
Distilling the Knowledge of BERT for CTC-based ASR
2022 · 6 citations
Multilingual End-to-End Speech Translation
2019 · 5 citations
Source and Target Bidirectional Knowledge Distillation for End-to-end Speech Translation
2021 · 5 citations
Speech-to-Speech Translation For A Real-world Unwritten Language
2022 · 5 citations
Efficient Monotonic Multihead Attention
2023 · 5 citations
End-to-end speech-to-dialog-act recognition
2020 · 4 citations
Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
2021 · 4 citations
Top co-authors
Shinji Watanabe
· 16
Tatsuya Kawahara
· 12
Juan Pino
· 9
Xutai Ma
· 7
Changhan Wang
· 6
Ilia Kulikov
· 6
Paden Tomasello
· 6
Yosuke Higuchi
· 6
Yun Tang
· 6
Anna Sun
· 5
Ann Lee
· 5
Hongyu Gong
· 5
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio
Speech Enhancement
Audio Generation
Speaker Analysis