Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yosuke Higuchi — most-cited papers & profile · Speech Audio
← authors
·
overview
Yosuke Higuchi
25
papers ·
106
citations ·
14
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Recent Developments on ESPnet Toolkit Boosted by Conformer
2020 · 40 citations
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models
2022 · 24 citations
A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation
2021 · 7 citations
Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict
2020 · 6 citations
Improved Mask-CTC for Non-Autoregressive End-to-End ASR
2020 · 6 citations
The 2020 ESPnet update: new features, broadened applications, performance improvements, and future plans
2020 · 6 citations
Momentum Pseudo-Labeling for Semi-Supervised Speech Recognition
2021 · 4 citations
Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
2021 · 4 citations
Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
2020 · 3 citations
Hierarchical Conditional End-to-End ASR with CTC and Multi-Granular Subword Units
2021 · 3 citations
Advancing Momentum Pseudo-Labeling with Conformer and Initialization Strategy
2021 · 1 citations
CTC Alignments Improve Autoregressive Translation
2022 · 1 citations
Predictive Speech Recognition and End-of-Utterance Detection Towards Spoken Dialog Systems
2024 · 1 citations
SpidR: Learning Fast and Stable Linguistic Units for Spoken Language Models Without Supervision
2025
SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation
2025
Top co-authors
Shinji Watanabe
· 15
Tetsunori Kobayashi
· 11
Hirofumi Inaguma
· 6
Tetsuji Ogawa
· 5
Tetsuji Ogawa
· 5
Takaaki Hori
· 3
Xuankai Chang
· 3
Chenda Li
· 2
Florian Boyer
· 2
Huaibo Zhao
· 2
Jiayi Shen
· 2
Jing Shi
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement