Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Naoki Makishima — most-cited papers & profile · Speech Audio
← authors
·
overview
Naoki Makishima
9
papers ·
29
citations ·
7
h-index
NTT (Japan) · NTT Medical Center
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Hierarchical Transformer-based Large-Context End-to-end ASR with Large-Context Knowledge Distillation
2021 · 21 citations
MAPGN: Masked Pointer-generator Network For Sequence-to-sequence Pre-training
2021 · 6 citations
End-to-End Rich Transcription-Style Automatic Speech Recognition with Semi-Supervised Learning
2021 · 1 citations
Speaker consistency loss and step-wise optimization for semi-supervised joint training of TTS and ASR using unpaired text data
2022 · 1 citations
Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model
2025
Zero-Shot Joint Modeling of Multiple Spoken-Text-Style Conversion Tasks using Switching Tokens
2021
Unified Autoregressive Modeling for Joint End-to-End Multi-Talker Overlapped Speech Recognition and Speaker Attribute Estimation
2021
Cross-Modal Transformer-Based Neural Correction Models for Automatic Speech Recognition
2021
End-to-End Joint Target and Non-Target Speakers ASR
2023
Top co-authors
Ryo Masumura
· 9
Mana Ihori
· 8
Akihiko Takashima
· 7
Shota Orihashi
· 7
Tomohiro Tanaka
· 5
Satoshi Suzuki
· 3
Tomohiro Tanaka
· 3
Taiga Yamane
· 2
Takafumi Moriya
· 2
Atsushi Ando
· 1
Atsushi Ando
· 1
Daiki Okamura
· 1
Topics
Speech Recognition
Text-to-Speech
Speaker Analysis
Speech Translation
Audio Generation
Music Generation
Audio Understanding
Multimodal Audio