Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Takanori Ashihara — most-cited papers & profile · Speech Audio
← authors
·
overview
Takanori Ashihara
29
papers ·
54
citations ·
9
h-index
NTT (Japan) · NTT Medical Center
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
What Do Self-Supervised Speech and Speaker Models Learn? New Findings From a Cross Model Layer-Wise Analysis
2024 · 10 citations
Leveraging Large Text Corpora for End-to-End Speech Summarization
2023 · 9 citations
Zero-shot text-to-speech synthesis conditioned using self-supervised speech representation model
2023 · 9 citations
Noise-robust zero-shot text-to-speech synthesis conditioned on self-supervised speech-representation model with adapters
2024 · 9 citations
Improving Scheduled Sampling for Neural Transducer-based ASR
2023 · 5 citations
End-to-End Automatic Speech Recognition with Deep Mutual Learning
2021 · 2 citations
Exploration of Language Dependency for Japanese Self-Supervised Speech Representation Models
2023 · 2 citations
Applying LLMs for Rescoring N-best ASR Hypotheses of Casual Conversations: Effects of Domain Adaptation and Context Carry-over
2024 · 2 citations
Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models
2022 · 1 citations
Knowledge Distillation for Neural Transducer-based Target-Speaker ASR: Exploiting Parallel Mixture/Single-Talker Speech Data
2023 · 1 citations
Transfer Learning from Pre-trained Language Models Improves End-to-End Speech Summarization
2023 · 1 citations
SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
2023 · 1 citations
Recursive Attentive Pooling for Extracting Speaker Embeddings from Multi-Speaker Recordings
2024 · 1 citations
Guided Speaker Embedding
2024 · 1 citations
Frontend Token Enhancement for Token-Based Speech Recognition
2026
Top co-authors
Takafumi Moriya
· 22
Marc Delcroix
· 21
Kohei Matsuura
· 13
Ryo Masumura
· 8
Tomohiro Tanaka
· 8
Tsubasa Ochiai
· 8
Atsunori Ogawa
· 7
Naohiro Tawara
· 7
Hiroshi Sato
· 6
Masato Mimura
· 6
Shota Horiguchi
· 6
Taichi Asami
· 6
Topics
Speech Recognition
Audio Understanding
Speaker Analysis
Speech Enhancement
Text-to-Speech
Speech Translation
Audio Generation
Multimodal Audio
Voice Cloning