Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Takuya Higuchi — most-cited papers & profile · Speech Audio
← authors
·
overview
Takuya Higuchi
11
papers ·
1
citations ·
18
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study
2023 · 1 citations
Which Data Matter? Embedding-Based Data Selection for Speech Recognition
2026
A Variational Framework for Improving Naturalness in Generative Spoken Language Models
2025
Speaker-IPL: Unsupervised Learning of Speaker Characteristics with i-Vector based Pseudo-Labels
2024
Multi-task Learning with Cross Attention for Keyword Spotting
2021
Improving Voice Trigger Detection with Metric Learning
2022
Multichannel Voice Trigger Detection Based on Transform-average-concatenate
2023
ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models
2024
Can you Remove the Downstream Model for Speaker Recognition with Self-Supervised Speech Features?
2024
Towards Automatic Assessment of Self-Supervised Speech Models using Rank
2024
Exploring Prediction Targets in Masked Pre-Training for Speech Foundation Models
2024
Top co-authors
Zakaria Aldeneh
· 7
Barry-John Theobald
· 6
Ahmed Hussen Abdelaziz
· 5
Shinji Watanabe
· 5
Stephen Shum
· 5
Tatiana Likhomanenko
· 5
Jee-weon Jung
· 4
Li-Wei Chen
· 3
Alexander Rudnicky
· 2
Anmol Gupta
· 2
Avamarie Brueggeman
· 2
Chandra Dhir
· 2
Topics
Audio Understanding
Speech Recognition
Speaker Analysis
Speech Enhancement
cs.SD
Audio Generation
Text-to-Speech
Voice Cloning
Speech Translation