Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hiroshi Saruwatari — most-cited papers & profile · Speech Audio
← authors
·
overview
Hiroshi Saruwatari
58
papers ·
477
citations ·
39
h-index
The University of Tokyo
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Statistical Parametric Speech Synthesis Incorporating Generative Adversarial Networks
2017 · 228 citations
JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis
2017 · 88 citations
JVS corpus: free Japanese multi-speaker voice corpus
2019 · 41 citations
Time-Domain Audio Source Separation Based on Wave-U-Net Combined with Discrete Wavelet Transform
2020 · 23 citations
The T05 System for The VoiceMOS Challenge 2024: Transfer Learning from Deep Image Classifier to Naturalness MOS Prediction of High-Quality Synthetic Speech
2024 · 22 citations
Sampling-based speech parameter generation using moment-matching networks
2017 · 11 citations
SelfRemaster: Self-Supervised Speech Restoration with Analysis-by-Synthesis Approach Using Channel Modeling
2022 · 9 citations
Utterance-level Sequential Modeling For Deep Gaussian Process Based Speech Synthesis Using Simple Recurrent Unit
2020 · 6 citations
DNN-based Speaker Embedding Using Subjective Inter-speaker Similarity for Multi-speaker Modeling in Speech Synthesis
2019 · 5 citations
J-MAC: Japanese multi-speaker audiobook corpus for speech synthesis
2022 · 5 citations
Coco-Nut: Corpus of Japanese Utterance and Voice Characteristics Description for Prompt-based Control
2023 · 5 citations
Human-in-the-loop Speaker Adaptation for DNN-based Multi-speaker TTS
2022 · 4 citations
Phase reconstruction from amplitude spectrograms based on von-Mises-distribution deep neural network
2018 · 3 citations
V2S attack: building DNN-based voice conversion from automatic speaker verification
2019 · 3 citations
Lifter Training and Sub-band Modeling for Computationally Efficient and High-Quality Voice Conversion Using Spectral Differentials
2020 · 3 citations
Top co-authors
Shinnosuke Takamichi
· 35
Detai Xin
· 9
Dong Yang
· 4
Shinji Watanabe
· 2
Xu Tan
· 2
Ashish Kulkarni
· 1
David M. Chan
· 1
Dongchao Yang
· 1
Hisashi Kawai
· 1
Jianing Yang
· 1
Jinyu Li
· 1
Kai Shen
· 1
Topics
Audio Generation
Text-to-Speech
Speech Enhancement
Speaker Analysis
Speech Recognition
Voice Cloning
Audio Understanding
Music Generation
eess.AS
cs.SD