Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tomoki Koriyama — most-cited papers & profile · Speech Audio
← authors
·
overview
Tomoki Koriyama
12
papers ·
63
citations ·
12
h-index
CyberAgent (Japan)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
JVS corpus: free Japanese multi-speaker voice corpus
2019 · 41 citations
Sampling-based speech parameter generation using moment-matching networks
2017 · 11 citations
Utterance-level Sequential Modeling For Deep Gaussian Process Based Speech Synthesis Using Simple Recurrent Unit
2020 · 6 citations
Multi-speaker Text-to-speech Synthesis Using Deep Gaussian Processes
2020 · 3 citations
Generative Moment Matching Network-based Random Modulation Post-filter for DNN-based Singing Voice Synthesis and Neural Double-tracking
2019 · 1 citations
Structured State Space Decoder for Speech Recognition and Synthesis
2022 · 1 citations
Instantaneous Pitch Estimation via Wave-U-Net-Based Fundamental Waveform Enhancement
2026
Speaker-Conditioned Phrase Break Prediction for Text-to-Speech with Phoneme-Level Pre-trained Language Model
2025
Duration-aware pause insertion using pre-trained language model for multi-speaker text-to-speech
2023
Frame-Wise Breath Detection with Self-Training: An Exploration of Enhancing Breath Naturalness in Text-to-Speech
2024
An Attribute Interpolation Method in Speech Synthesis by Model Merging
2024
VAE-based Phoneme Alignment Using Gradient Annealing and SSL Acoustic Features
2024
Top co-authors
Hiroshi Saruwatari
· 7
Yuki Saito
· 4
Dong Yang
· 3
Shinnosuke Takamichi
· 3
Detai Xin
· 2
Kentaro Mitsui
· 2
Koichi Miyazaki
· 2
Masato Murata
· 2
Takaaki Saeki
· 2
Hiroki Tamaru
· 1
Junya Koguchi
· 1
Naoko Tanji
· 1
Topics
Text-to-Speech
Audio Generation
Speaker Analysis
Speech Enhancement
Voice Cloning
Audio Understanding
Music Generation
Speech Recognition
Speech Translation