Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yukiya Hono — most-cited papers & profile · Speech Audio
← authors
·
overview
Yukiya Hono
12
papers ·
7
citations ·
7
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Sinsy: A Deep Neural Network-Based Singing Voice Synthesis System
2021 · 35 citations
PeriodNet: A non-autoregressive waveform generation model with a structure separating periodic and aperiodic components
2021 · 14 citations
PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model
2024 · 4 citations
Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis
2020 · 2 citations
Embedding a Differentiable Mel-cepstral Synthesis Filter to a Neural Speech Synthesis System
2022 · 1 citations
Singing Voice Synthesis Based on a Musical Note Position-Aware Attention Mechanism
2022 · 1 citations
Singing voice synthesis based on frame-level sequence-to-sequence models considering vocal timing deviation
2023 · 1 citations
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
2023 · 1 citations
Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis
2026
End-to-End Text-to-Speech Based on Latent Representation of Speaking Styles Using Spontaneous Dialogue
2022
UniFLG: Unified Facial Landmark Generator from Text or Speech
2023
PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems
2024
Top co-authors
Yoshihiko Nankaku
· 8
Kei Hashimoto
· 7
Keiichi Tokuda
· 7
Kei Sawada
· 5
Keiichiro Oura
· 4
Kentaro Mitsui
· 4
Koh Mitsuda
· 2
Shinji Takaki
· 2
Tianyu Zhao
· 2
Toshiaki Wakatsuki
· 2
and Keiichi Tokuda
· 1
Haesung Jeon
· 1
Topics
Audio Generation
Text-to-Speech
Music Generation
Speech Translation
Speech Recognition
Speech Enhancement
Multimodal Audio