Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuma Shirahata — most-cited papers & profile · Speech Audio
← authors
·
overview
Yuma Shirahata
13
papers ·
18
citations ·
2
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Cross-Speaker Emotion Transfer for Low-Resource Text-to-Speech Using Non-Parallel Voice Conversion with Pitch-Shift Data Augmentation
2022 · 16 citations
PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-to-Speech Using Natural Language Descriptions
2023 · 1 citations
Universal Score-based Speech Enhancement with High Content Preservation
2024 · 1 citations
Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech Generation from SSL features
2026
CC-G2PnP: Streaming Grapheme-to-Phoneme and prosody with Conformer-CTC for unsegmented languages
2026
SLASH: Self-Supervised Speech Pitch Estimation Leveraging DSP-derived Absolute Pitch
2025
BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing
2025
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
2025
Period VITS: Variational Inference with Explicit Pitch Modeling for End-to-end Emotional Speech Synthesis
2022
Lightweight and High-Fidelity End-to-End Text-to-Speech with Multi-Band Generation and Inverse Short-Time Fourier Transform
2022
LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning
2024
Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
2024
Description-based Controllable Text-to-Speech with Cross-Lingual Voice Control
2024
Top co-authors
Ryuichi Yamamoto
· 10
Kentaro Tachibana
· 7
Masaya Kawamura
· 7
Ryo Terashima
· 3
Byeongseon Park
· 2
Eunwoo Song
· 2
Hien Ohnaka
· 2
Jae-Min Kim
· 2
Takuya Hasumi
· 2
Tatsuya Komatsu
· 2
Hironori Doi
· 1
Hyun-Wook Yoon
· 1
Topics
Text-to-Speech
Audio Generation
Voice Cloning
Speech Recognition
Music Generation
Speech Enhancement
Speaker Analysis
Speech Translation
Audio Understanding