Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yingming Gao — most-cited papers & profile · Speech Audio
← authors
·
overview
Yingming Gao
9
papers ·
7
citations ·
11
h-index
Beijing University of Posts and Telecommunications · Chinese Academy of Medical Sciences & Peking Union Medical College
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
M2-CTTS: End-to-End Multi-scale Multi-modal Conversational Text-to-Speech Synthesis
2023 · 3 citations
Formant Tracking Using Dilated Convolutional Networks Through Dense Connection with Gating Mechanism
2020 · 1 citations
CONCSS: Contrastive-based Context Comprehension for Dialogue-appropriate Prosody in Conversational Speech Synthesis
2023 · 1 citations
Frame-level emotional state alignment method for speech emotion recognition
2023 · 1 citations
Retrieval Augmented Generation in Prompt-based Text-to-Speech Synthesis with Context-Aware Contrastive Language-Audio Pretraining
2024 · 1 citations
Multi-Loss Learning for Speech Emotion Recognition with Energy-Adaptive Mixup and Frame-Level Attention
2025
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
2025
Improving Audio Codec-based Zero-Shot Text-to-Speech Synthesis with Multi-Modal Context and Large Language Model
2024
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
2024
Top co-authors
Ya Li
· 7
Jinlong Xue
· 5
Yayue Deng
· 5
Fengping Wang
· 4
Bingsong Bai
· 2
Dengfeng Ke
· 2
Qifei Li
· 2
Yizhong Geng
· 2
Binghuai Lin
· 1
Chunfeng Wang
· 1
Cong Wang
· 1
Cong Wang
· 1
Topics
Audio Generation
Audio Understanding
Text-to-Speech
Speech Recognition
Voice Cloning
Music Generation
Multimodal Audio
Speech Enhancement