Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guangyan Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Guangyan Zhang
15
papers ·
24
citations ·
18
h-index
Tsinghua University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style
2021 · 9 citations
Recent Advances in Speech Language Models: A Survey
2024 · 5 citations
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
2025 · 4 citations
CUHK-EE Voice Cloning System for ICASSP 2021 M2VoC Challenge
2021 · 4 citations
Environment Aware Text-to-Speech Synthesis
2021 · 1 citations
Mixed-Phoneme BERT: Improving BERT with Mixed Phoneme and Sup-Phoneme Representations for Text to Speech
2022 · 1 citations
Entropy-based Coarse and Compressed Semantic Speech Representation Learning
2025
SpeechAccentLLM: A Unified Framework for Foreign Accent Conversion and Text to Speech
2025
Applying the Information Bottleneck Principle to Prosodic Representation Learning
2021
A study on the efficacy of model pre-training in developing neural text-to-speech system
2021
iEmoTTS: Toward Robust Cross-Speaker Emotion Transfer and Control for Speech Synthesis based on Disentanglement between Prosody and Timbre
2022
Creating Personalized Synthetic Voices from Post-Glossectomy Speech with Guided Diffusion Models
2023
Comparing normalizing flows and diffusion models for prosody and acoustic modelling in text-to-speech
2023
Enabling Beam Search for Language Model-Based Text-to-Speech Synthesis
2024
Top co-authors
Tan Lee
· 7
Daxin Tan
· 5
Sheng Zhao
· 3
Yiwen Guo
· 3
Zehai Tu
· 3
Jialong Zuo
· 2
Jingyu Li
· 2
Kaitao Song
· 2
Shengpeng Ji
· 2
Tao Qin
· 2
Xiaoqi Jiao
· 2
Xu Tan
· 2
Topics
Text-to-Speech
Audio Generation
Voice Cloning
Speech Recognition
Speaker Analysis
Speech Enhancement
Speech Translation
Audio Understanding
Multimodal Audio