Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guangzhi Sun — most-cited papers & profile · Speech Audio
← authors
·
overview
Guangzhi Sun
12
papers ·
149
citations ·
40
h-index
University of Cambridge · Shandong Provincial QianFoShan Hospital · Shandong First Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Fully-hierarchical fine-grained prosody modeling for interpretable speech synthesis
2020 · 93 citations
Tree-constrained Pointer Generator for End-to-end Contextual Speech Recognition
2021 · 30 citations
Generating diverse and natural text-to-speech samples using a quantized fine-grained VAE and auto-regressive prosody prior
2020 · 15 citations
Cross-Utterance Language Models with Acoustic Error Sampling
2020 · 5 citations
Speaker diarisation using 2D self-attentive combination of embeddings
2019 · 2 citations
Conditional Diffusion Model for Target Speaker Extraction
2023 · 2 citations
TorchAudio 2.1: Advancing speech recognition, self-supervised learning, and audio processing components for PyTorch
2023 · 1 citations
Enabling Auditory Large Language Models for Automatic Speech Quality Evaluation
2024 · 1 citations
Can Contextual Biasing Remain Effective with Whisper and GPT-2?
2023
Enhancing Quantised End-to-End ASR Models via Personalisation
2023
Speech-based Slot Filling using Large Language Models
2023
SOT Triggered Neural Clustering for Speaker Attributed ASR
2024
Top co-authors
Philip C. Woodland
· 6
Chao Zhang
· 5
Heiga Zen
· 2
Ron J. Weiss
· 2
Xianrui Zheng
· 2
Yonghui Wu
· 2
Yuan Cao
· 2
Andrew Rosenberg
· 1
Anurag Kumar
· 1
Bhuvana Ramabhadran
· 1
Caroline Chen
· 1
Changli Tang
· 1
Topics
Speech Recognition
Audio Understanding
Speaker Analysis
Text-to-Speech
Audio Generation
Speech Translation
Speech Enhancement