Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guangzhi Sun — most-cited papers & profile · Speech Audio
← authors
·
overview
Guangzhi Sun
21
papers ·
10
citations ·
40
h-index
University of Cambridge · Shandong Provincial QianFoShan Hospital · Shandong First Medical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Graph Neural Networks for Contextual ASR with the Tree-Constrained Pointer Generator
2023 · 1 citations
Whisper-PMFA: Partial Multi-Scale Feature Aggregation for Speaker Verification using Whisper Models
2024 · 1 citations
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
2024 · 1 citations
Investigation of Zero-shot Text-to-Speech Models for Enhancing Short-Utterance Speaker Verification
2025
Cross-Utterance Conditioned VAE for Non-Autoregressive Text-to-Speech
2022
Minimising Biasing Word Errors for Contextual ASR with the Tree-Constrained Pointer Generator
2022
Tree-constrained Pointer Generator with Graph Neural Network Encodings for Contextual Speech Recognition
2022
Spectral Clustering-aware Learning of Embeddings for Speaker Diarisation
2022
End-to-end Spoken Language Understanding with Tree-constrained Pointer Generator
2022
Knowledge-Aware Audio-Grounded Generative Slot Filling for Limited Annotated Data
2023
Cross-Utterance Conditioned VAE for Speech Generation
2023
Connecting Speech Encoder and Large Language Model for ASR
2023
Wav2Prompt: End-to-End Speech Prompt Generation and Tuning For LLM in Zero and Few-shot Learning
2024
SAML: Speaker Adaptive Mixture of LoRA Experts for End-to-End ASR
2024
Speaker Adaptation for Quantised End-to-End ASR Models
2024
Top co-authors
Chao Zhang
· 5
Philip C. Woodland
· 5
Thomas Fang Zheng
· 4
Mingxing Xu
· 3
Wenyi Yu
· 3
Chao Zhang
· 2
Chao Zhang
· 2
Fanglei Sun
· 2
Jun Wang
· 2
Qiuming Zhao
· 2
Shuai Wang
· 2
Weiqin Zu
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Speaker Analysis
Audio Generation
Speech Translation
Multimodal Audio
Speech Enhancement
Music Generation