Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Zhiyong Wu โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Zhiyong Wu
53
papers ยท
52
citations ยท
33
h-index
Shandong University of Technology ยท Shantou Central Hospital ยท Tsinghua University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Towards Multi-Scale Speaking Style Modelling with Hierarchical Context Information for Mandarin Speech Synthesis
2022 ยท 11 citations
Noise Robust TTS for Low Resource Speakers using Pre-trained Model and Speech Enhancement
2020 ยท 8 citations
Emotion controllable speech synthesis using emotion-unlabeled dataset with the assistance of cross-domain speech emotion recognition
2020 ยท 4 citations
Non-Autoregressive Transformer ASR with CTC-Enhanced Decoder Input
2020 ยท 4 citations
Towards Multi-Scale Style Control for Expressive Speech Synthesis
2021 ยท 4 citations
Speaker Independent and Multilingual/Mixlingual Speech-Driven Talking Head Generation Using Phonetic Posteriorgrams
2020 ยท 3 citations
Adversarially learning disentangled speech representations for robust multi-factor voice conversion
2021 ยท 3 citations
The Multi-speaker Multi-style Voice Cloning Challenge 2021
2021 ยท 3 citations
Transformer-S2A: Robust and Efficient Speech-to-Animation
2021 ยท 2 citations
Speech Representation Disentanglement with Adversarial Mutual Information Learning for One-shot Voice Conversion
2022 ยท 2 citations
AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis
2025 ยท 1 citations
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
2025 ยท 1 citations
Syntactic representation learning for neural network based TTS with syntactic parse tree traversal
2020 ยท 1 citations
Enhancing Speaking Styles in Conversational Text-to-Speech Synthesis with Graph-based Multi-modal Context Modeling
2021 ยท 1 citations
Focus on the Sound around You: Monaural Target Speaker Extraction via Distance and Speaker Information
2023 ยท 1 citations
Top co-authors
Helen Meng
ยท 35
Shiyin Kang
ยท 13
Shun Lei
ยท 9
Xixin Wu
ยท 9
Yixuan Zhou
ยท 9
Changhe Song
ยท 7
Liyang Chen
ยท 6
Jun Chen
ยท 5
Chao Weng
ยท 4
Dan Su
ยท 4
Jingbei Li
ยท 4
Wei Rao
ยท 4
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Voice Cloning
Speech Enhancement
Music Generation
Speaker Analysis
Speech Translation
Audio Understanding
Multimodal Audio