Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Xin Wang โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Xin Wang
286
papers ยท
1937
citations ยท
52
h-index
Wuhan University ยท Stony Brook University
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis
2018 ยท 51 citations
Transformation of low-quality device-recorded speech to high-quality speech using improved SEGAN model
2019 ยท 30 citations
Investigating self-supervised front ends for speech spoofing countermeasures
2021 ยท 12 citations
Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019 ยท 9 citations
WildSpoof Challenge Evaluation Plan
2025 ยท 8 citations
Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet
2019 ยท 8 citations
Neural source-filter waveform models for statistical parametric speech synthesis
2019 ยท 8 citations
Speaker Anonymization Using X-vector and Neural Waveform Models
2019 ยท 8 citations
Neural Harmonic-plus-Noise Waveform Model with Trainable Maximum Voice Frequency for Text-to-Speech Synthesis
2019 ยท 8 citations
A comparison of recent waveform generation and acoustic modeling methods for neural-network-based speech synthesis
2018 ยท 7 citations
Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama's voice using GAN, WaveNet and low-quality found data
2018 ยท 6 citations
Text-to-Speech Synthesis Techniques for MIDI-to-Audio Synthesis
2021 ยท 6 citations
Initial investigation of an encoder-decoder end-to-end TTS framework using marginalization of monotonic hard latent alignments
2019 ยท 5 citations
Investigation of enhanced Tacotron text-to-speech synthesis systems with self-attention for pitch accent language
2018 ยท 3 citations
Training Multi-Speaker Neural Text-to-Speech Systems using Speaker-Imbalanced Speech Corpora
2019 ยท 3 citations
Top co-authors
Nicholas Evans
ยท 6
Yihan Wu
ยท 3
Chang Zeng
ยท 2
Cheng Gong
ยท 2
Isao Echizen
ยท 2
Ji-Hoon Kim
ยท 2
Jinchuan Tian
ยท 2
Joon Son Chung
ยท 2
Shinji Watanabe
ยท 2
Shinnosuke Takamichi
ยท 2
Yi Zhao
ยท 2
Yuxuan Wang
ยท 2
Topics
Text-to-Speech
Audio Generation
Speech Enhancement
Speaker Analysis
Speech Recognition
eess.AS
Audio Understanding
cs.SD
Voice Cloning
Speech Translation