Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yang Ai — most-cited papers & profile · Speech Audio
← authors
·
overview
Yang Ai
49
papers ·
22
citations ·
10
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MP-SENet: A Speech Enhancement Model with Parallel Denoising of Magnitude and Phase Spectra
2023 · 113 citations
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
2024 · 36 citations
Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling
2019 · 30 citations
APNet: An All-Frame-Level Neural Vocoder Incorporating Direct Prediction of Amplitude and Phase Spectra
2023 · 17 citations
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
2024 · 15 citations
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
2024 · 10 citations
Zero-shot personalized lip-to-speech synthesis with face image based voice control
2023 · 6 citations
Explicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech Enhancement
2023 · 5 citations
Low-Latency Neural Speech Phase Prediction based on Parallel Estimation Architecture and Anti-Wrapping Losses for Speech Generation Tasks
2024 · 5 citations
APCodec+: A Spectrum-Coding-Based High-Fidelity and High-Compression-Rate Neural Audio Codec with Staged Training Paradigm
2024 · 4 citations
A Streamable Neural Audio Codec with Residual Scalar-Vector Quantization for Real-Time Communication
2025 · 3 citations
Knowledge-and-Data-Driven Amplitude Spectrum Prediction for Hierarchical Neural Vocoders
2020 · 3 citations
Long-frame-shift Neural Speech Phase Prediction with Spectral Continuity Enhancement and Interpolation Error Compensation
2023 · 3 citations
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
2024 · 2 citations
Enhancing Noise Robustness for Neural Speech Codecs through Resource-Efficient Progressive Quantization Perturbation Simulation
2025 · 1 citations
Top co-authors
Zhen-Hua Ling
· 40
Hui-Peng Du
· 24
Ye-Xin Lu
· 19
Rui-Chen Zheng
· 11
Xiao-Hang Jiang
· 11
Fei Liu
· 6
and Zhen-Hua Ling
· 4
Li-Rong Dai
· 4
Yu-Fei Shi
· 4
Ji Wu
· 3
Junichi Yamagishi
· 3
Zheng-Yan Sheng
· 3
Topics
Audio Generation
Speech Enhancement
eess.AS
Text-to-Speech
Speech Recognition
cs.SD
Music Generation
Audio Understanding
Speaker Analysis
Voice Cloning