Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yang Ai — most-cited papers & profile · Speech Audio
← authors
·
overview
Yang Ai
20
papers ·
184
citations ·
10
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
MP-SENet: A Speech Enhancement Model with Parallel Denoising of Magnitude and Phase Spectra
2023 · 113 citations
Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling
2019 · 30 citations
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
2024 · 15 citations
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
2024 · 10 citations
Low-Latency Neural Speech Phase Prediction based on Parallel Estimation Architecture and Anti-Wrapping Losses for Speech Generation Tasks
2024 · 5 citations
A Streamable Neural Audio Codec with Residual Scalar-Vector Quantization for Real-Time Communication
2025 · 3 citations
Knowledge-and-Data-Driven Amplitude Spectrum Prediction for Hierarchical Neural Vocoders
2020 · 3 citations
Long-frame-shift Neural Speech Phase Prediction with Spectral Continuity Enhancement and Interpolation Error Compensation
2023 · 3 citations
Reverberation Modeling for Source-Filter-based Neural Vocoder
2020 · 1 citations
Pitch-and-Spectrum-Aware Singing Quality Assessment with Bias Correction and Model Fusion
2024 · 1 citations
A Distilled Low-Latency Neural Vocoder with Explicit Amplitude and Phase Prediction
2025
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
2025
Neural Speech Separation with Parallel Amplitude and Phase Spectrum Estimation
2025
Universal Preference-Score-based Pairwise Speech Quality Assessment
2025
Vision-Integrated High-Quality Neural Speech Coding
2025
Top co-authors
Zhen-Hua Ling
· 13
Hui-Peng Du
· 7
Ye-Xin Lu
· 7
Xiao-Hang Jiang
· 4
Junichi Yamagishi
· 3
Rui-Chen Zheng
· 3
and Zhen-Hua Ling
· 2
Fei Liu
· 2
Haoyu Li
· 2
Yu-Fei Shi
· 2
Zhenhua Ling
· 2
Zhen-Hua Ling
· 2
Topics
Speech Enhancement
Audio Generation
Speech Recognition
Audio Understanding
Music Generation
Text-to-Speech
Speech Translation
math.IT
Multimodal Audio
Speaker Analysis