Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhen-Hua Ling — most-cited papers & profile · Speech Audio
← authors
·
overview
Zhen-Hua Ling
39
papers ·
166
citations ·
45
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
2024 · 36 citations
Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling
2019 · 30 citations
APNet: An All-Frame-Level Neural Vocoder Incorporating Direct Prediction of Amplitude and Phase Spectra
2023 · 17 citations
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
2024 · 15 citations
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
2024 · 10 citations
Voice Conversion by Cascading Automatic Speech Recognition and Text-to-Speech Synthesis with Prosody Transfer
2020 · 6 citations
Zero-shot personalized lip-to-speech synthesis with face image based voice control
2023 · 6 citations
Learning latent representations for style control and transfer in end-to-end speech synthesis
2018 · 5 citations
Low-Latency Neural Speech Phase Prediction based on Parallel Estimation Architecture and Anti-Wrapping Losses for Speech Generation Tasks
2024 · 5 citations
Recognition-Synthesis Based Non-Parallel Voice Conversion with Adversarial Learning
2020 · 4 citations
PoNet: Pooling Network for Efficient Token Mixing in Long Sequences
2021 · 4 citations
APCodec+: A Spectrum-Coding-Based High-Fidelity and High-Compression-Rate Neural Audio Codec with Staged Training Paradigm
2024 · 4 citations
A Streamable Neural Audio Codec with Residual Scalar-Vector Quantization for Real-Time Communication
2025 · 3 citations
Channel adversarial training for cross-channel text-independent speaker recognition
2019 · 3 citations
Knowledge-and-Data-Driven Amplitude Spectrum Prediction for Hierarchical Neural Vocoders
2020 · 3 citations
Top co-authors
Hui-Peng Du
· 13
Yang Ai
· 13
Ye-Xin Lu
· 12
Yang Ai
· 9
Xiao-Hang Jiang
· 5
Li-Rong Dai
· 4
Rui-Chen Zheng
· 4
Jing-Xuan Zhang
· 3
Rui-Chen Zheng
· 3
Yang Ai
· 3
Zheng-Yan Sheng
· 3
Fei Liu
· 2
Topics
Audio Generation
Speech Enhancement
Speech Recognition
Text-to-Speech
Audio Understanding
Speech Translation
Speaker Analysis
Voice Cloning
Multimodal Audio
Music Generation