Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hui-Peng Du — most-cited papers & profile · Speech Audio
← authors
·
overview
Hui-Peng Du
23
papers ·
68
citations ·
6
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
2024 · 36 citations
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
2024 · 15 citations
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
2024 · 10 citations
APCodec+: A Spectrum-Coding-Based High-Fidelity and High-Compression-Rate Neural Audio Codec with Staged Training Paradigm
2024 · 4 citations
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
2024 · 2 citations
Pitch-and-Spectrum-Aware Singing Quality Assessment with Bias Correction and Model Fusion
2024 · 1 citations
Ultra-Low-Bitrate Mel-Spectrogram-based Neural Speech Coding with Flow-Matching-based Refinement and Vocoding-driven Reconstruction
2026
CFMDCTCodec: A Low-Bitrate Neural Speech Codec with Noise-Prior-aware Conditional Flow Matching for MDCT-Spectral Enhancement
2026
QE-XVC: Zero-Shot Cross-Lingual Voice Conversion via Query-Enhancement and Conditional Flow Matching
2026
LatentFlowSR: High-Fidelity Audio Super-Resolution via Noise-Robust Latent Flow Matching
2026
CodeSep: Low-Bitrate Codec-Driven Speech Separation with Base-Token Disentanglement and Auxiliary-Token Serial Prediction
2026
Say More with Less: Variable-Frame-Rate Speech Tokenization via Adaptive Clustering and Implicit Duration Coding
2025
A Distilled Low-Latency Neural Vocoder with Explicit Amplitude and Phase Prediction
2025
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
2025
Is GAN Necessary for Mel-Spectrogram-based Neural Vocoder?
2025
Top co-authors
Ye-Xin Lu
· 13
Zhen-Hua Ling
· 13
Xiao-Hang Jiang
· 9
Yang Ai
· 7
Yang Ai
· 7
Yang Ai
· 4
Rui-Chen Zheng
· 3
Rui-Chen Zheng
· 3
Yu-Fei Shi
· 3
Zhen-Hua Ling
· 3
and Zhen-Hua Ling
· 2
Chong Deng
· 1
Topics
Audio Generation
Speech Enhancement
Text-to-Speech
Speech Recognition
Audio Understanding
Voice Cloning
Speech Translation
Music Generation
math.IT
Multimodal Audio