Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiao-Hang Jiang — most-cited papers & profile · Speech Audio
← authors
·
overview
Xiao-Hang Jiang
12
papers ·
49
citations ·
4
h-index
University of Science and Technology of China
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
2024 · 36 citations
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
2024 · 10 citations
A Streamable Neural Audio Codec with Residual Scalar-Vector Quantization for Real-Time Communication
2025 · 3 citations
Ultra-Low-Bitrate Mel-Spectrogram-based Neural Speech Coding with Flow-Matching-based Refinement and Vocoding-driven Reconstruction
2026
CFMDCTCodec: A Low-Bitrate Neural Speech Codec with Noise-Prior-aware Conditional Flow Matching for MDCT-Spectral Enhancement
2026
An Ultra-Low-Bitrate Neural Speech Codec with Plain-to-Pseudo Synergistic Vector Quantization
2026
VoCodec: A Low-bitrate Streamable Neural Speech Codec with Voicing-driven Quantization
2026
QE-XVC: Zero-Shot Cross-Lingual Voice Conversion via Query-Enhancement and Conditional Flow Matching
2026
CodeSep: Low-Bitrate Codec-Driven Speech Separation with Base-Token Disentanglement and Auxiliary-Token Serial Prediction
2026
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
2025
Vision-Integrated High-Quality Neural Speech Coding
2025
ESTVocoder: An Excitation-Spectral-Transformed Neural Vocoder Conditioned on Mel Spectrogram
2024
Top co-authors
Hui-Peng Du
· 9
Zhen-Hua Ling
· 5
Yang Ai
· 4
Rui-Chen Zheng
· 3
Ye-Xin Lu
· 3
Ji Wu
· 2
Rui-Chen Zheng
· 2
Yang Ai
· 2
Yang Ai
· 2
Zhen-Hua Ling
· 2
and Zhen-Hua Ling
· 1
En-Wei Zhang
· 1
Topics
Audio Generation
Speech Enhancement
Speech Recognition
Audio Understanding
Text-to-Speech
Voice Cloning
Speech Translation
math.IT
Multimodal Audio