Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yan Lu — most-cited papers & profile · Speech Audio
← authors
·
overview
Yan Lu
78
papers ·
123
citations ·
13
h-index
Zhejiang University of Science and Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Universal Speech Token Learning via Low-Bitrate Neural Codec and Pretrained Representations
2025 · 6 citations
Interactive Speech and Noise Modeling for Speech Enhancement
2020 · 4 citations
End-to-End Neural Speech Coding for Real-Time Communications
2022 · 2 citations
Latent-Domain Predictive Neural Speech Coding
2022 · 2 citations
StreamMel: Real-Time Zero-shot Text-to-Speech via Interleaved Continuous Autoregressive Modeling
2025 · 1 citations
Phoneme-based Distribution Regularization for Speech Enhancement
2021 · 1 citations
Disentangled Feature Learning for Real-Time Neural Speech Coding
2022 · 1 citations
DasFormer: Deep Alternating Spectrogram Transformer for Multi/Single-Channel Speech Separation
2023 · 1 citations
Low-latency Speech Enhancement via Speech Token Generation
2023 · 1 citations
Is Text All You Need? Text as a Universal Information Bottleneck for Speech LLMs
2026
Closing the Modality Reasoning Gap for Speech Large Language Models
2026
A Unified Neural Codec Language Model for Selective Editable Text to Speech Generation
2026
SpeechLLM-as-Judges: Towards General and Interpretable Speech Quality Evaluation
2025
Zero-Shot Streaming Text to Speech Synthesis with Transducer and Auto-Regressive Modeling
2025
Pseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis
2025
Top co-authors
Shujie Liu
· 9
Yuan Zhang
· 8
Jinyu Li
· 7
Hui Wang
· 6
Xue Jiang
· 6
Yifan Yang
· 6
Lingwei Meng
· 5
Yanqing Liu
· 5
Haoqin Sun
· 3
Jiaming Zhou
· 3
Jianwei Yu
· 3
Yong Qin
· 3
Topics
Speech Enhancement
Audio Generation
Audio Understanding
Speech Recognition
Text-to-Speech
Speaker Analysis
eess.AS
Speech Translation
Multimodal Audio
cs.SD