Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jiaming Zhou — most-cited papers & profile · Speech Audio
← authors
·
overview
Jiaming Zhou
36
papers ·
10
citations ·
2
h-index
Nankai University · China Institute of Water Resources and Hydropower Research
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
PB-LRDWWS System for the SLT 2024 Low-Resource Dysarthria Wake-Up Word Spotting Challenge
2024 · 3 citations
StreamMel: Real-Time Zero-shot Text-to-Speech via Interleaved Continuous Autoregressive Modeling
2025 · 1 citations
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
2025 · 1 citations
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
2026
H-SAGE: Holistic Speaker-Aware Guided Experts for MoE-based Multi-Talker ASR
2026
CosyEdit2: Speech-Editing-Oriented Reinforcement Learning Unlocks Better Zero-Shot TTS
2026
CosyEdit: Unlocking End-to-End Speech Editing Capability from Zero-Shot Text-to-Speech Models
2026
SpeechLLM-as-Judges: Towards General and Interpretable Speech Quality Evaluation
2025
GLAD: Global-Local Aware Dynamic Mixture-of-Experts for Multi-Talker ASR
2025
RealTalk-CN: A Realistic Chinese Speech-Text Dialogue Benchmark With Cross-Modal Interaction Analysis
2025
DIFFA: Large Language Diffusion Models Can Listen and Understand
2025
A Self-Training Approach for Whisper to Enhance Long Dysarthric Speech Recognition
2025
FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching
2025
CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition
2025
M2R-Whisper: Multi-stage and Multi-scale Retrieval Augmentation for Enhancing Whisper
2024
Top co-authors
Yong Qin
· 20
Shiwan Zhao
· 16
Hui Wang
· 11
Haoqin Sun
· 9
Aobo Kong
· 6
Shiyao Wang
· 5
Yuhang Jia
· 5
Yujie Guo
· 5
Jinyu Li
· 3
Junyang Chen
· 3
Shujie Liu
· 3
Xi Yang
· 3
Topics
Speech Recognition
cs.SD
Speech Translation
eess.AS
Speaker Analysis
Text-to-Speech
Audio Generation
Audio Understanding
cs.CL
Multimodal Audio