Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Nancy F. Chen — most-cited papers & profile · Speech Audio
← authors
·
overview
Nancy F. Chen
33
papers ·
11
citations ·
0
h-index
Agency for Science, Technology and Research · Institute of Materials Research and Engineering
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AudioBench: A Universal Benchmark for Audio Large Language Models
2024 · 4 citations
MultiGen: Child-Friendly Multilingual Speech Generator with LLMs
2025 · 3 citations
MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues
2024 · 2 citations
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
2026
Goodness-of-pronunciation without phoneme time alignment
2026
Benchmarking Contextual and Paralinguistic Reasoning in Speech-LLMs: A Case Study with In-the-Wild Data
2025
Prompt-Unseen-Emotion: Zero-shot Expressive Speech Synthesis with Prompt-LLM Contextual Knowledge for Mixed Emotions
2025
A correlation-permutation approach for speech-music encoders model merging
2025
Distilling a speech and music encoder with task arithmetic
2025
EPIC TTS Models: Empirical Pruning Investigations Characterizing Text-To-Speech Models
2022
SNIPER Training: Single-Shot Sparse Training for Text-to-Speech
2022
Noise robust distillation of self-supervised speech models via correlation metrics
2023
TTSlow: Slow Down Text-to-Speech with Efficiency Robustness Evaluations
2024
PRESENT: Zero-Shot Text-to-Prosody Control
2024
MoWE-Audio: Multitask AudioLLMs with Mixture of Weak Encoders
2024
Top co-authors
Xiaoxue Gao
· 7
Eng Siong Chng
· 3
Geyu Lin
· 3
Hung-yi Lee
· 3
Shuo Sun
· 3
Wenyu Zhang
· 3
Bin Wang
· 2
Dorien Herremans
· 2
Jinyang Wu
· 2
Yiming Chen
· 2
Chen Zhang
· 1
Dianwen Ng
· 1
Topics
Text-to-Speech
Audio Understanding
Speech Recognition
Multimodal Audio
Audio Generation
cs.AI
cs.CL
Speech Translation
eess.AS
eess.SP