Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuxuan Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Yuxuan Wang
11
papers ·
315
citations ·
14
h-index
Inner Mongolia University · Advanced Energy Materials (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
2017 · 184 citations
Trainable Frontend For Robust and Far-Field Keyword Spotting
2016 · 12 citations
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
2024 · 5 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
2024 · 2 citations
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
2024 · 1 citations
DiSTAR: Diffusion over a Scalable Token Autoregressive Representation for Speech Generation
2025
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
2025
Top co-authors
Yuping Wang
· 5
Zhuo Chen
· 5
Chumin Li
· 3
Dongya Jia
· 3
Jiawei Chen
· 3
Jitong Chen
· 3
Shouda Liu
· 3
Xiaobin Zhuang
· 3
Chao Yao
· 2
Chenpeng Du
· 2
Dejian Zhong
· 2
Hui Li
· 2
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Translation
Music Generation
Speech Enhancement
Audio Understanding
Multimodal Audio