Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tao Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Tao Wang
112
papers ·
460
citations ·
0
h-index
Ministry of Transport · Research Institute of Highway
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Noise2Music: Text-conditioned Music Generation with Diffusion Models
2023 · 50 citations
The Volctrans Neural Speech Translation System for IWSLT 2021
2021 · 8 citations
SecoustiCodec: Cross-Modal Aligned Streaming Single-Codecbook Speech Codec
2025 · 6 citations
Singing-Tacotron: Global duration control attention and dynamic filter for End-to-end singing voice synthesis
2022 · 3 citations
Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition
2025 · 2 citations
GigaST: A 10,000-hour Pseudo Speech Translation Corpus
2022 · 2 citations
Half-Truth: A Partially Fake Audio Detection Dataset
2021 · 1 citations
UnifySpeech: A Unified Framework for Zero-shot Text-to-Speech and Voice Conversion
2023 · 1 citations
Boosting Fast and High-Quality Speech Synthesis with Linear Diffusion
2023 · 1 citations
SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation
2026
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards
2026
JoyVoice: Long-Context Conditioning for Anthropomorphic Multi-Speaker Conversational Synthesis
2025
Selective Masking Adversarial Attack on Automatic Speech Recognition Systems
2025
VQ-CTAP: Cross-Modal Fine-Grained Sequence Representation Learning for Speech Processing
2024
CampNet: Context-Aware Mask Prediction for End-to-End Text-Based Speech Editing
2022
Top co-authors
Jianhua Tao
· 17
Zhengqi Wen
· 14
Yong Ren
· 4
Le Xu
· 3
Rong Ye
· 3
Xin Qi
· 3
Bowen Li
· 2
Cheng Gong
· 2
Chenxing Li
· 2
Chen Zhang
· 2
Chu Yuan Zhang
· 2
Haogeng Liu
· 2
Topics
Text-to-Speech
Audio Generation
Speech Recognition
Speech Enhancement
Speech Translation
Multimodal Audio
cs.SD
Voice Cloning
eess.AS
Music Generation