Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Hao Tang โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Hao Tang
128
papers ยท
1138
citations ยท
3
h-index
Northwestern Polytechnical University ยท Peking University ยท Nanjing General Hospital of Nanjing Military Command ยท Hebei Yiling Hospital
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
An Unsupervised Autoregressive Model for Speech Representation Learning
2019 ยท 350 citations
Frame-level speaker embeddings for text-independent speaker recognition and analysis of end-to-end model
2018 ยท 86 citations
A Study of Enhancement, Augmentation, and Autoencoder Methods for Domain Adaptation in Distant Speech Recognition
2018 ยท 24 citations
On Training Recurrent Networks with Truncated Backpropagation Through Time in Speech Recognition
2018 ยท 19 citations
VoiceID Loss: Speech Enhancement for Speaker Verification
2019 ยท 11 citations
Self-supervised Predictive Coding Models Encode Speaker and Phonetic Information in Orthogonal Subspaces
2023 ยท 9 citations
End-to-End Training Approaches for Discriminative Segmental Models
2016 ยท 7 citations
Supervised Attention in Sequence-to-Sequence Models for Speech Recognition
2022 ยท 5 citations
Efficient Segmental Cascades for Speech Recognition
2016 ยท 4 citations
Multitask Learning with Low-Level Auxiliary Tasks for Encoder-Decoder Based Speech Recognition
2017 ยท 3 citations
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
2022 ยท 2 citations
Estimating the Completeness of Discrete Speech Units
2024 ยท 2 citations
Effective Context in Neural Speech Models
2025 ยท 1 citations
Unsupervised Adaptation with Interpretable Disentangled Representations for Distant Conversational Speech Recognition
2018 ยท 1 citations
Analyzing Acoustic Word Embeddings from Pre-trained Self-supervised Speech Models
2022 ยท 1 citations
Top co-authors
Hung-yi Lee
ยท 7
James Glass
ยท 7
Tzu-Quan Lin
ยท 6
Liang Lu
ยท 2
Weiran Wang
ยท 2
Alexandre Mourachko
ยท 1
Chris Dyer
ยท 1
Duc Le
ยท 1
Hsi-Chun Cheng
ยท 1
Jay Mahadeokar
ยท 1
Lingpeng Kong
ยท 1
Noah A. Smith
ยท 1
Topics
Speech Recognition
Audio Understanding
Speech Translation
Speaker Analysis
Speech Enhancement
Text-to-Speech
eess.AS
cs.CL
cs.SD
Audio Generation