Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Yu Zhang โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Yu Zhang
54
papers ยท
658
citations ยท
40
h-index
Loughborough University ยท East University Of Heilongjiang ยท Fiberhome Technology Group (China)
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 ยท 112 citations
MAESTRO: Matched Speech Text Representations through Modality Matching
2022 ยท 71 citations
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021 ยท 30 citations
Self-supervised Learning with Random-projection Quantizer for Speech Recognition
2022 ยท 29 citations
Advances in Joint CTC-Attention based End-to-End Speech Recognition with a Deep CNN Encoder and RNN-LM
2017 ยท 21 citations
Efficient Domain Adaptation for Speech Foundation Models
2023 ยท 19 citations
FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech
2022 ยท 15 citations
Massively Multilingual Shallow Fusion with Large Language Models
2023 ยท 14 citations
Unsupervised Domain Adaptation for Robust Speech Recognition via Variational Autoencoder-Based Data Augmentation
2017 ยท 13 citations
Training Text-To-Speech Systems From Synthetic Data: A Practical Approach For Accent Transfer Tasks
2022 ยท 11 citations
E3 TTS: Easy End-to-End Diffusion-based Text to Speech
2023 ยท 10 citations
Bytes are All You Need: End-to-End Multilingual Speech Recognition and Synthesis with Bytes
2018 ยท 9 citations
Unsupervised Data Selection via Discrete Speech Representation for ASR
2022 ยท 9 citations
Residual Adapters for Few-Shot Text-to-Speech Speaker Adaptation
2022 ยท 9 citations
JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition
2023 ยท 8 citations
Top co-authors
Ankur Bapna
ยท 9
Bhuvana Ramabhadran
ยท 8
Zhehuai Chen
ยท 8
Bo Li
ยท 7
Trevor Strohman
ยท 7
Andrew Rosenberg
ยท 6
Heiga Zen
ยท 6
Tara N. Sainath
ยท 6
Rohit Prabhavalkar
ยท 5
Gary Wang
ยท 4
Nanxin Chen
ยท 4
Nobuyuki Morioka
ยท 4
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Speech Enhancement
Audio Understanding
Speaker Analysis
Multimodal Audio
Voice Cloning
Music Generation