Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Gary Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Gary Wang
15
papers ·
137
citations ·
29
h-index
Google (United States) · Simon Fraser University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 · 112 citations
Injecting Text in Self-Supervised Speech Pretraining
2021 · 25 citations
Accented Speech Recognition: Benchmarking, Pre-training, and Diverse Data
2022 · 12 citations
Deep Text-to-Speech System with Seq2Seq Model
2019 · 8 citations
Understanding Shared Speech-Text Representations
2023 · 5 citations
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
2024
Non-Parallel Voice Conversion for ASR Augmentation
2022
G-Augment: Searching for the Meta-Structure of Data Augmentation Policies for ASR
2022
Virtuoso: Massive Multilingual Speech-Text Joint Semi-Supervised Learning for Text-To-Speech
2022
Modular Hybrid Autoregressive Transducer
2022
High-precision Voice Search Query Correction via Retrievable Speech-text Embedings
2024
Extending Multilingual Speech Synthesis to 100+ Languages without Transcribed Data
2024
ASTRA: Aligning Speech and Text Representations for Asr without Sampling
2024
Zero-shot Cross-lingual Voice Transfer for TTS
2024
Top co-authors
Andrew Rosenberg
· 12
Bhuvana Ramabhadran
· 11
Zhehuai Chen
· 6
Kyle Kastner
· 5
Yu Zhang
· 5
Ankur Bapna
· 3
Fadi Biadsy
· 3
Chung-Cheng Chiu
· 2
Daniel S. Park
· 2
Fran\c{c}oise Beaufays
· 2
Heiga Zen
· 2
Isaac Elias
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Generation
Speech Translation
Audio Understanding
Multimodal Audio
Voice Cloning
Speaker Analysis