Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Andrew Rosenberg — most-cited papers & profile · Speech Audio
← authors
·
overview
Andrew Rosenberg
25
papers ·
174
citations ·
30
h-index
Google (United States) · University of Michigan
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 · 112 citations
End-to-End ASR-free Keyword Search from Speech
2017 · 91 citations
MAESTRO: Matched Speech Text Representations through Modality Matching
2022 · 71 citations
Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning
2019 · 26 citations
Injecting Text in Self-Supervised Speech Pretraining
2021 · 25 citations
Generating diverse and natural text-to-speech samples using a quantized fine-grained VAE and auto-regressive prosody prior
2020 · 15 citations
Accented Speech Recognition: Benchmarking, Pre-training, and Diverse Data
2022 · 12 citations
JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition
2023 · 8 citations
Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition
2018 · 5 citations
Understanding Shared Speech-Text Representations
2023 · 5 citations
Speech Recognition with Augmented Synthesized Speech
2019 · 4 citations
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
2024
Non-Parallel Voice Conversion for ASR Augmentation
2022
Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR
2022
G-Augment: Searching for the Meta-Structure of Data Augmentation Policies for ASR
2022
Top co-authors
Bhuvana Ramabhadran
· 22
Gary Wang
· 12
Yu Zhang
· 10
Zhehuai Chen
· 8
Ankur Bapna
· 6
Heiga Zen
· 5
Kyle Kastner
· 5
Pedro Moreno
· 5
Kartik Audhkhasi
· 4
Yonghui Wu
· 4
Zhong Meng
· 4
Fadi Biadsy
· 3
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Generation
Audio Understanding
Multimodal Audio
Voice Cloning
Speaker Analysis