Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ankur Bapna — most-cited papers & profile · Speech Audio
← authors
·
overview
Ankur Bapna
18
papers ·
534
citations ·
30
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
2019 · 158 citations
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 · 112 citations
MAESTRO: Matched Speech Text Representations through Modality Matching
2022 · 71 citations
mSLAM: Massively multilingual joint pre-training for speech and text
2022 · 59 citations
SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training
2021 · 50 citations
AudioPaLM: A Large Language Model That Can Speak and Listen
2023 · 41 citations
Leveraging unsupervised and weakly-supervised data to improve direct speech-to-speech translation
2022 · 15 citations
FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech
2022 · 15 citations
Mu$^2$SLAM: Multitask, Multilingual Speech and Language Models
2022 · 5 citations
Understanding Shared Speech-Text Representations
2023 · 5 citations
XTREME-S: Evaluating Cross-lingual Speech Representations
2022 · 1 citations
JOIST: A Joint Speech and Text Streaming Model For ASR
2022 · 1 citations
Miipher: A Robust Speech Restoration Model Integrating Self-Supervised Speech and Text Representations
2023 · 1 citations
Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR
2022
Virtuoso: Massive Multilingual Speech-Text Joint Semi-Supervised Learning for Text-To-Speech
2022
Top co-authors
Yu Zhang
· 9
Bhuvana Ramabhadran
· 7
Zhehuai Chen
· 6
Alexis Conneau
· 5
Andrew Rosenberg
· 5
Jason Riesa
· 5
Heiga Zen
· 4
Melvin Johnson
· 4
Nobuyuki Morioka
· 4
Vera Axelrod
· 4
Ye Jia
· 4
Colin Cherry
· 3
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Multimodal Audio
Audio Understanding
Speech Enhancement
Audio Generation