Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ann Lee — most-cited papers & profile · Speech Audio
← authors
·
overview
Ann Lee
23
papers ·
168
citations ·
33
h-index
Macquarie University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Seamless: Multilingual Expressive and Streaming Speech Translation
2023 · 41 citations
Semi-Supervised Speech Recognition via Local Prior Matching
2020 · 28 citations
Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training
2021 · 18 citations
Text-Free Prosody-Aware Generative Spoken Language Modeling
2021 · 16 citations
SeamlessM4T: Massively Multilingual & Multimodal Machine Translation
2023 · 13 citations
Sequence-to-Sequence Speech Recognition with Time-Depth Separable Convolutions
2019 · 11 citations
Direct speech-to-speech translation with discrete units
2021 · 7 citations
VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
2021 · 6 citations
Speech-to-Speech Translation For A Real-world Unwritten Language
2022 · 5 citations
fairseq S^2: A Scalable and Integrable Speech Synthesis Toolkit
2021 · 4 citations
SpeechMatrix: A Large-Scale Mined Corpus of Multilingual Speech-to-Speech Translations
2022 · 4 citations
Multilingual Speech-to-Speech Translation into Multiple Target Languages
2023 · 4 citations
UnitY: Two-pass Direct Speech-to-speech Translation with Discrete Units
2022 · 3 citations
Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation
2022 · 2 citations
Augmentation Invariant Discrete Representation for Generative Spoken Language Modeling
2022 · 2 citations
Top co-authors
Juan Pino
· 13
Peng-Jen Chen
· 10
Yossi Adi
· 7
Holger Schwenk
· 5
Paul-Ambroise Duquenne
· 5
Jiatao Gu
· 4
Ning Dong
· 4
Yun Tang
· 4
Adam Polyak
· 3
Gabriel Synnaeve
· 3
Ilia Kulikov
· 3
Shinji Watanabe
· 3
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Understanding
Multimodal Audio
Audio Generation
Speech Enhancement