Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yu-An Chung — most-cited papers & profile · Speech Audio
← authors
·
overview
Yu-An Chung
27
papers ·
2060
citations ·
29
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
AST: Audio Spectrogram Transformer
2021 · 989 citations
An Unsupervised Autoregressive Model for Speech Representation Learning
2019 · 350 citations
W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
2021 · 262 citations
Audio Word2Vec: Unsupervised Learning of Audio Segment Representations using Sequence-to-sequence Autoencoder
2016 · 194 citations
SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training
2021 · 50 citations
Unsupervised Cross-Modal Alignment of Speech and Text Embedding Spaces
2018 · 48 citations
Seamless: Multilingual Expressive and Streaming Speech Translation
2023 · 41 citations
Speech2Vec: A Sequence-to-Sequence Framework for Learning Word Embeddings from Speech
2018 · 34 citations
Towards Transfer Learning for End-to-End Speech Synthesis from Deep Pre-Trained Language Models
2019 · 26 citations
Learning Word Embeddings from Speech
2017 · 20 citations
SeamlessM4T: Massively Multilingual & Multimodal Machine Translation
2023 · 13 citations
Generative Pre-Training for Speech with Autoregressive Predictive Coding
2019 · 10 citations
Semi-Supervised Training for Improving Data Efficiency in End-to-End Speech Synthesis
2018 · 5 citations
Towards Unsupervised Speech-to-Text Translation
2018 · 5 citations
Speech-to-Speech Translation For A Real-world Unwritten Language
2022 · 5 citations
Top co-authors
James Glass
· 13
Changhan Wang
· 5
Juan Pino
· 5
Ann Lee
· 4
Christophe Ropers
· 4
Hirofumi Inaguma
· 4
Hongyu Gong
· 4
Paul-Ambroise Duquenne
· 4
Peng-Jen Chen
· 4
Artyom Kozhevnikov
· 3
Can Balioglu
· 3
Cynthia Gao
· 3
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Translation
Multimodal Audio
Speech Enhancement
Speaker Analysis
Audio Generation