Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
James Glass โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
James Glass
74
papers ยท
1823
citations ยท
66
h-index
Moscow Institute of Thermal Technology ยท IIT@MIT ยท Artificial Intelligence in Medicine (Canada) ยท Massachusetts Institute of Technology
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
AST: Audio Spectrogram Transformer
2021 ยท 989 citations
An Unsupervised Autoregressive Model for Speech Representation Learning
2019 ยท 350 citations
Frame-level speaker embeddings for text-independent speaker recognition and analysis of end-to-end model
2018 ยท 86 citations
Learning Latent Representations for Speech Generation and Transformation
2017 ยท 53 citations
Analyzing Hidden Representations in End-to-End Automatic Speech Recognition Systems
2017 ยท 48 citations
Unsupervised Cross-Modal Alignment of Speech and Text Embedding Spaces
2018 ยท 48 citations
Speech2Vec: A Sequence-to-Sequence Framework for Learning Word Embeddings from Speech
2018 ยท 34 citations
Towards Transfer Learning for End-to-End Speech Synthesis from Deep Pre-Trained Language Models
2019 ยท 26 citations
A Study of Enhancement, Augmentation, and Autoencoder Methods for Domain Adaptation in Distant Speech Recognition
2018 ยท 24 citations
Learning Word Embeddings from Speech
2017 ยท 20 citations
On Training Recurrent Networks with Truncated Backpropagation Through Time in Speech Recognition
2018 ยท 19 citations
Transfer Learning From Audio-visual Grounding To Speech Recognition
2019 ยท 18 citations
Unsupervised Domain Adaptation for Robust Speech Recognition via Variational Autoencoder-Based Data Augmentation
2017 ยท 13 citations
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition
2021 ยท 12 citations
Learning Hierarchical Discrete Linguistic Units From Visually-grounded Speech
2019 ยท 11 citations
Top co-authors
Sameer Khurana
ยท 8
Hao Tang
ยท 7
Hilde Kuehne
ยท 6
Rogerio Feris
ยท 6
Yuan Gong
ยท 6
Heng-Jui Chang
ยท 5
David Harwath
ยท 4
Hung-yi Lee
ยท 3
Leonid Karlinsky
ยท 3
Yonatan Belinkov
ยท 3
Yung-Sung Chuang
ยท 3
David Cox
ยท 2
Topics
Speech Recognition
Audio Understanding
Speech Translation
Text-to-Speech
Speaker Analysis
Speech Enhancement
Multimodal Audio
Audio Generation
eess.AS
cs.CL