Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
George Saon — most-cited papers & profile · Speech Audio
← authors
·
overview
George Saon
28
papers ·
141
citations ·
35
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Direct Acoustics-to-Word Models for English Conversational Speech Recognition
2017 · 114 citations
Embedding-Based Speaker Adaptive Training of Deep Neural Networks
2017 · 7 citations
Reducing Exposure Bias in Training Recurrent Neural Network Transducers
2021 · 6 citations
Distributed Deep Learning Strategies For Automatic Speech Recognition
2019 · 3 citations
Single headed attention based sequence-to-sequence model for state-of-the-art results on Switchboard
2020 · 3 citations
A Non-autoregressive Model for Joint STT and TTS
2025 · 2 citations
Extending RNN-T-based speech recognition systems with emotion and language classification
2022 · 2 citations
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
2025 · 1 citations
Language Modeling with Highway LSTM
2017 · 1 citations
Distributed Training of Deep Neural Network Acoustic Models for Automatic Speech Recognition
2020 · 1 citations
Asynchronous Decentralized Distributed Training of Acoustic Models
2021 · 1 citations
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
2026
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions
2026
NLE: Non-autoregressive LLM-based ASR by Transcript Editing
2026
Self-Speculative Decoding for LLM-based ASR with CTC Encoder Drafts
2026
Top co-authors
Brian Kingsbury
· 14
Xiaodong Cui
· 10
Samuel Thomas
· 6
Avihu Dekel
· 4
David Kung
· 4
Gakuto Kurata
· 4
Michael Picheny
· 4
Ulrich Finkler
· 4
Bhuvana Ramabhadran
· 3
Hagai Aronowitz
· 3
Hong-Kwang J. Kuo
· 3
Kartik Audhkhasi
· 3
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio
Speaker Analysis