Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Samuel Thomas — most-cited papers & profile · Speech Audio
← authors
·
overview
Samuel Thomas
23
papers ·
82
citations ·
32
h-index
IBM (United States) · Arvalis - Institut du Végétal
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Invariant Representations for Noisy Speech Recognition
2016 · 66 citations
Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition
2018 · 5 citations
End-to-End Spoken Language Understanding Without Full Transcripts
2020 · 3 citations
A Non-autoregressive Model for Joint STT and TTS
2025 · 2 citations
Extending RNN-T-based speech recognition systems with emotion and language classification
2022 · 2 citations
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
2025 · 1 citations
mWhisper-Flamingo for Multilingual Audio-Visual Noise-Robust Speech Recognition
2025 · 1 citations
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions
2026
NLE: Non-autoregressive LLM-based ASR by Transcript Editing
2026
Self-Speculative Decoding for LLM-based ASR with CTC Encoder Drafts
2026
Towards Audio Token Compression in Large Audio Language Models
2025
Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?
2025
Leveraging Unpaired Text Data for Training End-to-End Speech-to-Intent Systems
2020
RNN Transducer Models For Spoken Language Understanding
2021
Speak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs
2021
Top co-authors
Brian Kingsbury
· 11
George Saon
· 10
Hong-Kwang J. Kuo
· 6
Hilde Kuehne
· 5
James Glass
· 5
Rogerio Feris
· 5
Zvi Kons
· 5
Andrew Rouditchenko
· 4
Kartik Audhkhasi
· 4
Ron Hoory
· 4
Vishal Sunder
· 4
Avihu Dekel
· 3
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Translation
Multimodal Audio
eess.AS
cs.CL
cs.LG
cs.SD
cs.AI