Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Samuel Thomas — most-cited papers & profile · Speech Audio
← authors
·
overview
Samuel Thomas
17
papers ·
79
citations ·
32
h-index
IBM (United States) · Arvalis - Institut du Végétal
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Invariant Representations for Noisy Speech Recognition
2016 · 66 citations
Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition
2018 · 5 citations
A Non-autoregressive Model for Joint STT and TTS
2025 · 2 citations
Extending RNN-T-based speech recognition systems with emotion and language classification
2022 · 2 citations
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
2025 · 1 citations
mWhisper-Flamingo for Multilingual Audio-Visual Noise-Robust Speech Recognition
2025 · 1 citations
Towards Audio Token Compression in Large Audio Language Models
2025
Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?
2025
Leveraging Unpaired Text Data for Training End-to-End Speech-to-Intent Systems
2020
Speak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs
2021
Improving End-to-End Models for Set Prediction in Spoken Language Understanding
2022
Integrating Text Inputs For Training and Adapting RNN Transducer ASR Models
2022
Towards Reducing the Need for Speech Training Data To Build Spoken Language Understanding Systems
2022
Tokenwise Contrastive Pretraining for Finer Speech-to-BERT Alignment in End-to-End Speech-to-Intent Systems
2022
Comparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages
2023
Top co-authors
Brian Kingsbury
· 7
George Saon
· 6
Hilde Kuehne
· 5
James Glass
· 5
Andrew Rouditchenko
· 4
Hong-Kwang J. Kuo
· 4
Rogério Feris
· 4
Edmilson Morais
· 3
Hong-Kwang Kuo
· 3
Kartik Audhkhasi
· 3
Vishal Sunder
· 3
Zvi Kons
· 3
Topics
Speech Recognition
Audio Understanding
Speech Translation
Text-to-Speech
Multimodal Audio