Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kartik Audhkhasi — most-cited papers & profile · Speech Audio
← authors
·
overview
Kartik Audhkhasi
13
papers ·
194
citations ·
27
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Direct Acoustics-to-Word Models for English Conversational Speech Recognition
2017 · 114 citations
Invariant Representations for Noisy Speech Recognition
2016 · 66 citations
Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition
2018 · 5 citations
Analysis of Self-Attention Head Diversity for Conformer-based Automatic Speech Recognition
2022 · 4 citations
Single headed attention based sequence-to-sequence model for state-of-the-art results on Switchboard
2020 · 3 citations
LegoSLM: Connecting LLM with Speech Encoder using CTC Posteriors
2025
Building competitive direct acoustics-to-word models for English conversational speech recognition
2017
Acoustically Grounded Word Embeddings for Improved Acoustics-to-Word Speech Recognition
2019
Leveraging Unpaired Text Data for Training End-to-End Speech-to-Intent Systems
2020
Modular Hybrid Autoregressive Transducer
2022
O-1: Self-training with Oracle and 1-best Hypothesis
2023
STAB: Speech Tokenizer Assessment Benchmark
2024
Top co-authors
Bhuvana Ramabhadran
· 9
George Saon
· 3
Michael Picheny
· 3
Samuel Thomas
· 3
Yinghui Huang
· 3
Andrew Rosenberg
· 2
Brian Kingsbury
· 2
Andrew Rosenberg
· 1
Ankur Bapna
· 1
Brian Kingsbury
· 1
Chulayuth Asawaroengchai
· 1
David Nahamoo
· 1
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Multimodal Audio