Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Siddharth Dalmia — most-cited papers & profile · Speech Audio
← authors
·
overview
Siddharth Dalmia
21
papers ·
143
citations ·
14
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Branchformer: Parallel MLP-Attention Architectures to Capture Local and Global Context for Speech Recognition and Understanding
2022 · 40 citations
Searchable Hidden Intermediates for End-to-End Models of Decomposable Sequence Tasks
2021 · 34 citations
FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech
2022 · 15 citations
Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion
2019 · 11 citations
Differentiable Allophone Graphs for Language-Universal Speech Recognition
2021 · 7 citations
Token-level Sequence Labeling for Spoken Language Understanding using Compositional End-to-End Models
2022 · 6 citations
Cross-Attention End-to-End ASR for Two-Party Conversations
2019 · 4 citations
Multilingual Speech Recognition with Corpus Relatedness Sampling
2019 · 3 citations
ESPnet-ST IWSLT 2021 Offline Speech Translation System
2021 · 1 citations
CTC Alignments Improve Autoregressive Translation
2022 · 1 citations
Transforming LLMs into Cross-modal and Cross-lingual Retrieval Systems
2024 · 1 citations
Fast-MD: Fast Multi-Decoder End-to-End Speech Translation with Non-Autoregressive Hidden Intermediates
2021
ESPnet-SLU: Advancing Spoken Language Understanding through ESPnet
2021
Joint Modeling of Code-Switched and Monolingual ASR via Conditional Factorization
2021
LegoNN: Building Modular Encoder-Decoder Models
2022
Top co-authors
Shinji Watanabe
· 14
Florian Metze
· 8
Brian Yan
· 4
Brian Yan
· 4
Siddhant Arora
· 4
Yifan Peng
· 4
Alan W Black
· 3
Brian Yan
· 3
Hirofumi Inaguma
· 3
Xuankai Chang
· 3
Dan Berrebbi
· 2
Karthik Ganesan
· 2
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Speaker Analysis
Audio Generation
Music Generation
Multimodal Audio