Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Brian Yan — most-cited papers & profile · Speech Audio
← authors
·
overview
Brian Yan
37
papers ·
15
citations ·
15
h-index
Western University · London Health Sciences Centre
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Searchable Hidden Intermediates for End-to-End Models of Decomposable Sequence Tasks
2021 · 34 citations
ESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and Understanding
2022 · 28 citations
Improving Massively Multilingual ASR With Auxiliary CTC Objectives
2023 · 26 citations
Differentiable Allophone Graphs for Language-Universal Speech Recognition
2021 · 7 citations
Token-level Sequence Labeling for Spoken Language Understanding using Compositional End-to-End Models
2022 · 6 citations
Exploring Speech Recognition, Translation, and Understanding with Discrete Speech Units: A Comparative Study
2023 · 2 citations
OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
2024 · 2 citations
Joint Beam Search Integrating CTC, Attention, and Transducer Decoders
2024 · 2 citations
ESPnet-ST IWSLT 2021 Offline Speech Translation System
2021 · 1 citations
CTC Alignments Improve Autoregressive Translation
2022 · 1 citations
A Study on the Integration of Pipeline and E2E SLU systems for Spoken Semantic Parsing toward STOP Quality Challenge
2023 · 1 citations
CMU's IWSLT 2024 Simultaneous Speech Translation System
2024 · 1 citations
CS-YODAS: A Mined Dataset of In-the-Wild Code-Switched Speech
2026
Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech
2026
CS-FLEURS: A Massively Multilingual and Code-Switched Speech Dataset
2025
Top co-authors
Shinji Watanabe
· 36
Siddharth Dalmia
· 11
Siddhant Arora
· 10
Jiatong Shi
· 9
Yifan Peng
· 9
Xuankai Chang
· 8
William Chen
· 6
Dan Berrebbi
· 5
Florian Metze
· 4
Matthew Wiesner
· 4
Muhammad Shakeel
· 4
Soumi Maiti
· 4
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Audio Generation
Music Generation
cs.SD
cs.CL
Multimodal Audio
eess.AS