Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shuo-yiin Chang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shuo-yiin Chang
20
papers ·
386
citations ·
18
h-index
Google (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020 · 202 citations
Towards Fast and Accurate Streaming End-to-End ASR
2020 · 113 citations
Streaming End-to-end Speech Recognition For Mobile Devices
2018 · 23 citations
A Language Agnostic Multilingual Streaming On-Device ASR System
2022 · 10 citations
Improving the fusion of acoustic and text representations in RNN-T
2022 · 9 citations
Personal VAD: Speaker-Conditioned Voice Activity Detection
2019 · 8 citations
Multilingual and Fully Non-Autoregressive ASR with Large Language Model Fusion: A Comprehensive Study
2024 · 7 citations
FastEmit: Low-latency Streaming ASR with Sequence-level Emission Regularization
2020 · 6 citations
UML: A Universal Monolingual Output Layer for Multilingual ASR
2023 · 5 citations
A Better and Faster End-to-End Model for Streaming ASR
2020 · 2 citations
Improved Long-Form Speech Recognition by Jointly Modeling the Primary and Non-primary Speakers
2023 · 1 citations
On Neural Phone Recognition of Mixed-Source ECoG Signals
2019
E2E Segmenter: Joint Segmenting and Decoding for Long-Form ASR
2022
Turn-Taking Prediction for Natural Conversational Speech
2022
Streaming End-to-End Multilingual Speech Recognition with Joint Language Identification
2022
Top co-authors
Tara N. Sainath
· 15
Trevor Strohman
· 7
Bo Li
· 6
Ruoming Pang
· 6
Rohit Prabhavalkar
· 5
Yanzhang He
· 5
Yonghui Wu
· 5
Bo Li
· 4
Arun Narayanan
· 3
Chung-Cheng Chiu
· 3
David Rybach
· 3
Shaan Bijwadia
· 3
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Translation
Speaker Analysis
Speech Enhancement
Multimodal Audio
Voice Cloning
Audio Generation