Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ye Bai — most-cited papers & profile · Speech Audio
← authors
·
overview
Ye Bai
24
papers ·
61
citations ·
20
h-index
Shandong Institute of Automation · Institute of Automation
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition
2020 · 10 citations
Fast End-to-End Speech Recognition via Non-Autoregressive Models and Cross-Modal Knowledge Transferring from BERT
2021 · 8 citations
Synchronous Transformers for End-to-End Speech Recognition
2019 · 5 citations
Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition
2020 · 5 citations
One In A Hundred: Select The Best Predicted Sequence from Numerous Candidates for Streaming Speech Recognition
2020 · 5 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Learn Spelling from Teachers: Transferring Knowledge from Language Models to Sequence-to-Sequence Speech Recognition
2019 · 4 citations
Decoupling Pronunciation and Language for End-to-end Code-switching Automatic Speech Recognition
2020 · 4 citations
Rnn-transducer with language bias for end-to-end Mandarin-English code-switching speech recognition
2020 · 3 citations
TSNAT: Two-Step Non-Autoregressvie Transformer Models for Speech Recognition
2021 · 3 citations
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
2024 · 2 citations
Half-Truth: A Partially Fake Audio Detection Dataset
2021 · 1 citations
Parameter-Efficient Conformers via Sharing Sparsely-Gated Experts for End-to-End Speech Recognition
2022 · 1 citations
End-to-End Training for Discrete Token LLM based TTS System
2026
Top co-authors
Jianhua Tao
· 15
Zhengkun Tian
· 13
Zhengqi Wen
· 13
Shuai Zhang
· 10
Yuxuan Wang
· 4
Lu Lu
· 3
Yong Ren
· 3
Yuping Wang
· 3
Bo Wang
· 2
Chen Shen
· 2
Tao Wang
· 2
Xiaorui Wang
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Generation
Multimodal Audio
Audio Understanding
cs.SD
Voice Cloning
cs.AI
eess.AS