Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Linhao Dong — most-cited papers & profile · Speech Audio
← authors
·
overview
Linhao Dong
14
papers ·
74
citations ·
13
h-index
University of Shanghai for Science and Technology · Nankai University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Syllable-Based Sequence-to-Sequence Speech Recognition with the Transformer in Mandarin Chinese
2018 · 28 citations
A Comparison of Label-Synchronous and Frame-Synchronous End-to-End Models for Speech Recognition
2020 · 16 citations
A Comparison of Modeling Units in Sequence-to-Sequence Speech Recognition with the Transformer on Mandarin Chinese
2018 · 12 citations
CIF: Continuous Integrate-and-Fire for End-to-End Speech Recognition
2019 · 6 citations
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
2024 · 6 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Self-Attention Aligner: A Latency-Control End-to-End Model for ASR Using Self-Attention Network and Chunk-Hopping
2019 · 3 citations
Extending Recurrent Neural Aligner for Streaming End-to-End Speech Recognition in Mandarin
2018 · 2 citations
Language-specific Acoustic Boundary Learning for Mandarin-English Code-switching Speech Recognition
2023 · 2 citations
CIF-based Collaborative Decoding for End-to-end Contextual Speech Recognition
2020
Improving End-to-End Contextual Speech Recognition with Fine-Grained Contextual Knowledge Selection
2022
Token-level Speaker Change Detection Using Speaker Difference and Speech Content via Continuous Integrate-and-fire
2022
CIF-PT: Bridging Speech and Text Representations for Spoken Language Understanding via Continuous Integrate-and-Fire Pre-Training
2023
NEST-RQ: Next Token Prediction for Speech Self-Supervised Pre-Training
2024
Top co-authors
Bo Xu
· 9
Shiyu Zhou
· 7
Jun Zhang
· 5
Lu Lu
· 5
Zejun Ma
· 5
Minglun Han
· 4
Chen Shen
· 3
Shuang Xu
· 3
Zhenlin Liang
· 3
Zhiyun Fan
· 3
Meng Cai
· 2
Mingkun Huang
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speaker Analysis
Multimodal Audio
Speech Enhancement