Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Bo Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Bo Li
69
papers ·
1157
citations ·
44
h-index
North University of China · Northwestern Polytechnical University · Henan University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023 · 112 citations
Controllable Time-Delay Transformer for Real-Time Punctuation Prediction and Disfluency Detection
2020 · 33 citations
Efficient Domain Adaptation for Speech Foundation Models
2023 · 19 citations
Massively Multilingual Shallow Fusion with Large Language Models
2023 · 14 citations
A Language Agnostic Multilingual Streaming On-Device ASR System
2022 · 10 citations
Bytes are All You Need: End-to-End Multilingual Speech Recognition and Synthesis with Bytes
2018 · 9 citations
Improving the fusion of acoustic and text representations in RNN-T
2022 · 9 citations
JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition
2023 · 8 citations
UML: A Universal Monolingual Output Layer for Multilingual ASR
2023 · 5 citations
Large vocabulary speech recognition for languages of Africa: multilingual modeling and self-supervised learning
2022 · 2 citations
Exploring Speech Enhancement with Generative Adversarial Networks for Robust Speech Recognition
2017 · 1 citations
JOIST: A Joint Speech and Text Streaming Model For ASR
2022 · 1 citations
Massive End-to-end Models for Short Search Queries
2023 · 1 citations
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
2025
Turn-Taking Prediction for Natural Conversational Speech
2022
Top co-authors
Tara N. Sainath
· 15
Trevor Strohman
· 9
Rohit Prabhavalkar
· 8
Yu Zhang
· 7
Shuo-yiin Chang
· 6
Dongseong Hwang
· 4
Khe Chai Sim
· 4
Tara Sainath
· 4
Weiran Wang
· 4
Ke Hu
· 3
Yanzhang He
· 3
Zhong Meng
· 3
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Speech Enhancement
Multimodal Audio
Audio Generation
Music Generation
Speaker Analysis