Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yangze Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Yangze Li
9
papers ·
16
citations ·
3
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Unveiling the Potential of LLM-Based ASR on Chinese Open-Source Datasets
2024 · 12 citations
Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
2024 · 3 citations
A Transcription Prompt-based Efficient Audio Large Language Model for Robust Speech Recognition
2024 · 1 citations
OSUM: Advancing Open Speech Understanding Models with Limited Resources in Academia
2025
CASA-ASR: Context-Aware Speaker-Attributed ASR
2023
BA-SOT: Boundary-Aware Serialized Output Training for Multi-Talker ASR
2023
SA-Paraformer: Non-autoregressive End-to-End Speaker-Attributed ASR
2023
MMGER: Multi-modal and Multi-granularity Generative Error Correction with LLM for Joint Accent and Speech Recognition
2024
Top co-authors
Lei Xie
· 7
Pengcheng Guo
· 4
Fan Yu
· 3
Shiliang Zhang
· 3
Long Ma
· 2
Qian Chen
· 2
Tianyi Xu
· 2
Xiong Wang
· 2
Zhihao Du
· 2
Chaoyou Fu
· 1
He Wang
· 1
Jie Zhang
· 1
Topics
Speech Recognition
Text-to-Speech
Multimodal Audio
Audio Understanding
Speaker Analysis
Speech Translation