Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shiliang Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shiliang Zhang
38
papers ·
133
citations ·
0
h-index
Xihua University · Xi'an University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
2023 · 18 citations
Speaker Overlap-aware Neural Diarization for Multi-party Meeting Analysis
2022 · 15 citations
Deep-FSMN for Large Vocabulary Continuous Speech Recognition
2018 · 13 citations
Streaming Chunk-Aware Multihead Attention for Online End-to-End Speech Recognition
2020 · 13 citations
M2MeT: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge
2021 · 11 citations
Universal ASR: Unifying Streaming and Non-Streaming ASR Using a Single Encoder-Decoder Model
2020 · 10 citations
Automatic Spelling Correction with Transformer for CTC-based End-to-End Speech Recognition
2019 · 9 citations
SAN-M: Memory Equipped Self-Attention for End-to-End Speech Recognition
2020 · 9 citations
Deep Feed-forward Sequential Memory Networks for Speech Synthesis
2018 · 8 citations
FunASR: A Fundamental End-to-End Speech Recognition Toolkit
2023 · 4 citations
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 · 4 citations
Neural Zero-Inflated Quality Estimation Model For Automatic Speech Recognition System
2019 · 3 citations
CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
2024 · 3 citations
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
2025 · 2 citations
Paraformer: Fast and Accurate Parallel Transformer for Non-autoregressive End-to-End Speech Recognition
2022 · 2 citations
Top co-authors
Zhihao Du
· 15
Qian Chen
· 10
Siqi Zheng
· 10
Ming Lei
· 9
Zhifu Gao
· 9
Zhijie Yan
· 9
Fan Yu
· 7
Lei Xie
· 6
Zhijie Yan
· 6
Ziyang Ma
· 6
Guanrou Yang
· 4
Ian McLoughlin
· 4
Topics
Speech Recognition
Text-to-Speech
Speech Translation
Audio Understanding
Speaker Analysis
Audio Generation
Multimodal Audio
Speech Enhancement
Voice Cloning