Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jun Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Jun Zhang
184
papers ·
1325
citations ·
0
h-index
North China Institute of Science and Technology · Hong Kong University of Science and Technology · North China Institute of Aerospace Engineering
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep LSTM for Large Vocabulary Continuous Speech Recognition
2017 · 23 citations
SALMONN-omni: A Standalone Speech LLM without Codec Injection for Full-duplex Conversation
2025 · 16 citations
DIHARD II is Still Hard: Experimental Results and Discussions from the DKU-LENOVO Team
2020 · 8 citations
Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice
2025 · 6 citations
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
2024 · 6 citations
Improving RNN transducer with normalized jointer network
2020 · 5 citations
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024 · 5 citations
Exponential Moving Average Model in Parallel Speech Recognition Training
2017 · 3 citations
Dynamic latency speech recognition with asynchronous revision
2020 · 2 citations
Language-specific Acoustic Boundary Learning for Mandarin-English Code-switching Speech Recognition
2023 · 2 citations
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
2025 · 1 citations
Enabling Auditory Large Language Models for Automatic Speech Quality Evaluation
2024 · 1 citations
SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
2024 · 1 citations
Anatomy of the Modality Gap: Dissecting the Internal States of End-to-End Speech LLMs
2026
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
2025
Top co-authors
Lu Lu
· 13
Zejun Ma
· 10
Yuxuan Wang
· 9
Xiaohai Tian
· 5
Yang Zhang
· 4
Yi He
· 4
Chao Zhang
· 3
Chen Shen
· 3
Guangzhi Sun
· 3
Siyin Wang
· 3
Wei Li
· 3
Bin Wang
· 2
Topics
Speech Recognition
Audio Understanding
Speech Translation
Text-to-Speech
Speech Enhancement
Speaker Analysis
Multimodal Audio
cs.CL
Audio Generation
eess.AS