Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jian Zhu — most-cited papers & profile · Speech Audio
← authors
·
overview
Jian Zhu
15
papers ·
267
citations ·
6
h-index
University of Science and Technology of China · University of British Columbia · State Grid Corporation of China (China) · Zhejiang Lab
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
POWSM: A Phonetic Open Whisper-Style Speech Foundation Model
2025 · 7 citations
Phone-to-audio alignment without text: A Semi-supervised Approach
2021 · 2 citations
Developing multilingual speech synthesis system for Ojibwe, Mi'kmaq, and Maliseet
2025 · 1 citations
Internalizing ASR with Implicit Chain of Thought for Efficient Speech-to-Speech Conversational LLM
2024 · 1 citations
Synchronising speech segments with musical beats in Mandarin and English singing
2021
Bootstrapping meaning through listening: Unsupervised learning of spoken sentence embeddings
2022
The taste of IPA: Towards open-vocabulary keyword spotting and forced alignment in any language
2023
Top co-authors
Changbing Yang
· 2
Cong Zhang
· 2
Chad Quinn
· 1
Chia-wen Lo
· 1
Chin-Jou Li
· 1
Christopher Hammerly
· 1
Cong Zhang
· 1
David Jurgens
· 1
David Mortensen
· 1
Eunjung Yeo
· 1
Farhan Samir
· 1
Jahurul Islam
· 1
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Translation
Music Generation
Speech Enhancement
Multimodal Audio