Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tian-Hao Zhang — most-cited papers & profile · Speech Audio
← authors
·
overview
Tian-Hao Zhang
10
papers ·
3
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Breaking Through the Spike: Spike Window Decoding for Accelerated and Precise Automatic Speech Recognition
2025 · 1 citations
Non-autoregressive Transformer with Unified Bidirectional Decoder for Automatic Speech Recognition
2021 · 1 citations
InterFormer: Interactive Local and Global Features Fusion for Automatic Speech Recognition
2023 · 1 citations
IKFST: IOO and KOO Algorithms for Accelerated and Precise WFST-based End-to-End Automatic Speech Recognition
2026
SEAL: Speech Embedding Alignment Learning for Speech Large Language Model with Retrieval-Augmented Generation
2025
FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles
2025
Improving Zero-Shot Chinese-English Code-Switching ASR with kNN-CTC and Gated Monolingual Datastores
2024
Rethinking Speech Recognition with A Multimodal Perspective via Acoustic and Semantic Cooperative Decoding
2023
CIF-T: A Novel CIF-based Transducer Architecture for Automatic Speech Recognition
2023
I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception
2024
Top co-authors
Xu-Cheng Yin
· 6
Xinyuan Qian
· 5
Feng Chen
· 3
Song-Lu Chen
· 3
Dinghao Zhou
· 2
Jiawei Zhang
· 2
Jun Wang
· 2
Anbin Qi
· 1
Baoxiang Li
· 1
Bingyu Liu
· 1
Chao Luo
· 1
Chao Luo
· 1
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Multimodal Audio
Audio Generation
Music Generation
Audio Understanding