Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shiyin Kang — most-cited papers & profile · Speech Audio
← authors
·
overview
Shiyin Kang
22
papers ·
136
citations ·
21
h-index
Sensimetrics Corporation
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DurIAN: Duration Informed Attention Network For Multimodal Synthesis
2019 · 94 citations
Maximizing Mutual Information for Tacotron
2019 · 13 citations
Towards Multi-Scale Speaking Style Modelling with Hierarchical Context Information for Mandarin Speech Synthesis
2022 · 11 citations
Audio-visual Recognition of Overlapped speech for the LRS2 dataset
2020 · 10 citations
Speaker Independent and Multilingual/Mixlingual Speech-Driven Talking Head Generation Using Phonetic Posteriorgrams
2020 · 3 citations
Transferring Source Style in Non-Parallel Voice Conversion
2020 · 2 citations
Disentangleing Content and Fine-grained Prosody Information via Hybrid ASR Bottleneck Features for Voice Conversion
2022 · 1 citations
MSStyleTTS: Multi-Scale Style Modeling with Hierarchical Context Information for Expressive Speech Synthesis
2023 · 1 citations
SCNet: Sparse Compression Network for Music Source Separation
2024 · 1 citations
SemaVoice: Semantic-Aware Continuous Autoregressive Speech Synthesis
2026
Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model
2026
VAENAR-TTS: Variational Auto-Encoder based Non-AutoRegressive Text-to-Speech Synthesis
2021
FullSubNet+: Channel Attention FullSubNet with Complex Spectrograms for Speech Enhancement
2022
Towards Expressive Speaking Style Modelling with Hierarchical Context Information for Mandarin Speech Synthesis
2022
Context-aware Coherent Speaking Style Prediction with Hierarchical Transformers for Audiobook Speech Synthesis
2023
Top co-authors
Helen Meng
· 16
Zhiyong Wu
· 13
Shun Lei
· 8
Deyi Tuo
· 5
Dong Yu
· 5
Liyang Chen
· 5
Yixuan Zhou
· 5
Dan Su
· 4
Xixin Wu
· 4
Jun Chen
· 3
Xunying Liu
· 3
Guangzhi Lei
· 2
Topics
Audio Generation
Text-to-Speech
Speech Recognition
Music Generation
Speaker Analysis
Multimodal Audio
Voice Cloning
Audio Understanding
Speech Enhancement