Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sining Sun — most-cited papers & profile · Speech Audio
← authors
·
overview
Sining Sun
13
papers ·
49
citations ·
13
h-index
Northwestern Polytechnical University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Domain Adversarial Training for Accented Speech Recognition
2018 · 15 citations
Training Augmentation with Adversarial Examples for Robust Speech Recognition
2018 · 14 citations
Multi-head Monotonic Chunkwise Attention For Online Speech Recognition
2020 · 13 citations
Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning
2021 · 13 citations
Efficient Conformer with Prob-Sparse Attention Mechanism for End-to-EndSpeech Recognition
2021 · 5 citations
Tiny Transducer: A Highly-efficient Speech Recognition Model on Edge Devices
2021 · 2 citations
Investigating Generative Adversarial Networks based Speech Dereverberation for Robust Speech Recognition
2018
Conversational Speech Recognition By Learning Conversation-level Characteristics
2022
Leveraging Acoustic Contextual Representation by Audio-textual Cross-modal Learning for Conversational ASR
2022
Key Frame Mechanism For Efficient Conformer Based End-to-end Speech Recognition
2023
Self-Supervised Disentangled Representation Learning for Robust Target Speech Extraction
2023
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition
2024
Streaming Decoder-Only Automatic Speech Recognition with Discrete Speech Units: A Pilot Study
2024
Top co-authors
Lei Xie
· 7
Long Ma
· 6
Qing Yang
· 4
Changhao Shan
· 3
Yike Zhang
· 3
Ching-Feng Yeh
· 2
Kun Wei
· 2
Mari Ostendorf
· 2
Mei-Yuh Hwang
· 2
Peng Fan
· 2
Songjun Cao
· 2
Baiji Liu
· 1
Topics
Speech Recognition
Speech Translation
Speech Enhancement
Audio Understanding
Multimodal Audio
Speaker Analysis
Text-to-Speech