Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Siqi Zheng — most-cited papers & profile · Speech Audio
← authors
·
overview
Siqi Zheng
16
papers ·
62
citations ·
51
h-index
East China Jiaotong University · Japan Real Estate Institute
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
2023 · 18 citations
Speaker Overlap-aware Neural Diarization for Multi-party Meeting Analysis
2022 · 15 citations
M2MeT: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge
2021 · 11 citations
3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
2023 · 5 citations
An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
2024 · 4 citations
CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking
2023 · 3 citations
CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
2024 · 3 citations
Improving Speaker Diarization using Semantic Information: Joint Pairwise Constraints Propagation
2023 · 2 citations
Speaker Embedding-aware Neural Diarization: an Efficient Framework for Overlapping Speech Diarization in Meeting Scenarios
2022 · 1 citations
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
2024
Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
2022
Graph Convolutional Network Based Semi-Supervised Learning on Multi-Speaker Meeting Data
2022
Contextual Expressive Text-to-Speech
2022
FunCodec: A Fundamental, Reproducible and Integrable Open-source Toolkit for Neural Speech Codec
2023
Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR
2023
Top co-authors
Shiliang Zhang
· 10
Zhihao Du
· 8
Qian Chen
· 7
Zhijie Yan
· 6
Kai Hu
· 4
Luyao Cheng
· 4
Yafeng Chen
· 4
Hui Wang
· 3
Zhifu Gao
· 3
Ziyang Ma
· 3
Bin Ma
· 2
Chang Zhou
· 2
Topics
Speech Recognition
Speaker Analysis
Audio Understanding
Audio Generation
Text-to-Speech
Multimodal Audio
Speech Enhancement
Voice Cloning
Speech Translation