Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zejun MA — most-cited papers & profile · Speech Audio
← authors
·
overview
Zejun MA
22
papers ·
50
citations ·
20
h-index
QT Ultrasound (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023 · 16 citations
Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis
2023 · 8 citations
LiteG2P: A fast, light and high accuracy model for grapheme-to-phoneme conversion
2023 · 5 citations
PolyVoice: Language Models for Speech to Speech Translation
2023 · 5 citations
Improving Pseudo-label Training For End-to-end Speech Recognition Using Gradient Mask
2021 · 3 citations
Towards High-fidelity Singing Voice Conversion with Acoustic Reference and Contrastive Predictive Coding
2021 · 2 citations
Language-specific Acoustic Boundary Learning for Mandarin-English Code-switching Speech Recognition
2023 · 2 citations
HTS-AT: A Hierarchical Token-Semantic Audio Transformer for Sound Classification and Detection
2022 · 1 citations
S3T: Self-Supervised Pre-training with Swin Transformer for Music Classification
2022 · 1 citations
Direct Speech-to-speech Translation without Textual Annotation using Bottleneck Features
2022 · 1 citations
Improving End-to-End Contextual Speech Recognition with Fine-Grained Contextual Knowledge Selection
2022
The Volcspeech system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge
2022
Language Adaptive Cross-lingual Speech Representation Learning with Sparse Sharing Sub-networks
2022
Random Utterance Concatenation Based Data Augmentation for Improving Short-video Speech Recognition
2022
Token-level Speaker Change Detection Using Speaker Difference and Speech Content via Continuous Integrate-and-fire
2022
Top co-authors
Xiang Yin
· 5
Chen Shen
· 3
Chunfeng Wang
· 3
Linhao Dong
· 3
Meng Cai
· 3
Zhenlin Liang
· 3
Bo Xu
· 2
Chen Zhang
· 2
Jinglin Liu
· 2
Lu Lu
· 2
Shiyu Zhou
· 2
Wei Li
· 2
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Speech Translation
Voice Cloning
Audio Generation
Music Generation
Speaker Analysis
Speech Enhancement
Multimodal Audio