Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jixun Yao — most-cited papers & profile · Speech Audio
← authors
·
overview
Jixun Yao
8
papers ·
57
citations ·
7
h-index
Northwestern Polytechnical University · Northwestern Polytechnic University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DiffRhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
2025 · 51 citations
MeanVC: Lightweight and Streaming Zero-Shot Voice Conversion via Mean Flows
2025 · 4 citations
PSCodec: A Series of High-Fidelity Low-bitrate Neural Speech Codecs Leveraging Prompt Encoders
2024 · 2 citations
DiffRhythm 2: Efficient and High Fidelity Song Generation via Block Flow Matching
2025
Aligning Generative Speech Enhancement with Perceptual Feedback
2025
S2ST-Omni: Hierarchical Language-Aware SpeechLLM Adaptation for Multilingual Speech-to-Speech Translation
2025
The NPU-HWC System for the ISCSLP 2024 Inspirational and Convincing Audio Generation Challenge
2024
KALL-E:Autoregressive Speech Synthesis with Next-Distribution Prediction
2024
Top co-authors
Lei Xie
· 3
Yuepeng Jiang
· 3
Ziqian Ning
· 3
Guobin Ma
· 2
Huakang Chen
· 2
Jianjun Zhao
· 2
Kangxiang Xia
· 2
Lei Ma
· 2
Xinfa Zhu
· 2
Yuguang Yang
· 2
Chunbo Hao
· 1
Dake Guo
· 1
Topics
Audio Generation
Voice Cloning
Music Generation
Speech Enhancement
Text-to-Speech
Multimodal Audio
Speech Translation
Speech Recognition
Speaker Analysis