Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ya Li — most-cited papers & profile · Speech Audio
← authors
·
overview
Ya Li
19
papers ·
23
citations ·
28
h-index
Beijing University of Posts and Telecommunications · Harbin Institute of Technology · Heilongjiang Institute of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improving speech emotion recognition via Transformer-based Predictive Coding through transfer learning
2018 · 13 citations
A pairwise discriminative task for speech emotion recognition
2018 · 3 citations
M2-CTTS: End-to-End Multi-scale Multi-modal Conversational Text-to-Speech Synthesis
2023 · 3 citations
CONCSS: Contrastive-based Context Comprehension for Dialogue-appropriate Prosody in Conversational Speech Synthesis
2023 · 1 citations
Frame-level emotional state alignment method for speech emotion recognition
2023 · 1 citations
Retrieval Augmented Generation in Prompt-based Text-to-Speech Synthesis with Context-Aware Contrastive Language-Audio Pretraining
2024 · 1 citations
AffectCodec: Emotion-Preserving Neural Speech Codec with Block-Diagonal Residual FSQ
2026
Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models
2026
Multi-Loss Learning for Speech Emotion Recognition with Energy-Adaptive Mixup and Frame-Level Attention
2025
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
2025
SynParaSpeech: Automated Synthesis of Paralinguistic Datasets for Speech Generation and Understanding
2025
MGFF-TDNN: A Multi-Granularity Feature Fusion TDNN Model with Depth-Wise Separable Module for Speaker Verification
2025
Speech Emotion Recognition via Contrastive Loss under Siamese Networks
2019
ECAPA-TDNN for Multi-speaker Text-to-speech Synthesis
2022
Rhythm-controllable Attention with High Robustness for Long Sentence Speech Synthesis
2023
Top co-authors
Yingming Gao
· 10
Jinlong Xue
· 7
Yayue Deng
· 7
Jianhua Tao
· 4
Cong Wang
· 3
Jian Huang
· 3
Yizhong Geng
· 3
Zheng Lian
· 3
and Xiaoyu Shen
· 1
Binghuai Lin
· 1
Bo Hu
· 1
Chunfeng Wang
· 1
Topics
Audio Understanding
Text-to-Speech
Audio Generation
Speech Recognition
Speaker Analysis
Speech Enhancement
cs.SD
cs.CL
cs.AI
eess.AS