Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Haizhou Li โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Haizhou Li
137
papers ยท
395
citations ยท
74
h-index
National University of Singapore ยท Shenzhen Research Institute of Big Data ยท Chinese University of Hong Kong, Shenzhen
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
2020 ยท 26 citations
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
2024 ยท 24 citations
Expressive TTS Training with Frame and Style Reconstruction Loss
2020 ยท 20 citations
Leveraging Acoustic and Linguistic Embeddings from Pretrained speech and language Models for Intent Classification
2021 ยท 19 citations
SpEx+: A Complete Time Domain Speaker Extraction Network
2020 ยท 18 citations
Converting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion
2020 ยท 15 citations
Seen and Unseen emotional style transfer for voice conversion with a new emotional speech dataset
2020 ยท 14 citations
On the End-to-End Solution to Mandarin-English Code-switching Speech Recognition
2018 ยท 13 citations
GraphSpeech: Syntax-Aware Graph Attention Network For Neural Speech Synthesis
2020 ยท 13 citations
Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
2022 ยท 13 citations
Generative x-vectors for text-independent speaker verification
2018 ยท 11 citations
A Modularized Neural Network with Language-Specific Output Layers for Cross-lingual Voice Conversion
2019 ยท 11 citations
VAW-GAN for Singing Voice Conversion with Non-parallel Training Data
2020 ยท 10 citations
Self-and-Mixed Attention Decoder with Deep Acoustic Structure for Transformer-based LVCSR
2020 ยท 9 citations
Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet
2019 ยท 8 citations
Top co-authors
Berrak Sisman
ยท 21
Ruijie Tao
ยท 11
Junyi Ao
ยท 10
Eng Siong Chng
ยท 9
Chenglin Xu
ยท 8
Kong Aik Lee
ยท 8
Kun Zhou
ยท 8
Rohan Kumar Das
ยท 8
Xiaoxue Gao
ยท 8
Zexu Pan
ยท 8
Emre Y{\i}lmaz
ยท 7
Mingyang Zhang
ยท 7
Topics
Speech Recognition
Audio Understanding
Audio Generation
Speaker Analysis
Text-to-Speech
Speech Enhancement
Multimodal Audio
Speech Translation
Voice Cloning
Music Generation