Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hung-yi Lee — most-cited papers & profile · Speech Audio
← authors
·
overview
Hung-yi Lee
176
papers ·
981
citations ·
43
h-index
National Taiwan University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Audio Word2Vec: Unsupervised Learning of Audio Segment Representations using Sequence-to-sequence Autoencoder
2016 · 194 citations
Multi-target Voice Conversion without Parallel Data by Adversarially Learning Disentangled Audio Representations
2018 · 135 citations
Segmental Audio Word2Vec: Representing Utterances as Sequences of Vectors with Applications in Spoken Term Detection
2018 · 51 citations
SUPERB: Speech processing Universal PERformance Benchmark
2021 · 51 citations
Audio ALBERT: A Lite BERT for Self-supervised Learning of Audio Representation
2020 · 39 citations
SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering
2019 · 29 citations
Rhythm-Flexible Voice Conversion without Parallel Data Using Cycle-GAN over Phoneme Posteriorgram Sequences
2018 · 25 citations
Towards Robust Neural Vocoding for Speech Generation: A Survey
2019 · 23 citations
One-shot Voice Conversion by Separating Speaker and Content Representations with Instance Normalization
2019 · 19 citations
DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden-unit BERT
2021 · 18 citations
Noise Adaptive Speech Enhancement using Domain Adversarial Training
2018 · 16 citations
Ensemble knowledge distillation of self-supervised speech models
2023 · 16 citations
SpeechPrompt v2: Prompt Tuning for Speech Classification Tasks
2023 · 16 citations
Completely Unsupervised Speech Recognition By A Generative Adversarial Network Harmonized With Iteratively Refined Hidden Markov Models
2019 · 15 citations
VQVC+: One-Shot Voice Conversion by Vector Quantization and U-Net architecture
2020 · 15 citations
Top co-authors
Shinji Watanabe
· 20
Lin-shan Lee
· 19
Shang-Wen Li
· 18
Guan-Ting Lin
· 13
Haibin Wu
· 13
Kai-Wei Chang
· 12
Sung-Feng Huang
· 12
Andy T. Liu
· 10
Kuan-Po Huang
· 10
Heng-Jui Chang
· 9
Jiatong Shi
· 9
Yu-Kuan Fu
· 9
Topics
Speech Recognition
Audio Understanding
Speech Translation
Text-to-Speech
Audio Generation
Speech Enhancement
Speaker Analysis
Multimodal Audio
Voice Cloning
Music Generation