Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Alexander H. Liu — most-cited papers & profile · Speech Audio
← authors
·
overview
Alexander H. Liu
16
papers ·
29
citations ·
12
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition
2021 · 12 citations
DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation Learning
2023 · 6 citations
Towards audio language modeling -- an overview
2024 · 4 citations
End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training
2020 · 2 citations
Semi-supervised Learning for Multi-speaker Text-to-speech Synthesis Using Discrete Speech Representation
2020 · 2 citations
Worse WER, but Better BLEU? Leveraging Word Embedding as Intermediate in Multitask End-to-End Speech Translation
2020 · 2 citations
Towards Unsupervised Speech Recognition and Synthesis with Quantized Speech Representation Learning
2019 · 1 citations
USAD: Universal Speech and Audio Representation via Distillation
2025
UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation
2025
Adversarial Training of End-to-end Speech Recognition Using a Criticizing Language Model
2018
Sequence-to-sequence Automatic Speech Recognition with Word Embedding Regularization and Fused Decoding
2019
Non-Autoregressive Predictive Coding for Learning Speech Representations from Local Dependencies
2020
On the Interplay Between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis
2021
Self-supervised Fine-tuning for Improved Content Representations by Speaker-invariant Clustering
2023
Generative Pre-training for Speech with Flow Matching
2023
Top co-authors
Hung-yi Lee
· 7
James Glass
· 6
Heng-Jui Chang
· 4
Lin-shan Lee
· 4
Cheng-I Jeff Lai
· 2
David Cox
· 2
James R. Glass
· 2
Kaizhi Qian
· 2
Shiyu Chang
· 2
Shun-Po Chuang
· 2
Tzu-Wei Sung
· 2
Wei-Ning Hsu
· 2
Topics
Speech Recognition
Speech Translation
Text-to-Speech
Audio Understanding
Speech Enhancement
Audio Generation
Speaker Analysis
Multimodal Audio