Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Kai-Wei Chang — most-cited papers & profile · Speech Audio
← authors
·
overview
Kai-Wei Chang
49
papers ·
3923
citations ·
53
h-index
University of California, Los Angeles · Los Angeles Medical Center
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Ensemble knowledge distillation of self-supervised speech models
2023 · 16 citations
SpeechPrompt v2: Prompt Tuning for Speech Classification Tasks
2023 · 16 citations
SpeechNet: A Universal Modularized Model for Speech Processing Tasks
2021 · 6 citations
MiniSUPERB: Lightweight Benchmark for Self-supervised Speech Models
2023 · 6 citations
Towards audio language modeling -- an overview
2024 · 4 citations
Toward Degradation-Robust Voice Conversion
2021 · 3 citations
EMO-SUPERB: An In-depth Look at Speech Emotion Recognition
2024 · 2 citations
Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks
2023 · 1 citations
Towards Unsupervised Speech Recognition at the Syllable-Level
2025
SUPERB @ SLT 2022: Challenge on Generalization and Efficiency of Self-Supervised Speech Representation Learning
2022
ABC-KD: Attention-Based-Compression Knowledge Distillation for Deep Learning-Based Noise Suppression
2023
Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
2023
Towards General-Purpose Text-Instruction-Guided Voice Conversion
2023
Prompting and Adapter Tuning for Self-supervised Encoder-Decoder Speech Model
2023
Top co-authors
Hung-yi Lee
· 12
Haibin Wu
· 4
Shang-Wen Li
· 4
Chien-yu Huang
· 3
Chun-Yi Kuan
· 2
Ho-Lam Chung
· 2
Jiatong Shi
· 2
Shinji Watanabe
· 2
Tsu-Yuan Hsu
· 2
Tzu-hsun Feng
· 2
Wei-Cheng Tseng
· 2
Abdelrahman Mohamed
· 1
Topics
Speech Recognition
Audio Understanding
Text-to-Speech
Speech Enhancement
Audio Generation
Voice Cloning
Speech Translation
Speaker Analysis
Multimodal Audio