Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chien-yu Huang — most-cited papers & profile · Speech Audio
← authors
·
overview
Chien-yu Huang
10
papers ·
50
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
2025 · 46 citations
Toward Degradation-Robust Voice Conversion
2021 · 3 citations
Investigating on Incorporating Pretrained and Learnable Speaker Representations for Multi-Speaker Multi-Style Text-to-Speech
2021 · 1 citations
How Far Are We from Robust Voice Conversion: A Survey
2020
Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
2023
Prompting and Adapter Tuning for Self-supervised Encoder-Decoder Speech Model
2023
SpeechCaps: Advancing Instruction-Based Universal Speech Models with Multi-Talker Speaking Style Captioning
2024
Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
2024
Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition
2024
Top co-authors
Hung-yi Lee
· 8
Chi-Yuan Hsiao
· 3
Kai-Wei Chang
· 3
Ke-Han Lu
· 3
Shinji Watanabe
· 3
Chee-En Yu
· 2
Chung-Ming Chien
· 2
Chun-Yi Kuan
· 2
Haibin Wu
· 2
Jheng-hao Lin
· 2
Jiatong Shi
· 2
Kai-Wei Chang
· 2
Topics
Audio Understanding
Speech Recognition
Multimodal Audio
Voice Cloning
Speaker Analysis
Text-to-Speech
Audio Generation
Speech Enhancement
Speech Translation