Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sung-Feng Huang — most-cited papers & profile · Speech Audio
← authors
·
overview
Sung-Feng Huang
19
papers ·
57
citations ·
8
h-index
Nvidia (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Towards Unsupervised Automatic Speech Recognition Trained by Unaligned Speech and Text only
2018 · 14 citations
Almost-unsupervised Speech Recognition with Close-to-zero Resource Based on Phonetic Structures Learned from Very Small Unpaired Speech and Text Data
2018 · 14 citations
Non-autoregressive Mandarin-English Code-switching Speech Recognition
2021 · 9 citations
Meta-TTS: Meta-Learning for Few-Shot Speaker Adaptive Text-to-Speech
2021 · 8 citations
SpeechNet: A Universal Modularized Model for Speech Processing Tasks
2021 · 6 citations
Improved Audio Embeddings By Adjacency-based Clustering With Applications In Spoken Term Detection
2018 · 5 citations
From Semi-supervised to Almost-unsupervised Speech Recognition with Very-low Resource by Jointly Learning Phonetic Structures from Audio and Text Embeddings
2019 · 1 citations
One Model, Many Latencies: Universal Speech Enhancement for Diverse Real-Time Applications
2026
MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translation
2026
Joint Fullband-Subband Modeling for High-Resolution SingFake Detection
2026
VIBE: Voice-Induced open-ended Bias Evaluation for Large Audio-Language Models via Real-World Speech
2026
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
2025
HighRateMOS: Sampling-Rate Aware Modeling for Speech Quality Assessment
2025
Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training
2020
Few-Shot Cross-Lingual TTS Using Transferable Phoneme Embedding
2022
Top co-authors
Hung-yi Lee
· 12
Yi-Chen Chen
· 5
Da-Rong Liu
· 4
Lin-shan Lee
· 3
Shun-Po Chuang
· 3
Szu-Wei Fu
· 3
Chia-Hao Shen
· 2
Hung-yi Lee
· 2
Wei-Ping Huang
· 2
Wenze Ren
· 2
Yi-Cheng Lin
· 2
Alexander H. Liu
· 1
Topics
Speech Recognition
Text-to-Speech
Audio Understanding
Speech Translation
Speech Enhancement
Multimodal Audio
cs.SD
eess.AS
Speaker Analysis
Audio Generation