Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Huaming Wang — most-cited papers & profile · Speech Audio
← authors
·
overview
Huaming Wang
14
papers ·
352
citations ·
19
h-index
Ningbo University · Microsoft (United States) · Ningbo University of Technology · Chinese Academy of Sciences · Beihang University · Nanjing University of Aeronautics and Astronautics
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
2023 · 163 citations
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation
2022 · 32 citations
Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
2023 · 25 citations
Human Listening and Live Captioning: Multi-Task Training for Speech Enhancement
2021 · 24 citations
Pengi: An Audio Language Model for Audio Tasks
2023 · 20 citations
Cracking the cocktail party problem by multi-beam deep attractor network
2018 · 3 citations
Real-Time Joint Personalized Speech Enhancement and Acoustic Echo Cancellation
2022 · 2 citations
Personalized Speech Enhancement: New Models and Comprehensive Evaluation
2021 · 1 citations
One model to enhance them all: array geometry agnostic multi-channel personalized speech enhancement
2021
Real-Time Audio-Visual End-to-End Speech Enhancement
2023
Natural Language Supervision for General-Purpose Audio Representations
2023
NOTSOFAR-1 Challenge: New Datasets, Baseline, and Tasks for Distant Meeting Transcription
2024
Top co-authors
Sefik Emre Eskimez
· 6
Takuya Yoshioka
· 6
Zhuo Chen
· 6
Min Tang
· 4
Jinyu Li
· 3
Xiaofei Wang
· 3
Benjamin Elizalde
· 2
Chengyi Wang
· 2
Furu Wei
· 2
Hemin Yang
· 2
Lei He
· 2
Long Zhou
· 2
Topics
Speech Enhancement
Speech Recognition
Audio Understanding
Audio Generation
Multimodal Audio
Speaker Analysis
Speech Translation
Text-to-Speech
Voice Cloning