Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Daniel Povey — most-cited papers & profile · Speech Audio
← authors
·
overview
Daniel Povey
32
papers ·
357
citations ·
61
h-index
Xiaomi (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
2021 · 201 citations
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
2020 · 97 citations
Zipformer: A faster and better encoder for automatic speech recognition
2023 · 28 citations
Lhotse: a speech data representation library for the modern deep learning ecosystem
2021 · 6 citations
Acoustic data-driven lexicon learning based on a greedy pronunciation selection framework
2017 · 5 citations
GPU-accelerated Guided Source Separation for Meeting Transcription
2022 · 5 citations
Frustratingly Easy Noise-aware Training of Acoustic Models
2020 · 4 citations
Speaker Diarization with Region Proposal Network
2020 · 2 citations
speechocean762: An Open-Source Non-native English Speech Corpus For Pronunciation Assessment
2021 · 2 citations
Alternative Pseudo-Labeling for Semi-Supervised Automatic Speech Recognition
2023 · 2 citations
PyChain: A Fully Parallelized PyTorch Implementation of LF-MMI for End-to-End ASR
2020 · 1 citations
Delay-penalized transducer for low-latency streaming ASR
2022 · 1 citations
SURT 2.0: Advances in Transducer-based Multi-talker Speech Recognition
2023 · 1 citations
Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
2023 · 1 citations
Learning from Flawed Data: Weakly Supervised Automatic Speech Recognition
2023 · 1 citations
Top co-authors
Sanjeev Khudanpur
· 16
Zengwei Yao
· 10
Wei Kang
· 9
Liyong Guo
· 8
Long Lin
· 8
Fangjun Kuang
· 7
Yifan Yang
· 7
Desh Raj
· 6
Xiaoyu Yang
· 6
Dongji Gao
· 4
Fangjun Kuang
· 4
Zengrui Jin
· 4
Topics
Speech Recognition
Speech Translation
Audio Understanding
Text-to-Speech
Speech Enhancement
Speaker Analysis
Audio Generation
Music Generation
Multimodal Audio