Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Vimal Manohar — most-cited papers & profile · Speech Audio
← authors
·
overview
Vimal Manohar
14
papers ·
176
citations ·
21
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
2020 · 97 citations
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
2023 · 45 citations
On lattice-free boosted MMI training of HMM and CTC-based full-context ASR models
2021 · 11 citations
Accent-Robust Automatic Speech Recognition Using Supervised and Unsupervised Wav2vec Embeddings
2021 · 11 citations
Automatic Speech Recognition and Topic Identification for Almost-Zero-Resource Languages
2018 · 6 citations
Acoustic data-driven lexicon learning based on a greedy pronunciation selection framework
2017 · 5 citations
Kaizen: Continuously improving teacher using Exponential Moving Average for semi-supervised speech recognition
2021 · 1 citations
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment
2025
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
2025
Using of heterogeneous corpora for training of an ASR system
2017
Large scale weakly and semi-supervised learning for low-resource video ASR
2020
Towards zero-shot Text-based voice editing using acoustic context conditioning, utterance embeddings, and reference encoders
2022
Self-Supervised Representations for Singing Voice Conversion
2023
Acoustic modeling for Overlapping Speech Recognition: JHU Chime-5 Challenge System
2024
Top co-authors
Sanjeev Khudanpur
· 5
Yatharth Saraf
· 4
Jan Trmal
· 3
Xiaohui Zhang
· 3
Abdelrahman Mohamed
· 2
Daniel Povey
· 2
Frank Zhang
· 2
Geoffrey Zweig
· 2
Jilong Wu
· 2
Julian Chan
· 2
Kainan Peng
· 2
Mike Seltzer
· 2
Topics
Speech Recognition
Speech Translation
Audio Generation
Voice Cloning
Text-to-Speech
Speech Enhancement
Music Generation
Audio Understanding
Speaker Analysis