Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Joseph Keshet — most-cited papers & profile · Speech Audio
← authors
·
overview
Joseph Keshet
22
papers ·
7
citations ·
29
h-index
Technion – Israel Institute of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
FlowTSE: Target Speaker Extraction with Flow Matching
2025 · 2 citations
Dr.VOT : Measuring Positive and Negative Voice Onset Time in the Wild
2019 · 2 citations
DiffAR: Denoising Diffusion Autoregressive Model for Raw Speech Waveform Generation
2023 · 1 citations
Tradition or Innovation: A Comparison of Modern ASR Methods for Forced Alignment
2024 · 1 citations
WhisperNER: Unified Open Named Entity and Speech Recognition
2024 · 1 citations
Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming
2026
Drax: Speech Recognition with Discrete Flow Matching
2025
UmbraTTS: Adapting Text-to-Speech to Environmental Contexts with Flow Matching
2025
Sequence Segmentation Using Joint RNN and Structured Prediction Models
2016
Domain Adaptation For Formant Estimation Using Deep Learning
2016
Phoneme Boundary Detection using Learnable Segmental Features
2020
CNN-based Spoken Term Detection and Localization without Dynamic Programming
2021
DeepFry: Identifying Vocal Fry Using Deep Neural Networks
2022
Correcting Mispronunciations in Speech using Spectrogram Inpainting
2022
Unsupervised Word Segmentation using K Nearest Neighbors
2022
Top co-authors
Aviv Navon
· 7
Aviv Shamsian
· 7
Gill Hetz
· 5
Neta Glazer
· 5
Yael Segal
· 4
Felix Kreuk
· 3
Matthew Goldrick
· 3
Yael Segal-Feldman
· 3
Yehoshua Dissen
· 3
Yosi Shrem
· 3
Bronya R. Chernyak
· 2
Eleanor Chodroff
· 2
Topics
Speech Recognition
Audio Understanding
Speaker Analysis
Speech Translation
Speech Enhancement
Text-to-Speech
Audio Generation
Multimodal Audio
Voice Cloning
Music Generation