Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Jan \v{C}ernock\'y — most-cited papers & profile · Speech Audio
← authors
·
overview
Jan \v{C}ernock\'y
18
papers ·
3
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Neural Target Speech Extraction: An Overview
2023 · 3 citations
SE-DiCoW: Self-Enrolled Diarization-Conditioned Whisper
2026
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
2025
Unsupervised Speech Enhancement using Data-defined Priors
2025
DeCRED: Decoder-Centric Regularization for Encoder-Decoder Based Speech Recognition
2025
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
2025
Approaching Dialogue State Tracking via Aligning Speech Encoders and LLMs
2025
BUT System for the MLC-SLM Challenge
2025
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
2025
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
2025
Semi-supervised Sequence-to-sequence ASR using Unpaired Speech and Text
2019
Revisiting joint decoding based multi-talker speech recognition with DNN acoustic model
2021
MultiSV: Dataset for Far-Field Multi-Channel Speaker Verification
2021
Parameter-efficient transfer learning of pre-trained Transformer models for speaker verification using adapters
2022
Improving Speaker Verification with Self-Pretrained Transformer Models
2023
Top co-authors
Luk\'a\v{s} Burget
· 12
Old\v{r}ich Plchot
· 6
Alexander Polok
· 5
Santosh Kesiraju
· 5
Bolaji Yusuf
· 3
Dominik Klement
· 3
Junyi Peng
· 3
Ladislav Mo\v{s}ner
· 3
Marc Delcroix
· 3
Tsubasa Ochiai
· 3
\v{S}imon Sedl\'a\v{c}ek
· 3
J\'an \v{S}vec
· 2
Topics
Speech Recognition
Speaker Analysis
Audio Understanding
Speech Translation
Speech Enhancement
Text-to-Speech
Multimodal Audio