Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Olivier Siohan — most-cited papers & profile · Speech Audio
← authors
·
overview
Olivier Siohan
9
papers ·
15
citations ·
22
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Best of Both Worlds: Multi-task Audio-Visual Automatic Speech Recognition and Active Speaker Detection
2022 · 9 citations
End-to-end multi-talker audio-visual ASR using an active speaker attention module
2022 · 3 citations
Bridging the gap between streaming and non-streaming ASR systems bydistilling ensembles of CTC and RNN-T models
2021 · 2 citations
Revisiting the Entropy Semiring for Neural Speech Recognition
2023 · 1 citations
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
2025
End-to-End Multi-Person Audio/Visual Automatic Speech Recognition
2022
A Closer Look at Audio-Visual Multi-Person Speech Recognition and Active Speaker Selection
2022
Cascaded encoders for fine-tuning ASR models on overlapped speech
2023
Audio-visual fine-tuning of audio-only ASR models
2023
Top co-authors
Otavio Braga
· 4
Oscar Chang
· 2
Ankit Parag Shah
· 1
Avner May
· 1
Chung-Cheng Chiu
· 1
Dmitriy Serdyuk
· 1
Dongseong Hwang
· 1
Florian Metze
· 1
Hank Liao
· 1
Liangliang Cao
· 1
Li Wan
· 1
Ming Sun
· 1
Topics
Speech Recognition
Multimodal Audio
Speech Translation
Speaker Analysis