DISPLACE Challenge: Diarization Of Speaker And Language In Conversational Environments
2023 Β· Shikha Baghel, Shreyas Ramoji, Sidharth, et al.
Abstract
In multilingual societies, social conversations often involve code-mixed speech. The current speech technology may not be well equipped to extract information from multi-lingual multi-speaker conversations. The DISPLACE challenge entails a first-of-kind task to benchmark speaker and language diarization on the same data, as the data contains multi-speaker conversations in multilingual code-mixed speech. The challenge attempts to highlight outstanding issues in speaker diarization (SD) in multilingual settings with code-mixing. Further, language diarization (LD) in multi-speaker settings also introduces new challenges, where the system has to disambiguate speaker switches with code switches. For this challenge, a natural multilingual, multi-speaker conversational dataset is distributed for development and evaluation purposes. The systems are evaluated on single-channel far-field recordings. We also release a baseline system and report the highlights of the system submissions.
Authors
(none)
Tags
Stats
Related papers
- Summary Of The DISPLACE Challenge 2023 - Diarization Of Speaker And Language In Conversational Environments (2023)0.00
- The Second DISPLACE Challenge : Diarization Of Speaker And Language In Conversational Environments (2024)5.84
- Taltech-irit-lis Speaker And Language Diarization Systems For DISPLACE 2024 (2024)4.52
- The Second DIHARD Diarization Challenge: Dataset, Task, And Baselines (2019)15.00
- Exploring Speaker-related Information In Spoken Language Understanding For Better Speaker Diarization (2023)0.00
- Spot The Conversation: Speaker Diarisation In The Wild (2020)15.31
- TSUP Speaker Diarization System For Conversational Short-phrase Speaker Diarization Challenge (2022)5.24
- Integrating Audio, Visual, And Semantic Information For Enhanced Multimodal Speaker Diarization (2024)0.00