The Dku-dukeece System For The Self-supervision Speaker Verification Task Of The 2021 Voxceleb Speaker Recognition Challenge
2021 Β· Danwei Cai, Ming Li
Abstract
This report describes the submission of the DKU-DukeECE team to the self-supervision speaker verification task of the 2021 VoxCeleb Speaker Recognition Challenge (VoxSRC). Our method employs an iterative labeling framework to learn self-supervised speaker representation based on a deep neural network (DNN). The framework starts with training a self-supervision speaker embedding network by maximizing agreement between different segments within an utterance via a contrastive loss. Taking advantage of DNN's ability to learn from data with label noise, we propose to cluster the speaker embedding obtained from the previous speaker network and use the subsequent class assignments as pseudo labels to train a new DNN. Moreover, we iteratively train the speaker network with pseudo labels generated from the previous step to bootstrap the discriminative power of a DNN. Also, visual modal data is incorporated in this self-labeling framework. The visual pseudo label and the audio pseudo label are f
Authors
(none)
Tags
Stats
Related papers
- An Iterative Framework For Self-supervised Deep Speaker Representation Learning (2020)10.61
- The DKU-MSXF Speaker Verification System For The Voxceleb Speaker Recognition Challenge 2023 (2023)0.00
- The Phonexia Voxceleb Speaker Recognition Challenge 2021 System Description (2021)0.00
- Beijing ZKJ-NPU Speaker Verification System For Voxceleb Speaker Recognition Challenge 2021 (2021)0.00
- Curriculum Learning For Self-supervised Speaker Verification (2022)8.09
- The DKU-MSXF Diarization System For The Voxceleb Speaker Recognition Challenge 2023 (2023)5.24
- The Kriston AI System For The Voxceleb Speaker Recognition Challenge 2022 (2022)0.00
- Self-distillation Prototypes Network: Learning Robust Speaker Representations Without Supervision (2023)4.52