Summary On The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Grand Challenge
2022 Β· Fan Yu, Shiliang Zhang, Pengcheng Guo, et al.
Abstract
The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Grand Challenge (M2MeT) focuses on one of the most valuable and the most challenging scenarios of speech technologies. The M2MeT challenge has particularly set up two tracks, speaker diarization (track 1) and multi-speaker automatic speech recognition (ASR) (track 2). Along with the challenge, we released 120 hours of real-recorded Mandarin meeting speech data with manual annotation, including far-field data collected by 8-channel microphone array as well as near-field data collected by each participants' headset microphone. We briefly describe the released dataset, track setups, baselines and summarize the challenge results and major techniques used in the submissions.
Authors
(none)
Tags
Stats
Related papers
- The Second Multi-channel Multi-party Meeting Transcription Challenge (m2met) 2.0): A Benchmark For Speaker-attributed ASR (2023)6.77
- The Volcspeech System For The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (2022)5.84
- The CUHK-TENCENT Speaker Diarization System For The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (2022)7.81
- The Ustc-ximalaya System For The ICASSP 2022 Multi-channel Multi-party Meeting Transcription (m2met) Challenge (2022)6.34
- Royalflush Speaker Diarization System For ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (2022)0.00
- The Xmuspeech System For Multi-channel Multi-party Meeting Transcription Challenge (2022)0.00
- Pp-met: A Real-world Personalized Prompt Based Meeting Transcription System (2023)4.52
- Cross-channel Attention-based Target Speaker Voice Activity Detection: Experimental Results For M2met Challenge (2022)10.07