论文标题
ICASSP 2022多渠道多党会议转录大挑战的摘要
Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
论文作者
论文摘要
ICASSP 2022多渠道多党会议大会大挑战(M2MET)着重于言语技术最有价值,最具挑战性的场景之一。 M2MET挑战特别设置了两个轨道,扬声器诊断(曲目1)和多扬声器自动语音识别(ASR)(轨道2)。除挑战外,我们发布了120个小时的真实录制的普通话会议语音数据,其中包括手动注释,包括由8通道麦克风阵列收集的远场数据以及每个参与者的耳机麦克风收集的近场数据。我们简要描述已发布的数据集,轨道设置,基线,并总结提交中使用的挑战结果和主要技术。
The ICASSP 2022 Multi-channel Multi-party Meeting Transcription Grand Challenge (M2MeT) focuses on one of the most valuable and the most challenging scenarios of speech technologies. The M2MeT challenge has particularly set up two tracks, speaker diarization (track 1) and multi-speaker automatic speech recognition (ASR) (track 2). Along with the challenge, we released 120 hours of real-recorded Mandarin meeting speech data with manual annotation, including far-field data collected by 8-channel microphone array as well as near-field data collected by each participants' headset microphone. We briefly describe the released dataset, track setups, baselines and summarize the challenge results and major techniques used in the submissions.