Overview
Who this competition fits
This suits NLP practitioners who can build sentence-level classifiers for Mandarin meeting transcripts and fine-tune publicly available pre-trained language models. It is also suitable for academic or industry researchers, except challenge organizers, Alibaba employees, and Zhejiang University faculty or students, who are explicitly ineligible.
Read the original official blurb
本项目为ICASSP2023 信号处理大挑战的通用会议理解及生成挑战赛(MUG challenge)。赛事构建并发布了目前为止规模最大的中文会议数据集,并基于会议人工转写结果进行了多项口语语言处理(SLP)任务的标注;目标是推动SLP在会议文本处理场景的研究并应对其中的多项关键挑战,包括 人人交互场景下多样化的口语现象、会议场景下的长篇章文档建模 等。 Official category: 算法竞赛. Participants: 55. Organizers: 阿里巴巴达摩院, 魔搭modelscope社区, 浙江大学.
Preparation
From registration to a first submission
- 01
Mandarin NLP
- 02
meeting transcript processing
- 03
sentence classification
- 04
pre-trained language model fine-tuning
- 05
long-document modeling
- 06
precision–recall evaluation
Before you commit: The core difficulty is recognizing actionable-task sentences in long, spoken meeting transcripts containing disfluencies, redundancies, grammar errors, coreference, colloquial language, and fragments. Participants must also work within a constrained-track setup: evaluation data cannot be used for training or fine-tuning, and only specified public models and corpora are allowed.
Source
How this page was assembled
Competition information is structured from the official page. Scores are platform estimates for decision support; official rules take precedence.
- Official competition page
- ModelScope
- Last checked
- Dec 2, 2022