Overview
Who this competition fits
This suits researchers or engineers who can build keyphrase extraction systems for Mandarin meeting transcripts, especially those familiar with long-document NLP and fine-tuning pretrained models. It welcomes academia and industry participants, but Alibaba employees and Zhejiang University faculty/students cannot participate.
Read the original official blurb
本项目为ICASSP2023 信号处理大挑战的通用会议理解及生成挑战赛(MUG challenge)。赛事构建并发布了目前为止规模最大的中文会议数据集,并基于会议人工转写结果进行了多项口语语言处理(SLP)任务的标注;目标是推动SLP在会议文本处理场景的研究并应对其中的多项关键挑战,包括 人人交互场景下多样化的口语现象、会议场景下的长篇章文档建模 等。 Official category: 算法竞赛. Participants: 63. Organizers: 阿里巴巴达摩院, 魔搭modelscope社区, 浙江大学.
Preparation
From registration to a first submission
- 01
Mandarin text processing
- 02
keyphrase extraction
- 03
long-document modeling
- 04
Transformer fine-tuning
- 05
spoken-language transcript handling
- 06
ModelScope pre-trained models
Before you commit: The main challenge is extracting salient phrases from long, conversational meeting transcripts rather than clean written text. Disfluencies, redundancies, grammar errors, coreference, and fragmented utterances compound the computational difficulty of processing documents with several thousand words or more.
Source
How this page was assembled
Competition information is structured from the official page. Scores are platform estimates for decision support; official rules take precedence.
- Official competition page
- ModelScope
- Last checked
- Dec 2, 2022