Overview
Who this competition fits
Builders comfortable with Chinese algorithm-competition pages hosted by iFlytek.
Read the original official blurb
在现实世界中,信息往往以多模态的形式呈现,尤其是图像与文本的结合。例如,包含图示的教科书、带图解的说明书、信息图表、科学论文、新闻报道,甚至是社交媒体帖子,都要求我们不仅能理解单独的图片或文字,更要将两者结合起来进行深层次的逻辑推理。当前的AI模型在处理单一模态(如纯文本阅读理解或纯图像识别)方面已取得显著进展,但在需要跨模态、多步骤、甚至涉及常识知识… Official track: 多模态技术. Sponsor: 科大讯飞xDatawhale. Teams entered: 48.
Preparation
From registration to a first submission
- 01
Can read Chinese competition pages
- 02
Comfortable submitting a baseline
- 03
Basic familiarity with LLM tooling or agent/skill workflows
Before you commit: Task details, eligibility, and scoring rules live mainly on the official iFlytek page and change each season.
Source
How this page was assembled
Competition information is structured from the official page. Scores are platform estimates for decision support; official rules take precedence.
- Official competition page
- iFlytek Challenge
- Last checked
- Jul 25, 2025