Skip to main navigation Skip to search Skip to main content

CCL24-Eval任务9总结报告:中文图文多模态理解评测

Translated title of the contribution: Overview of CCL24-Eval Task 9: Chinese Vision-Language Understanding
  • Zhejiang Lab
  • Harbin Institute of Technology

Research output: Contribution to conferencePaperpeer-review

Abstract

The Chinese Vision-Language Understanding Task aims to evaluate the vision-language modeling and understanding capabilities of Chinese vision-language pre-training models from multiple perspectives. This task includes five sub-tasks: image retrieval, text retrieval, visual question answering, visual grounding, and visual dialogue. The final score is calculated based on the combined results of these five tasks. This paper first introduces the background and motivation of the task, then presents and demonstrates relevant information about this task from aspects including task description, evaluation metrics, submitted results, and participant methods. A total of 11 teams registered to participate in this task. And 3 teams eventually submitted their results.

Translated title of the contributionOverview of CCL24-Eval Task 9: Chinese Vision-Language Understanding
Original languageChinese (Traditional)
Pages364-371
Number of pages8
StatePublished - 2024
Event23rd Chinese National Conference on Computational Linguistics, CCL 2024 - Taiyuan, China
Duration: 24 Jul 202428 Jul 2024

Conference

Conference23rd Chinese National Conference on Computational Linguistics, CCL 2024
Country/TerritoryChina
CityTaiyuan
Period24/07/2428/07/24

Fingerprint

Dive into the research topics of 'Overview of CCL24-Eval Task 9: Chinese Vision-Language Understanding'. Together they form a unique fingerprint.

Cite this