Abstract
The Chinese Vision-Language Understanding Task aims to evaluate the vision-language modeling and understanding capabilities of Chinese vision-language pre-training models from multiple perspectives. This task includes five sub-tasks: image retrieval, text retrieval, visual question answering, visual grounding, and visual dialogue. The final score is calculated based on the combined results of these five tasks. This paper first introduces the background and motivation of the task, then presents and demonstrates relevant information about this task from aspects including task description, evaluation metrics, submitted results, and participant methods. A total of 11 teams registered to participate in this task. And 3 teams eventually submitted their results.
| Translated title of the contribution | Overview of CCL24-Eval Task 9: Chinese Vision-Language Understanding |
|---|---|
| Original language | Chinese (Traditional) |
| Pages | 364-371 |
| Number of pages | 8 |
| State | Published - 2024 |
| Event | 23rd Chinese National Conference on Computational Linguistics, CCL 2024 - Taiyuan, China Duration: 24 Jul 2024 → 28 Jul 2024 |
Conference
| Conference | 23rd Chinese National Conference on Computational Linguistics, CCL 2024 |
|---|---|
| Country/Territory | China |
| City | Taiyuan |
| Period | 24/07/24 → 28/07/24 |
Fingerprint
Dive into the research topics of 'Overview of CCL24-Eval Task 9: Chinese Vision-Language Understanding'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver