Skip to main navigation Skip to search Skip to main content

CATSLU: The 1st Chinese audio-textual spoken language understanding challenge

  • Su Zhu
  • , Zijian Zhao
  • , Tiejun Zhao
  • , Chengqing Zong
  • , Kai Yu*
  • *Corresponding author for this work
  • Shanghai Jiao Tong University
  • Chinese Academy of Sciences

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Spoken language understanding (SLU) is a key component of conversational dialogue systems, which converts user utterances into semantic representations. The previous works almost focus on parsing semantic from textual inputs (top hypothesis of speech recognition and even manual transcripts) while losing information hidden in the audio. We herein describe the 1st Chinese Audio-Textual Spoken Language Understanding Challenge (CATSLU) which introduces a new dataset with audio-textual information, multiple domains and domain knowledge. We introduce two scenarios of audio-textual SLU in which participants are encouraged to utilize data of other domains or not. In this paper, we will describe the challenge and results.

Original languageEnglish
Title of host publicationICMI 2019 - Proceedings of the 2019 International Conference on Multimodal Interaction
EditorsWen Gao, Helen Mei Ling Meng, Matthew Turk, Susan R. Fussell, Bjorn Schuller, Bjorn Schuller, Yale Song, Kai Yu
PublisherAssociation for Computing Machinery, Inc
Pages521-525
Number of pages5
ISBN (Electronic)9781450368605
DOIs
StatePublished - 14 Oct 2019
Event21st ACM International Conference on Multimodal Interaction, ICMI 2019 - Suzhou, China
Duration: 14 Oct 201918 Oct 2019

Publication series

NameICMI 2019 - Proceedings of the 2019 International Conference on Multimodal Interaction

Conference

Conference21st ACM International Conference on Multimodal Interaction, ICMI 2019
Country/TerritoryChina
CitySuzhou
Period14/10/1918/10/19

Keywords

  • Datasets
  • Spoken language understanding

Fingerprint

Dive into the research topics of 'CATSLU: The 1st Chinese audio-textual spoken language understanding challenge'. Together they form a unique fingerprint.

Cite this