Skip to main navigation Skip to search Skip to main content

Unsupervised Time-Aware Sampling Network With Deep Reinforcement Learning for EEG-Based Emotion Recognition

  • Yongtao Zhang
  • , Yue Pan
  • , Yulin Zhang
  • , Min Zhang
  • , Linling Li
  • , Li Zhang
  • , Gan Huang
  • , Lei Su
  • , Honghai Liu
  • , Zhen Liang*
  • , Zhiguo Zhang*
  • *Corresponding author for this work
  • Shenzhen University
  • Harbin Institute of Technology Shenzhen
  • Marshall Laboratory of Biomedical Engineering
  • Peng Cheng Laboratory

Research output: Contribution to journalArticlepeer-review

Abstract

Recognizing human emotions from complex, multivariate, and non-stationary electroencephalography (EEG) time series is essential in affective brain-computer interface. However, because continuous labeling of ever-changing emotional states is not feasible in practice, existing methods can only assign a fixed label to all EEG timepoints in a continuous emotion-evoking trial, which overlooks the highly dynamic emotional states and highly non-stationary EEG signals. To solve the problems of high reliance on fixed labels and ignorance of time-changing information, in this paper we propose a time-aware sampling network (TAS-Net) using deep reinforcement learning (DRL) for unsupervised emotion recognition, which is able to detect key emotion fragments and disregard irrelevant and misleading parts. Specifically, we formulate the process of mining key emotion fragments from EEG time series as a Markov decision process and train a time-aware agent through DRL without label information. First, the time-aware agent takes deep features from a feature extractor as input and generates sample-wise importance scores reflecting the emotion-related information each sample contains. Then, based on the obtained sample-wise importance scores, our method preserves top-X continuous EEG fragments with relevant emotion and discards the rest. Finally, we treat these continuous fragments as key emotion fragments and feed them into a hypergraph decoding model for unsupervised clustering. Extensive experiments are conducted on three public datasets (SEED, DEAP, and MAHNOB-HCI) for emotion recognition using leave-one-subject-out cross-validation, and the results demonstrate the superiority of the proposed method against previous unsupervised emotion recognition methods. The proposed TAS-Net has great potential in achieving a more practical and accurate affective brain-computer interface in a dynamic and label-free circumstance.

Original languageEnglish
Pages (from-to)1090-1103
Number of pages14
JournalIEEE Transactions on Affective Computing
Volume15
Issue number3
DOIs
StatePublished - 2024
Externally publishedYes

Keywords

  • Electroencephalography
  • affective brain-computer interface
  • deep reinforcement learning
  • emotion recognition
  • unsupervised learning

Fingerprint

Dive into the research topics of 'Unsupervised Time-Aware Sampling Network With Deep Reinforcement Learning for EEG-Based Emotion Recognition'. Together they form a unique fingerprint.

Cite this