Skip to main navigation Skip to search Skip to main content

Bridge the semantic gap between pop music acoustic feature and emotion: Build an interpretable model

  • Jiang Long Zhang*
  • , Xiang Lin Huang
  • , Lifang Yang
  • , Liqiang Nie
  • *Corresponding author for this work
  • Communication University of China
  • National University of Singapore

Research output: Contribution to journalArticlepeer-review

Abstract

Music emotion recognition (MER) is an important topic in music understanding, recommendation, retrieval and human computer interaction. Great success has been achieved by machine learning methods in estimating human emotional response to music. However, few of them pay much attention in semantic interpret for emotion response. In our work, we first train an interpretable model between acoustic audio and emotion. Filter, wrapper and shrinkage methods are applied to select important features. We then apply statistical models to build and explain the emotion model. Extensive experimental results reveal that the shrinkage methods outperform the wrapper methods and the filter methods in arousal emotion. In addition, we observed that only a small set of the extracted features have the key effects to arousal. While, most of our extracted features have small contribution to valence music perception. Ultimately, we obtain a higher average accuracy rate in arousal, compared to that in valence.

Original languageEnglish
Pages (from-to)333-341
Number of pages9
JournalNeurocomputing
Volume208
DOIs
StatePublished - 5 Oct 2016
Externally publishedYes

Keywords

  • Feature selection
  • Interpretable model
  • Music emotion
  • Semantic gap
  • Shrinkage methods

Fingerprint

Dive into the research topics of 'Bridge the semantic gap between pop music acoustic feature and emotion: Build an interpretable model'. Together they form a unique fingerprint.

Cite this