Skip to main navigation Skip to search Skip to main content

Adaptive Across-Subcenter Representation Learning for Imbalanced Anomalous Sound Detection

  • Faculty of Computing, Harbin Institute of Technology

Research output: Contribution to journalConference articlepeer-review

Abstract

Anomalous Sound Detection requires constructing a distribution using only normal sounds. However, collecting sufficient normal samples across diverse conditions is challenging, leading to sample imbalance within subclasses. Existing subcenter angular margin loss methods use multiple subcenters to capture intra-class diversity but still suffer from under-representation or overfitting. To address this issue, we propose Adaptive Across-Subcenter Representation Learning (AASRL). Unlike existing methods that use either a single or all subcenters, AASRL adaptively selects subcenters based on the representation quality of samples and optimizes their representation across the most relevant subcenters. This ensures efficient representation of each sample and prevents the majority subclass from dominating the representation space. Experiments on the DCASE2023 Challenge Task2 dataset and a constructed imbalanced dataset demonstrate the effectiveness of AASRL.

Original languageEnglish
Pages (from-to)3379-3383
Number of pages5
JournalProceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
DOIs
StatePublished - 2025
Externally publishedYes
Event26th Interspeech Conference 2025 - Rotterdam, Netherlands
Duration: 17 Aug 202521 Aug 2025

Keywords

  • angular margin loss
  • anomalous sound detection
  • data imbalance
  • subcenter

Fingerprint

Dive into the research topics of 'Adaptive Across-Subcenter Representation Learning for Imbalanced Anomalous Sound Detection'. Together they form a unique fingerprint.

Cite this