Skip to main navigation Skip to search Skip to main content

VCD: VIEW-CONSTRAINT DISENTANGLEMENT FOR ACTION RECOGNITION

  • Xian Zhong
  • , Zhuo Zhou
  • , Wenxuan Liu
  • , Kui Jiang
  • , Xuemei Jia
  • , Wenxin Huang
  • , Zheng Wang
  • Wuhan University of Technology
  • Peking University
  • Wuhan University
  • Hubei University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Action recognition is a hot topic in computer vision due to its wide range of applications in urban surveillance. Although some methods are more advanced from an invariant view perspective, those approaches do not perform well for the viewpoint change. To address this issue, one possible solution is tantamount to track the view-invariant representation as it evolves with the performed action. However, the views' and actions' performance always complement each other, once simply looking for the view-invariant representation may cause some behavior information to be lost. In this paper, we propose the View-Constraint Disentanglement (VCD) framework for cross-view action recognition. Specifically, Constraint Disentanglement Module (CDM) is utilized to learn an action-invariant representation by discretizing view-specific representation and its normal distribution, which resolves the entangled relationship between view and action. Moreover, a novel Adaptive Distribution Module (ADM) is intended to befit enhance the high-correlation viewpoint variation information and refine the suitable weight. Extensive experiments are conducted on public benchmarks, indicating that our approach achieves better performance than other state-of-the-art approaches.

Original languageEnglish
Title of host publication2022 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2022 - Proceedings
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages2170-2174
Number of pages5
ISBN (Electronic)9781665405409
DOIs
StatePublished - 2022
Externally publishedYes
Event2022 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2022 - Hybrid, Singapore
Duration: 22 May 202227 May 2022

Publication series

NameICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings
Volume2022-May
ISSN (Print)1520-6149

Conference

Conference2022 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2022
Country/TerritorySingapore
CityHybrid
Period22/05/2227/05/22

Keywords

  • Action recognition
  • Adaptive distribution
  • Cross-view
  • Disentangled representation learning

Fingerprint

Dive into the research topics of 'VCD: VIEW-CONSTRAINT DISENTANGLEMENT FOR ACTION RECOGNITION'. Together they form a unique fingerprint.

Cite this