Skip to main navigation Skip to search Skip to main content

Sentiment classification of online Cantonese reviews by supervised machine learning approaches

  • School of Management, Harbin Institute of Technology
  • Hong Kong Polytechnic University

Research output: Contribution to journalArticlepeer-review

Abstract

Cantonese is an important Chinese dialect spoken in some regions of Southern China. Local online users often represent their opinions and experiences with written Cantonese on the web. With two supervised machine learning approaches, this paper conducts a series of experiments to explore appropriate methods for automatic sentiment classification in the very noisy domain of online Cantonese-written reviews. Findings indicate that the support vector machine classifier based on a Mandarin Chinese word segmentation tool performs surprisingly well. The accuracy, precision and recall respectively for positive and negative reviews all reach above 85% when the training corpus contains 5,000 or more reviews.

Original languageEnglish
Pages (from-to)382-397
Number of pages16
JournalInternational Journal of Web Engineering and Technology
Volume5
Issue number4
DOIs
StatePublished - Mar 2009
Externally publishedYes

Keywords

  • Cantonese
  • Online reviews
  • Sentiment classification
  • Text mining

Fingerprint

Dive into the research topics of 'Sentiment classification of online Cantonese reviews by supervised machine learning approaches'. Together they form a unique fingerprint.

Cite this