Article ID | Journal | Published Year | Pages | File Type |
---|---|---|---|---|
385297 | Expert Systems with Applications | 2008 | 8 Pages |
Abstract
Up to now, there are very few researches conducted on sentiment classification for Chinese documents. In order to remedy this deficiency, this paper presents an empirical study of sentiment categorization on Chinese documents. Four feature selection methods (MI, IG, CHI and DF) and five learning methods (centroid classifier, K-nearest neighbor, winnow classifier, Naïve Bayes and SVM) are investigated on a Chinese sentiment corpus with a size of 1021 documents. The experimental results indicate that IG performs the best for sentimental terms selection and SVM exhibits the best performance for sentiment classification. Furthermore, we found that sentiment classifiers are severely dependent on domains or topics.
Related Topics
Physical Sciences and Engineering
Computer Science
Artificial Intelligence
Authors
Songbo Tan, Jin Zhang,