Article ID | Journal | Published Year | Pages | File Type |
---|---|---|---|---|
405343 | Knowledge-Based Systems | 2010 | 6 Pages |
Abstract
Identification of meaningful clusters from categorical data is one key problem in data mining. Recently, Average Normalized Mutual Information (ANMI) has been used to define categorical data clustering as an optimization problem. To find globally optimal or near-optimal partition determined by ANMI, a genetic clustering algorithm (G-ANMI) is proposed in this paper. Experimental results show that G-ANMI is superior or comparable to existing algorithms for clustering categorical data in terms of clustering accuracy.
Related Topics
Physical Sciences and Engineering
Computer Science
Artificial Intelligence
Authors
Shengchun Deng, Zengyou He, Xiaofei Xu,