A Clustering Method Based on Rough Sets and Its Application to Knowledge Discovery in the Medical Database
A Clustering Method Based on Rough Sets and Its Application to Knowledge Discovery in the Medical Database
复制标题
基于粗糙集的聚类方法及其在医学数据库知识发现中的应用
DOI:
10.3233/978-1-60750-928-8-206
复制
发表时间:
2001
影响因子:
--
通讯作者:
Y. Hata
中科院分区:
文献类型:
--
作者:
S. Hirano;S. Tsumoto;Tomohiro Okuzaki;Y. Hata
This paper proposes a clustering method for nominal and numerical data based on Rough Sets and its application to knowledge discovery in the medical database. Classification is performed according to the indiscernibility relations defined on the basis of relative similarity between objects. The similarity is defined as a combination of two types of similarity measures: the Hamming distance for nominal attributes and the Mahalanobis distance for numerical attributes. Excessive generation of small category is suppressed by modifying similar equivalence relations into the same equivalence relation. An analysis of the meningoencephalitis diagnosis database was performed to validate this method. The result showed that this method could deal well with both types of attributes and discover the primary factors for diagnosis.