Research on K-means Text Clustering Algorithm Based on Semantic
Research on K-means Text Clustering Algorithm Based on Semantic
复制标题
基于语义的K-means文本聚类算法研究
DOI:
10.1109/ccie.2010.39
复制
发表时间:
2010
期刊:
影响因子:
--
通讯作者:
Shuicai Shi
中科院分区:
文献类型:
--
作者:
Yufang Liu;Shibin Xiao;Xueqiang Lv;Shuicai Shi
Through research on K-means algorithm of text clustering and semantic-based vector space model, a semantic-based K-means text clustering model is proposed to solve the problem on high-dimensional and sparse characteristics of text data set. The model reduces the semantic loss of the text data and improves the quality of text clustering. Experiments prove that semantic-based text clustering increases by more 6 percent than non-semantic-based one in the final evaluation of the F1 index value.