CLUSTERING THE SCIENCE CITATION INDEX USING CO-CITATIONS .1. A COMPARISON OF METHODS

CLUSTERING THE SCIENCE CITATION INDEX USING CO-CITATIONS .1. A COMPARISON OF METHODS
复制标题

DOI:
10.1007/bf02017157
复制
发表时间:
1985-01-01
期刊:
影响因子:
3.9
通讯作者:
SWEENEY, E
SWEENEY, E
中科院分区:
管理学3区
文献类型:
--
作者:
SMALL, H;SWEENEY, E

文献摘要

被引文献

相似文献

回顾了使用共引对科学引文索引(SCI)数据库进行聚类的早期实验。对该方法提出了两点改进建议:分数引文计数和具有最大聚类大小限制的可变水平聚类。描述了使用1979年SCI的实验结果,并将新方法与以前使用的方法进行了比较。分数引文计数有助于减少在使用整数引文计数阈值时固有的对高引用领域(如生物医学和生物化学)的偏见,并增加集群覆盖的主题范围。另一方面,可变级别的聚类提高了查全率,这是通过聚类中高被引项的百分比来衡量的。这两种新方法的结合使用将提高我们生成德里克·普莱斯设想的全面科学地图的能力。
Earlier experiments in the use of co-citations to cluster the Science Citation Index (SCI) database are reviewed. Two proposed improvements in the methodology are introduced: fractional citation counting and variable level clustering with a maximum cluster size limit. Results of an experiment using the 1979 SCI are described comparing the new methods with those previously employed. Fractional citation counting helps reduce the bias toward high referencing fields such as biomedicine and biochemistry inherent in the use of an integer citation count threshold, and increases the range of subject matters covered by clusters. Variable level clustering, on the other hand, increases recall as measured by the percentage of highly cited items included in clusters. The 2 new methods used in combination will improve our ability to generate comprehensive maps of science as envisioned by Derek Price.