Automatic categorization of web sites based on source types

Automatic categorization of web sites based on source types
复制标题

根据来源类型自动对网站进行分类

DOI:
10.1145/1012807.1012821
复制
发表时间:
2004
期刊:
--
影响因子:
--
通讯作者:
R. Krishnapuram
R. Krishnapuram
中科院分区:
--
文献类型:
--
作者:
Shourya Roy;Sachindra Joshi;R. Krishnapuram

文献摘要

被引文献

相似文献

Web的一个重要问题是验证与Web站点相关的信息的准确性、时效性和真实性。解决这个问题的一种方法是确定Web站点的“来源”或“赞助者”。但是,源标识并不简单,因为Web站点的源不能总是由站点的URL或内容来确定。在本文中,我们提出了一种源识别方法,该方法使用由于Web站点之间和站点内部的超链接而产生的各种类型的入站、出站和内部交互。
An important issue with the Web is verification of the accuracy, currency and authenticity of the information associated with Web sites. One way to address this problem is to identify the "source" or "sponsor" of the Web site. However, source identification is non-trivial because the source of a Web site cannot always be determined by the URL or content of the site. In this paper, we propose a method for source identification that uses various types of inbound, outbound and internal interactions that arise due to hyperlinks between and within Web sites.