"Power tags" in information retrieval

"Power tags" in information retrieval
复制标题

DOI:
10.1108/07378831011026706
复制
发表时间:
2010-01-01
期刊:
影响因子:
3.4
通讯作者:
Stock, Wolfgang G.
Stock, Wolfgang G.
中科院分区:
管理学4区
文献类型:
--
作者:
Peters, Isabella;Stock, Wolfgang G.

文献摘要

被引文献

相似文献

用途-许多Web 2.0服务(包括Library 2.0目录)使用大众分类法。本文的目的是切断特定于文档的标签分布的长尾中的所有标签。其余的标签在开始时的标签分布被认为是电力标签,并形成一个新的,额外的搜索选项在信息检索systems.Design/方法/方式-在理论上的方法,本文讨论了文件特定的标签分布(幂律和逆逻辑形状),这种分布的发展(尤尔-西蒙过程和洗牌理论),并介绍了搜索标签(除了众所周知的索引标签)作为生成标签分布的可能性。搜索标签与广义和狭义的大众分类法以及所有知识组织系统(例如分类系统和叙词表)兼容,而索引标签仅适用于广义的大众分类法。根据这些发现,本文提出了一个草图的算法挖掘和处理电力标签在信息检索系统。研究限制/影响-这种概念性的方法是需要在一个具体的检索系统的实证评估。实际影响-电力标签是一个新的搜索选择检索系统,以限制点击量。独创性/价值-本文介绍了电力标签作为一种手段,提高精度的搜索结果的信息检索系统,应用大众分类法,例如在图书馆2.0环境中的目录。
Purpose - Many Web 2.0 services (including Library 2.0 catalogs) make use of folksonomies. The purpose of this paper is to cut off all tags in the long tail of a document-specific tag distribution. The remaining tags at the beginning of a tag distribution are considered power tags and form a new, additional search option in information retrieval systems.Design/methodology/approach - In a theoretical approach the paper discusses document-specific tag distributions (power law and inverse-logistic shape), the development of such distributions (Yule-Simon process and shuffling theory) and introduces search tags (besides the well-known index tags) as a possibility for generating tag distributions.Findings - Search tags are compatible with broad and narrow folksonomies and with all knowledge organization systems (e.g. classification systems and thesauri), while index tags are only applicable in broad folksonomies. Based on these findings, the paper presents a sketch of an algorithm for mining and processing power tags in information retrieval systems.Research limitations/implications - This conceptual approach is in need of empirical evaluation in a concrete retrieval system.Practical implications - Power tags are a new search option for retrieval systems to limit the amount of hits.Originality/value - The paper introduces power tags as a means for enhancing the precision of search results in information retrieval systems that apply folksonomies, e.g. catalogs in Library 2.0 environments.