Kikori-KS: An Effective and Efficient Keyword Search System for Digital Libraries in XML

Kikori-KS: An Effective and Efficient Keyword Search System for Digital Libraries in XML
复制标题

DOI:
10.1007/11931584_42
复制
发表时间:
2006-11
期刊:
--
影响因子:
--
通讯作者:
Toshiyuki Shimizu;N. Terada;Masatoshi Yoshikawa
Toshiyuki Shimizu;N. Terada;Masatoshi Yoshikawa
中科院分区:
其他
文献类型:
--
作者:
Toshiyuki Shimizu;N. Terada;Masatoshi Yoshikawa

文献摘要

相似文献

识别有意义的文档片段是用XML编码文档的一个主要优点。在学术文章中,这样的文档片段包括节、子节和段落。XML信息检索系统需要从数字图书馆中的XML文档集中检索与查询相关的文档片段。我们提出了Kikori-KS,一个有效的和高效的XML信息检索系统的学术文章。Kikori-KS接受一组关键字作为查询。这种查询形式简单而有用,因为用户不需要理解XML查询语言或XML模式。为了满足在学术论文中搜索相关片段的实际需求,我们开发了一个用户友好的界面来显示搜索结果。Kikori-KS是在我们小组开发的关系XML数据库系统之上实现的。通过仔细设计数据库模式,Kikori-KS有效地处理了大量的文档片段。我们使用INEX测试集的实验表明,Kikori-KS实现了可接受的搜索时间和相对较高的精度。
Identifying meaningful document fragments is a major advantage achieved by encoding documents in XML. In scholarly articles, such document fragments include sections, subsections and paragraphs. XML information retrieval systems need to search document fragments relevant to queries from a set of XML documents in a digital library. We present Kikori-KS, an effective and efficient XML information retrieval system for scholartic articles. Kikori-KS accepts a set of keywords as a query. This form of query is simple yet useful because users are not required to understand XML query languages or XML schema. To meet practical demands for searching relevant fragments in scholartic articles, we have developed a user-friendly interface for displaying search results. Kikori-KS was implemented on top of a relational XML database system developed by our group. By carefully designing the database schema, Kikori-KS handles a huge number of document fragments efficiently. Our experiments using INEX test collection show that Kikori-KS achieved an acceptable search time and with relatively high precision.