ISCAS at Subtopic Mining Task in NTCIR9

ISCAS at Subtopic Mining Task in NTCIR9
复制标题

DOI:
--
复制
发表时间:
2011
期刊:
--
影响因子:
--
通讯作者:
Xue Jiang;Xianpei Han;Le Sun-
Xue Jiang;Xianpei Han;Le Sun-
中科院分区:
其他
文献类型:
--
作者:
Xue Jiang;Xianpei Han;Le Sun-

文献摘要

被引文献

相似文献

本文描述了我们在NTCIR-9中的子主题挖掘子任务中的工作。为了找到特定查询的可能子主题,我们选择查询日志记录的相关查询,或Google和百度提供的搜索结果标题,或百度百科中相应条目的目录,这些查询在词汇上与原始查询相似,然后使用k-means算法对这些具有不同k(k=5,10)的候选查询进行聚类,并考虑相似性和聚类对这些查询进行排序。关键词子主题挖掘,查询日志
paper, we describe our work at subtopic mining subtask in NTCIR-9 in simplified Chinese. To find possible subtopics of a specific query, we select related queries recorded by query log, or titles of searching results provided by Google and Baidu, or the catalog of corresponding entry in Baidu encyclopedia, which are lexically similar as the original query, then we apply k-means algorithm to cluster these candidate queries with different k (k=5, 10), and rank these queries with consideration of similarities and clusters. Keywordssubtopic mining, query log