Deriving query intents from web search engine queries

Deriving query intents from web search engine queries
复制标题

从网络搜索引擎查询中派生查询意图

DOI:
10.1002/asi.22706
复制
发表时间:
2012
期刊:
J. Assoc. Inf. Sci. Technol.
影响因子:
--
通讯作者:
S. Mach
S. Mach
中科院分区:
--
文献类型:
--
作者:
D. Lewandowski;Jessica Drechsler;S. Mach

文献摘要

被引文献

相似文献

本文的目的是测试从查询派生的查询意图的可靠性,无论是由输入查询的用户还是由另一个陪审员。我们报告了三项研究的结果。首先,我们使用众包方法进行了大规模的分类研究(约50,000个查询)。接下来,我们使用了来自搜索引擎日志的点击数据,并验证了众包研究中陪审员的判断。最后,我们在一个商业搜索引擎的门户网站上进行了在线调查。因为我们在所有三项研究中使用了相同的查询,所以我们也能够比较不同方法的结果和有效性。我们发现,无论是众包的方法,使用陪审员谁分类查询来自其他用户,也不是问卷调查的方法,使用搜索者被问及他们自己的查询,他们刚刚进入一个网络搜索引擎,导致令人满意的结果。这让我们得出结论,尽管两组陪审员都得到了详细的指示,但他们对分类任务知之甚少。虽然我们使用的是手动分类,但我们的研究对自动分类也有重要意义。我们必须质疑使用自动分类并将其性能与人类陪审员的基线进行比较的方法是否成功。© 2012 Wiley Periodicals,Inc.
The purpose of this article is to test the reliability of query intents derived from queries, either by the user who entered the query or by another juror. We report the findings of three studies. First, we conducted a large-scale classification study (~50,000 queries) using a crowdsourcing approach. Next, we used clickthrough data from a search engine log and validated the judgments given by the jurors from the crowdsourcing study. Finally, we conducted an online survey on a commercial search engine's portal. Because we used the same queries for all three studies, we also were able to compare the results and the effectiveness of the different approaches. We found that neither the crowdsourcing approach, using jurors who classified queries originating from other users, nor the questionnaire approach, using searchers who were asked about their own query that they just entered into a Web search engine, led to satisfying results. This leads us to conclude that there was little understanding of the classification tasks, even though both groups of jurors were given detailed instructions. Although we used manual classification, our research also has important implications for automatic classification. We must question the success of approaches using automatic classification and comparing its performance to a baseline from human jurors. © 2012 Wiley Periodicals, Inc.