课题基金 / 基金详情

Enriching, repairing and merging taxonomies by inducing qualitative spatial representations from the web

Enriching, repairing and merging taxonomies by inducing qualitative spatial representations from the web
通过从网络中引入定性空间表示来丰富、修复和合并分类法
批准号:
EP/K021788/1
负责人:
Steven Schockaert
金额:
$12.61万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2013
资助国家:
英国
项目状态:
已结题
起止时间:
2013 至 --

项目摘要

项目成果

Steven Schockaert的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Taxonomies encode how different terms or concepts from a given domain are related to each other. They are used to standardise vocabularies (e.g. biologists use taxonomies to organise species into broader categories such as family and order), and to categorise content such that it can be more easily searched (e.g. librarians assigning categories from a taxonomy to books). While taxonomies are traditionally the result of a careful and time-consuming manual process, recent developments in the world wide web have led to a proliferation of taxonomies of a more informal nature. Online retailers such as Amazon, for instance, organise their products using an ad hoc taxonomy, which reflects how customers use their website, rather than any commitment on the semantics of the underlying product categories. Similarly, applications such as Foursquare allow users to contribute to a taxonomy of place types.While these informal taxonomies are useful to organise online content (e.g. products on Amazon, or venues on Foursquare), they are often of poor quality, and difficult to reuse among different applications. Moreover, like traditional taxonomies, they focus on a very limited set of semantic relations; usually only the relation "is a sub-category of" is considered. In contrast, in practice the semantic relationship between two categories may not be so clear-cut, among others because of the existence of borderline cases (e.g. should a pub which serves food be categorised as a restaurant?). Nonetheless, the widespread availability of taxonomies is of potentially great interest, provided that they can be improved using automated methods. The goal of this project is to study how such an improvement can be realised, by statistically analysing meta-data that is available on the web, and in particular from so-called Web 2.0 websites such as Flickr, where users describe photos using short textual annotations called tags.The proposed approach is built on the idea of discovering semantic relationships between categories by statistically analysing such meta-data. On the one hand, these relations will encode information about typicality and similarity. To see why such relations are useful, consider an application which allows a user to search for restaurants in Cardiff. The search engine may rank venues of type "restaurant" by taking into account features such as distance to the city centre and average ratings (if available). However, as another criterion, one would also want to see "normal" restaurants before venues such as breakfast places, coffee houses, or pubs, which may be considered as restaurants, broadly speaking, but are not what users would typically be interested in when querying about restaurants. Similarly, when the user's query asks about "Sichuan restaurants in Cardiff", and no such restaurants are known, instances of the most similar categories may be shown instead (e.g. Cantonese restaurants). On the other hand, the relations that are discovered will also encode information that can help us to pinpoint likely errors in existing taxonomies and that can help us to merge different taxonomies to get a single coherent view of a given domain. In particular, these relations will allow us to detect irregularities in existing taxonomies. For example, given the assumption that similar categories usually have similar properties, and the knowledge that Cantonese and Sichuan restaurants are very similar, a taxonomy in which Cantonese and Sichuan restaurants are both sub-categories of Chinese restaurants will be considered more regular than a taxonomy in which they have different super-categories.Our approach is unique in its data-driven approach to enrich taxonomies with semantic relations for common-sense reasoning, as well as in the proposed methods for repairing and merging existing taxonomies. Regarding applications, the results of this project will form a crucial stepping-stone towards more intelligent search engines.
期刊论文(6)
专著(0)
科研奖励(0)
会议论文
DOI: 10.1016/j.artint.2015.07.002
发表时间: 2015-11
期刊: Artif. Intell.
影响因子: --
作者: [J. Derrac;Steven Schockaert]
通讯作者: J. Derrac;Steven Schockaert
Realizing RCC8 networks using convex regions
使用凸区域实现 RCC8 网络
DOI: 10.48550/arxiv.1410.2442
发表时间: 2014
期刊:
影响因子: --
作者: [Schockaert S]
通讯作者: Schockaert S
Commonsense reasoning based on betweenness and direction in distributional models
分布模型中基于介数和方向的常识推理
DOI: --
发表时间: 2015
期刊:
影响因子: --
作者: [Steven Schockaert]
通讯作者: Steven Schockaert
DOI: 10.3233/978-1-61499-419-0-243
发表时间: 2014-08
期刊:
影响因子: --
作者: [J. Derrac;Steven Schockaert]
通讯作者: J. Derrac;Steven Schockaert
Reasoning about Structured Story Representations
  • 批准号:
    EP/W003309/1
  • 项目类别:
    Fellowship
  • 资助金额:
    $163.22万
  • 财政年份:
    2022
  • 负责人:
    Steven Schockaert
  • 依托单位:
Encyclopedic Lexical Representations for Natural Language Processing
  • 批准号:
    EP/V025961/1
  • 项目类别:
    Research Grant
  • 资助金额:
    $76.1万
  • 财政年份:
    2021
  • 负责人:
    Steven Schockaert
  • 依托单位:
海外基金