Exploitation of Diverse Data via Automatic Adaptation of Knowledge Extraction Software
Exploitation of Diverse Data via Automatic Adaptation of Knowledge Extraction Software
批准号:
100934
负责人:
金额:
$21.01万
依托单位:
依托单位国家:
英国
项目类别:
Collaborative R&D
财政年份:
2011
资助国家:
英国
项目状态:
已结题
起止时间:
2011 至 --
中文摘要
点击翻译按钮获取中文摘要
英文摘要
The current generation of language processing has had considerable success in extracting useful information from large amounts of unstructured text, whether this is research literature or social media. However, adapting to a new domain is often a laborious process, with respect both to diverse types of data (e.g. newswire vs. patent literature) and to the terminology used in a given domain (e.g. in medical practice vs. pharmaceutical research). Humans can perform these tasks on small data sets, but face a challenge in the face of massively increasing amounts of electronic text. The EVOKES project is exploiting distributional similarity techniques to accelerate key components of customisation - the recognition of concepts, and the creation or adaptation of terminologies that link terms to concepts.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金