A Data-Driven Approach for Extracting "the Most Specific Term" for Ontology Development
A Data-Driven Approach for Extracting "the Most Specific Term" for Ontology Development
复制标题
一种数据驱动的方法,为本体开发提取“最具体的术语”
DOI:
--
复制
发表时间:
2003
期刊:
影响因子:
--
通讯作者:
C. Chute
中科院分区:
文献类型:
--
作者:
G. Savova;M. Harris;Thomas M. Johnson;Serguei V. S. Pakhomov;C. Chute
We present a data-driven approach to extract the "most specific" terms relevant to an ontology of functioning, disability and health. The algorithm is a combination of statistical and linguistic approaches. The statistical filter is based on the frequency of the content words in a given text string; the linguistic heuristic is an extension of existing algorithms but goes beyond noun phrases and is formulated as a "complete syntactic node". Thus, it can be applied to any syntactic node of interest in the particular domain. Two test sets were marked by three experts. Test set 1 is a well-constructed text from pain abstracts; test set 2 is actual medical reports. Results are reported as recall, precision, F-score and rate of valid terms in false positives. A limitation of the current research is the relatively small test set.
影响因子:
5.4
作者:
VERBRUGGE, LM;JETTE, AM
通讯作者:
JETTE, AM