OpenDMAP: an open source, ontology-driven concept analysis engine, with applications to capturing knowledge regarding protein transport, protein interactions and cell-type-specific gene expression.
OpenDMAP: an open source, ontology-driven concept analysis engine, with applications to capturing knowledge regarding protein transport, protein interactions and cell-type-specific gene expression.
复制标题
OPENDMAP:开源,本体驱动的概念分析引擎,应用于捕获有关蛋白质转运,蛋白质相互作用和细胞类型特异性基因表达的知识。
DOI:
10.1186/1471-2105-9-78
复制
发表时间:
2008-01-31
影响因子:
3
通讯作者:
Cohen KB
中科院分区:
文献类型:
--
作者:
Hunter L;Lu Z;Firby J;Baumgartner WA Jr;Johnson HL;Ogren PV;Cohen KB
Information extraction (IE) efforts are widely acknowledged to be important in harnessing the rapid advance of biomedical knowledge, particularly in areas where important factual information is published in a diverse literature. Here we report on the design, implementation and several evaluations of OpenDMAP, an ontology-driven, integrated concept analysis system. It significantly advances the state of the art in information extraction by leveraging knowledge in ontological resources, integrating diverse text processing applications, and using an expanded pattern language that allows the mixing of syntactic and semantic elements and variable ordering. OpenDMAP information extraction systems were produced for extracting protein transport assertions (transport), protein-protein interaction assertions (interaction) and assertions that a gene is expressed in a cell type (expression). Evaluations were performed on each system, resulting in F-scores ranging from .26 – .72 (precision .39 – .85, recall .16 – .85). Additionally, each of these systems was run over all abstracts in MEDLINE, producing a total of 72,460 transport instances, 265,795 interaction instances and 176,153 expression instances. OpenDMAP advances the performance standards for extracting protein-protein interaction predications from the full texts of biomedical research articles. Furthermore, this level of performance appears to generalize to other information extraction tasks, including extracting information about predicates of more than two arguments. The output of the information extraction system is always constructed from elements of an ontology, ensuring that the knowledge representation is grounded with respect to a carefully constructed model of reality. The results of these efforts can be used to increase the efficiency of manual curation efforts and to provide additional features in systems that integrate multiple sources for information extraction. The open source OpenDMAP code library is freely available at
登录
查看更多内容
影响因子:
3
作者:
Chen H;Sharp BM
通讯作者:
Sharp BM
影响因子:
--
作者:
Blaschke, C;Valencia, A
通讯作者:
Valencia, A
影响因子:
2.9
作者:
Blaschke, Christian;Oliveros, Juan C.;Valencia, Alfonso
通讯作者:
Valencia, Alfonso
影响因子:
5.8
作者:
Corney, DPA;Buxton, BF;Jones, DT
通讯作者:
Jones, DT
影响因子:
7.5
作者:
Bunescu, R;Ge, RF;Wong, YW
通讯作者:
Wong, YW