Moara: a Java library for extracting and normalizing gene and protein mentions.

Moara: a Java library for extracting and normalizing gene and protein mentions.
复制标题

DOI:
10.1186/1471-2105-11-157
复制
发表时间:
2010-03-26
期刊:
影响因子:
3
通讯作者:
Pascual-Montano A
Pascual-Montano A
中科院分区:
生物学4区
文献类型:
--
作者:
Neves ML;Carazo JM;Pascual-Montano A

文献摘要

参考文献

被引文献

相似文献

Gene/protein recognition and normalization are important preliminary steps for many biological text mining tasks, such as information retrieval, protein-protein interactions, and extraction of semantic information, among others. Despite dedication to these problems and effective solutions being reported, easily integrated tools to perform these tasks are not readily available. This study proposes a versatile and trainable Java library that implements gene/protein tagger and normalization steps based on machine learning approaches. The system has been trained for several model organisms and corpora but can be expanded to support new organisms and documents. Moara is a flexible, trainable and open-source system that is not specifically orientated to any organism and therefore does not requires specific tuning in the algorithms or dictionaries utilized. Moara can be used as a stand-alone application or can be incorporated in the workflow of a more general text mining system.
DOI: 10.1186/1471-2105-6-s1-s15
发表时间: 2005
期刊: BMC bioinformatics
影响因子: 3
作者:
Fundel K;Güttler D;Zimmer R;Apostolakis J
通讯作者: Apostolakis J
DOI: 10.1093/nar/gkn664
发表时间: 2009-01
影响因子: 14.9
作者:
UniProt Consortium
通讯作者: UniProt Consortium
DOI: 10.1093/bioinformatics/btm557
发表时间: 2008-01-15
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Rebholz-Schuhmann, Dietrich;Arregui, Miguel;Jimeno, Antonio
通讯作者: Jimeno, Antonio
DOI: 10.1007/s00439-001-0615-0
发表时间: 2001-12-01
期刊: HUMAN GENETICS
影响因子: 5.3
作者:
Povey, S;Lovering, R;Wain, H
通讯作者: Wain, H
DOI: 10.1186/1471-2105-6-s1-s13
发表时间: 2005
期刊: BMC bioinformatics
影响因子: 3
作者:
Crim J;McDonald R;Pereira F
通讯作者: Pereira F