A Large Scale Terminology Resource for Biomedical Text Processing
A Large Scale Terminology Resource for Biomedical Text Processing
复制标题
用于生物医学文本处理的大规模术语资源
DOI:
--
复制
发表时间:
2004
期刊:
影响因子:
--
通讯作者:
Yikun Guo
中科院分区:
文献类型:
--
作者:
H. Harkema;R. Gaizauskas;Mark Hepple;A. Roberts;Ian Roberts;Neil Davis;Yikun Guo
In this paper we discuss the design, implementation, and use of Termino, a large scale terminological resource for text processing. Dealing with terminology is a difficult but unavoidable task for language processing applications, such as Information Extraction in technical domains. Complex, heterogeneous information must be stored about large numbers of terms. At the same time term recognition must be performed in realistic times. Termino attempts to reconcile this tension by maintaining a flexible, extensible relational database for storing terminological information and compiling finite state machines from this database to do term lookup. While Termino has been developed for biomedical applications, its general design allows it to be used for term processing in any domain.