From medical language processing to BioNLP domain

From medical language processing to BioNLP domain
复制标题

从医学语言处理到 BioNLP 领域

DOI:
--
复制
发表时间:
2012
期刊:
--
影响因子:
--
通讯作者:
S. Biagioni
S. Biagioni
中科院分区:
--
文献类型:
--
作者:
Gabriella Pardelli;M. Sassi;S. Goggi;S. Biagioni

文献摘要

被引文献

相似文献

本文介绍了在生物医学领域的参考语料库上进行的术语学工作的结果。特别是,这项研究倾向于分析生物医学中某些术语的使用情况,以核实它们随时间的变化,目的是从网上检索文件的本质。术语样本包含BioNLP和生物医学中使用的词汇,并确定哪些术语是从科学出版物传递到日报上的,哪些是相当保留用于科学生产的。这项工作的最终范围是确定如何将科学传播到社会的更大部分,使普通公民的公众能够接触到关于生物医学研究和开发的交流;其主要来源是一个参考语料库,由三个主要储存库组成,从中提取与BioNLP和Biomedicine有关的信息。本文件分为三个部分:1)专门介绍从科学文献中提取的数据;2)第二部分专门介绍方法和数据描述;3)第三部分包含从档案中提取的术语的统计表示:索引和语料库可以反映这一领域中某些术语的使用情况,并为在数字时代获取知识提供可能的关键。
This paper presents the results of a terminological work on a reference corpus in the domain of Biomedicine. In particular, the research tends to analyse the use of certain terms in Biomedicine in order to verify their change over the time with the aim of retrieving from the net the very essence of documentation. The terminological sample contains words used in BioNLP and biomedicine and identifies which terms are passing from scientific publications to the daily press and which are rather reserved to scientific production. The final scope of this work is to determine how scientific dissemination to an ever larger part of the society enables a public of common citizens to approach communication on biomedical research and development; and its main source is a reference corpus made up of three main repositories from which information related to BioNLP and Biomedicine is extracted. The paper is divided in three sections: 1) an introduction dedicated to data extracted from scientific documentation; 2) the second section devoted to methodology and data description; 3) the third part containing a statistical representation of terms extracted from the archive: indexes and concordances allow to reflect on the use of certain terms in this field and give possible keys for having access to the extraction of knowledge in the digital era.