Introduction to the tm Package Text Mining in R

Introduction to the tm Package Text Mining in R
复制标题

R中tm包文本挖掘简介

DOI:
--
复制
发表时间:
2007
期刊:
影响因子:
--
通讯作者:
Ingo Feinerer
Ingo Feinerer
中科院分区:
--
文献类型:
--
作者:
Ingo Feinerer

文献摘要

被引文献

相似文献

这篇短文简要介绍了如何利用tm包提供的文本挖掘框架在R中进行文本挖掘。我们介绍了数据导入、语料库处理、预处理、元数据管理和术语文档矩阵创建的方法。我们的重点是开始在r中进行文本挖掘的主要方面,tm提供的文本挖掘基础设施的深入描述发表在统计软件杂志上(Feinerer et al., 2008)。一篇关于R文本挖掘的介绍性文章发表在R News (Feinerer, 2008)上。
This vignette gives a short introduction to text mining in R utilizing the text mining framework provided by the tm package. We present methods for data import, corpus handling, preprocessing, metadata management, and creation of term-document matrices. Our focus is on the main aspects of getting started with text mining in R—an in-depth description of the text mining infrastructure offered by tm was published in the Journal of Statistical Software (Feinerer et al., 2008). An introductory article on text mining in R was published in R News (Feinerer, 2008).