Document Summarization and Information Extraction for Generation of Presentation Slides
Document Summarization and Information Extraction for Generation of Presentation Slides
复制标题
用于生成演示幻灯片的文档摘要和信息提取
DOI:
10.1109/artcom.2009.74
复制
发表时间:
2009
期刊:
影响因子:
--
通讯作者:
T. Geetha
中科院分区:
文献类型:
--
作者:
Harish Mathivanan;M. Jayaprakasam;K. Prasad;T. Geetha
In this paper, a semi automated technique to generate slide presentations from english text documents is proposed. The technique discussed in this paper is considered to be a pioneering attempt in the field of NLP (Natural Language Processing). The technique involves an information extractor and a slide generator, which combines certain NLP methods such as segmentation, chunking, summarization etc.., with certain special linguistic features of the text such as the ontology of the words, noun phrases found, semantic links, sentence centrality etc., In order to aid the language processing task, two tools can be utilized namely, MontyLingua which helps in chunking and Doddle helps in creating an ontology for the input text represented as an OWL (Ontology Web Language) file. The process of the technique comprises of extracting text, creating an ontology, identifying important phrases for bullets and generating slides.