Exploring Collections of Tagged Text for Literary Scholarship

Exploring Collections of Tagged Text for Literary Scholarship
复制标题

DOI:
10.1111/j.1467-8659.2011.01922.x
复制
发表时间:
2011-01-01
影响因子:
2.5
通讯作者:
Gleicher, M.
Gleicher, M.
中科院分区:
计算机科学4区
文献类型:
--
作者:
Correll, M.;Witmore, M.;Gleicher, M.

文献摘要

被引文献

相似文献

现代文学学者必须将获取大量文本与传统的对其领域的密切分析结合起来。在本文中,我们讨论了支持这项工作的工具的设计和开发。基于对文学学者需求的分析,我们构建了一套可视化工具,用于分析大量标记文本(即一个或多个单词被注释为属于特定类别的文本)。这些工具统一了学者们工作的各个方面:大规模概述工具有助于识别语料库范围内的统计模式,而精细尺度分析工具有助于找到支持这些观察结果的具体细节。我们设计了可视化工具来支持和集成这些层次的分析。结果是第一个可以支持学者进行多层次文本分析的工具套件,将标准的视觉元素与选择单个文本和识别其中代表性段落的新方法相结合。
Modern literary scholars must combine access to vast collections of text with the traditional close analysis of their field. In this paper, we discuss the design and development of tools to support this work. Based on analysis of the needs of literary scholars, we constructed a suite of visualization tools for the analysis of large collections of tagged text (i.e. text where one or more words have been annotated as belonging to a specific category). These tools unite the aspects of the scholars' work: large scale overview tools help to identify corpus-wide statistical patterns while fine scale analysis tools assist in finding specific details that support these observations. We designed visual tools that support and integrate these levels of analysis. The result is the first tool suite that can support the multi-level text analysis performed by scholars, combining standard visual elements with novel methods for selecting individual texts and identifying represenative passages in them.