Computer-Assisted Methods of Stemmatic Analysis

Computer-Assisted Methods of Stemmatic Analysis
复制标题

计算机辅助的词干分析方法

DOI:
--
复制
发表时间:
1993
期刊:
影响因子:
--
通讯作者:
P. Robinson
P. Robinson
中科院分区:
--
文献类型:
--
作者:
R. O’Hara;P. Robinson

文献摘要

参考文献

被引文献

相似文献

在这篇文章中,我们回顾了可用于坎特伯雷故事计划的计算机辅助词干分析方法。1我们相信,这些技术将使我们能够比曼利和里克特(1940)更准确地重建坎特伯雷故事的历史,这对我们决定开展这项工作至关重要。这些技术有两个主要方面。第一,分支分析,用来快速了解手稿之间的广泛关系。第二,数据库分析,用于在对个别变种及其分布进行审查的基础上,提炼关于特定手稿和群体之间确切关系的结论。除了对这些技巧的讨论外,我们在这里简要报告我们在巴斯的前言手稿的妻子身上测试这些工具的结果,以及其他材料。分支分析对大量中世纪白话传统手稿的整理产生了大量关于手稿之间的一致和分歧的信息,即使在拼写规则之后也是如此。在坎特伯雷故事项目的初步研究期间,在整理的46份手稿中,提供了大约13,000份独立的实质性变体读物。由于每个变体平均出现在15个手稿中,这就提供了大约20万个单独的信息项供审查。乘以所有手稿,然后乘以《坎特伯雷故事集》的所有部分,我们得到的数据量远远超出了人工分类技术的能力。事实上,曼利和里克特似乎没有能力设计出任何应对这种信息洪流的方法,这是他们未能从基因上重建对手稿传统有用的编辑目的的原因。2重建手稿词干的困难,以及校对产生的数据高度结构化的特点,向一些作者表明,计算机辅助技术在迅速指出可能的关系方面可能是有价值的,然后可以用其他手段彻底检查这些关系。3这些方法中最成功和最合适的是支系分析(来自希腊支系)。这项技术是由系统学领域的研究人员在过去30年里发展起来的,系统学是进化生物学的一个分支,专门研究
In this essay, we review the methods of computer-assisted stemmatic analysis available to the Canterbury Tales Project. 1 Our belief that these techniques will permit us to arrive at a more exact reconstruction of the history of the Canterbury Tales than could Manly and Rickert (1940) is vital to our decision to undertake this work. There are two major strands to these techniques. The first, cladistic analysis, is used to gain a rapid overview of the broad relations among the manuscripts. The second, database analysis, is used to refine conclusions about the exact relationships of particular manuscripts and groups, on the basis of scrutiny of individual variants and their distribution. In addition to discussion of these techniques, we briefly report here the results of our testing of these tools on the Wife of Bath’s Prologue manuscripts, among other materials. Cladistic analysis The collation of the manuscripts of a large medieval vernacular tradition yields enormous amounts of information concerning the agreements and disagreements among the manuscripts, even after regularization of spelling. Collation of transcripts of the Wife of Bath’s Prologue manuscripts by the computer collation program Collate, during preliminary studies for the Canterbury Tales Project, supplied around 13,000 separate substantive variant readings among the forty-six manuscripts collated. With each variant occurring in an average of fifteen manuscripts, this gives about two hundred thousand separate items of information to be examined. Multiply this by all the manuscripts, then by all the parts of the Canterbury Tales, and we have a quantity of data far beyond the capacity of manual sorting techniques. Indeed, it appears that the inability of Manly and Rickert to devise any means of coping with this flood of information lies behind their failure to arrive at a genetic reconstruction of the manuscript tradition useful for editorial purposes. 2 The difficulty of reconstructing manuscript stemmata, and the highly structured character of the data that result from collation, have suggested to a number of authors that computer-assisted techniques might prove valuable in pointing quickly to possible relationships which could then be thoroughly examined by other means. 3 The most successful and appropriate of these methods is cladistic analysis (from the Greek clados ‘branch’). This technique has been developed over the last thirty years by researchers in the field of systematics, the branch of evolutionary biology which specializes in the
DOI: 10.1126/science.1590849
发表时间: 1992-02-07
期刊: SCIENCE
影响因子: 56.9
作者:
TEMPLETON, AR
通讯作者: TEMPLETON, AR