Methods and Results in Language Documentation using Literary Digital Editions
Methods and Results in Language Documentation using Literary Digital Editions
批准号:
2109679
负责人:
Allison Bigelow
金额:
$24.91万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2022
资助国家:
美国
项目状态:
已结题
起止时间:
2022-01-01 至 2024-06-30
中文摘要
殖民档案对学者提出了一系列方法上的挑战。一方面,它们包含了关于前欧洲和早期殖民时期土著语言,历史和社会的宝贵信息。另一方面,这些文件往往是由宗教和政治官员撰写的,他们试图改变或根除他们所描述的传统。为了解决这个众所周知的问题,学者们必须开发出一种方法,将殖民干预与土著语言,历史和文化数据分开。这个项目提供了这样一个解决方案。该项目与社区的土著学者合作,使用开源、符合标准的工具,建立了一个数字收藏,其中包括1492年以前美洲最长、最完整的五个版本。通过展示原住民和非原住民作者以不同的方式和在不同的时刻记录词汇,句法结构和形态,该项目生成了一个语料库和方法,可以应用于其他语言和历史背景,包括未来的教育工作。由于有关数字材料对遗产学习者和非遗产发言者的土著语言习得的有效性的数据很少,该项目为未来的研究创造了重要的材料。该项目将文本编码倡议(TEI)开发的文本编码的国际学术标准应用于最长和最完整的1492年前的书,以幸存下来。其他数字项目以手稿传真的形式呈现文本,这种格式对非专家来说是无法访问的,不容易用于语言学习,并且优先考虑殖民时代的写作行为,而不是原始的口头传统。相比之下,该项目的五项数字收集的历史和现代文本,视频和翻译,由土著学者与非土著教师和学生研究人员合作开发,使用编码工具来揭示殖民干预,并突出土著学者对正字法,形态句法和词汇的纠正。这些工具包括引用字符串分析属性rs ana来标记文本中的字符,地点和对象,空间元素来保留诗意的口头性并根据形态学平行性和短语结尾标记对齐概念,以及注释,通过880个文化主题的关系数据库进行管理,以呈现对文本段落,语音和句法的竞争性学术解释。该集合位于一个平台上,该平台使用开源静态站点生成器框架来避免基于服务器的网站的长加载时间,与渐进式Web应用程序框架相结合,为移动的读者提供类似应用程序的体验,而不会牺牲手机存储。这种历史文本分析的数字人文方法允许各种用户,包括研究人员,教师,学生和感兴趣的公众成员,以无法通过印刷传播的方式表示语料库和代码。通过这样做,它为母语者、传统学习者和非母语研究人员和学生的语言学习、文本分析和语言研究创造了一个新的工具。该奖项反映了NSF的法定使命,并被认为值得通过使用基金会的智力价值和更广泛的影响审查标准进行评估来支持。
英文摘要
Colonial archives present a range of methodological challenges for scholars. On the one hand, they contain valuable information about Indigenous languages, histories, and societies in the Pre-European and early colonial periods. On the other hand, these documents are often written by religious and political officials who sought to change or eradicate the traditions they described. To address this well-known problem, scholars must develop methods that disentangle colonial interventions from Indigenous linguistic, historical, and cultural data. This project offers one such solution. In collaboration with Indigenous scholars from the community, this project uses open source, standards-compliant tools to build a digital collection of five versions of the longest and most complete pre-1492 book of the Americas. By showing where Indigenous and non-Native authors record vocabularies, syntactic structures, and morphologies in different ways and at different moments, this project generates a corpus and methodology that can be applied to other linguistic and historical contexts, including future educational efforts. Because there is little data on the effectiveness of digital materials on Indigenous language acquisition for heritage learners and non-heritage speakers, this project creates important materials for future research.This project applies the international, scholarly standards of textual encoding developed by the Text Encoding Initiative (TEI) to the longest and most complete pre-1492 book to survive the conquest. Other digital projects represent the text in manuscript facsimiles, a format that is inaccessible to non-experts, cannot easily be used for language learning, and privileges the colonial-era act of writing rather than the original oral tradition. In contrast, this project’s five-item digital collection of historical and modern texts, videos, and translations, developed by Indigenous scholars in collaboration with non-Native faculty and student researchers, uses encoding tools to reveal colonial interventions and highlight Indigenous scholars’ corrections to orthography, morphosyntax, and vocabulary. These tools include referencing string analytical attributes rs ana to tag characters, places, and objects in the text, the space element to preserve poetic orality and align concepts according to morphological parallelism and phrase-final markers, and annotations, managed through a relational database of 880 cultural topics, to present competing scholarly interpretations of textual passages, phonetics, and syntax. The collection is housed on a platform that uses an open-source Static Site Generator framework to avoid the long loading time of server-based websites, paired with a Progressive Web Application framework to provide mobile readers with an app-like experience that does not sacrifice phone storage. This digital humanities approach to historical textual analytics allows a variety of users, including researchers, teachers, students, and interested members of the public, to represent the corpus and code in ways that cannot be disseminated through print. In so doing it creates a new tool for language learning, textual analysis, and linguistic study by native speakers, heritage learners, and non-Native researchers and students.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金