课题基金 / 基金详情

Language Preservation 2.0: Crowdsourcing Oral Language Documentation using Mobile Devices

Language Preservation 2.0: Crowdsourcing Oral Language Documentation using Mobile Devices
语言保存2.0:使用移动设备众包口语文档
批准号:
1160639
负责人:
Mark Liberman
金额:
$10.15万
依托单位:
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2012
资助国家:
美国
项目状态:
已结题
起止时间:
2012-07-01 至 2014-12-31

项目摘要

项目成果

Mark Liberman的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Language Preservation 2.0The purpose of this pilot project is to demonstrate the feasibility of a new approach to documenting endangered languages.To allow wide-ranging investigation of a language even after it is no longer spoken, we need the equivalent of the million words of extant biblical Hebrew texts, or the five million words of extant classical Latin. But for endangered languages without a significant culture of literacy, diverse text collections on this scale seem out of reach. Given typical speaking rates of about 10,000 word-equivalents per hour, a hundred hours of recorded speech -- conversations, narratives, or oral histories -- would give us the equivalent of a million words of text. With community involvement, hundreds of hours of such recordings are easily within reach.However, transcribing such large audio collections is a daunting task, given the small number of literate native speakers and the time-consuming nature of such transcription, which can take 200 hours of work for every hour of audio. We propose to solve this problem by substituting re-speaking and verbal translation: one or more native speakers repeats each phrase of a recording, speaking slowly and carefully, and then translates it into a better-documented language.The utility of translated passages as a way to analyze otherwise-unknown languages has been demonstrated many times, starting with the Rosetta Stone. This aspect of our task is easier, since at least a grammatical sketch will in general be available. Our goal in this project is to demonstrate the utility of re-speaking. We believe that linguists, starting out with relatively little knowledge of a language, can produce phonetic transcriptions that will be good enough to support subsequent analysis resulting in coherent texts, in a process analogous to (but easier than) the process that allowed previous generations of scholars to learn to read ancient Egyptian or Sumerian.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
CI-NEW: NIEUW: Novel Incentives and Workflows in Linguistic Data Collection and Annotation
  • 批准号:
    1730377
  • 项目类别:
    Standard Grant
  • 资助金额:
    $121.85万
  • 财政年份:
    2017
  • 负责人:
    Mark Liberman
  • 依托单位:
EAGER: Mining a Year of Speech
  • 批准号:
    1048900
  • 项目类别:
    Standard Grant
  • 资助金额:
    $9.99万
  • 财政年份:
    2010
  • 负责人:
    Mark Liberman
  • 依托单位:
Prosodic Systems in New Guinea: Integrating computational and typological approaches to linguistic analysis
  • 批准号:
    0951651
  • 项目类别:
    Standard Grant
  • 资助金额:
    $29.93万
  • 财政年份:
    2010
  • 负责人:
    Mark Liberman
  • 依托单位:
Collaborative Research: OLAC: Accessing the World's Language Resources
  • 批准号:
    0723357
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $14.7万
  • 财政年份:
    2007
  • 负责人:
    Mark Liberman
  • 依托单位:
海外基金