课题基金 / 基金详情

Automatic speech recognition for mobile and personal applications

Automatic speech recognition for mobile and personal applications
适用于移动和个人应用的自动语音识别
批准号:
336627-2006
负责人:
OShaughnessy, Douglas
金额:
$8.74万
依托单位国家:
加拿大
项目类别:
Strategic Projects - Group
财政年份:
2007
资助国家:
加拿大
项目状态:
已结题
起止时间:
2007-01-01 至 2008-12-31

项目摘要

项目成果

OShaughnessy, Douglas的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
In this proposal, we examine ways to improve the performance of automatic speech recognition (ASR), whichconverts ordinary speech into text. This allows unsupervised programs to interact with people via voice, thusgiving people access to a broad range of data services, over telephones or PDAs (personal digital assistants).ASR exists now only for limited services, as its accuracy is much lower than that of human listeners. For ASRto gain widespread use, we need to make it much more efficient. This proposal addresses several areas oftargeted interest, e.g., accessing databases over wireless channels. The research deals with intelligent signalprocessing, as we need to develop algorithms to interpret voice signals in terms of their intent.The simplicity of models used in current ASR, when compared to the complex, nonlinear processing done byhumans, suggests that there remain many ways to improve ASR. At each level (i.e., feature estimation,temporal and acoustic modeling, language modeling, decision making), compromises have been made in ASRto have simple and fast processing, at the expense of lower recognition accuracy. We intend to examineseveral of these compromises and alternative ways to approach the performance of human speech recognition.A major application of our work is to render search and access to information via voice, e.g., large databasesas found on the World Wide Web. Typing on small portable devices is awkward; thus access via voice ishighly desired.Human-computer interactions often use a mouse+keyboard as machine input, and a computer screen orprinter as output. Speech, however, has always had a high priority in human communication, developed longbefore writing. While less economical in communication bandwidth than text, speech is more universal andeasy to use than text processing, e.g., people speak more quickly and easily than typing. The proposed researchhere deals with improving ASR to render it more practical and hence more useful.On recherche nouvelles méthodes afin d'améliorer la performance de systèmes de reconnaissance automatiquede parole (RAP), ce qui convertit la parole ordinaire en texte. Ceci permet aux algorithmes non dirigés àinteragir avec les gens par la parole, ce qui leur donne accès à une grande gamme de services de données, partéléphone ou par assistants numériques.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
More efficient and accurate automatic speech recognition
More efficient and accurate automatic speech recognition
More efficient and accurate automatic speech recognition
More efficient and accurate automatic speech recognition
国内基金
海外基金
儿童植入耳蜗后听觉行为与言语发展进程的关联性研究
  • 批准号:
    81170916
  • 项目类别:
    面上项目
  • 资助金额:
    65.0万元
  • 批准年份:
    2011
  • 负责人:
    刘莎
  • 依托单位:
儿童植入人工耳蜗后开放式听觉言语发育特性研究
  • 批准号:
    30872859
  • 项目类别:
    面上项目
  • 资助金额:
    30.0万元
  • 批准年份:
    2008
  • 负责人:
    刘莎
  • 依托单位: