课题基金 / 基金详情

Automatic evaluation of speech quality

Automatic evaluation of speech quality
语音质量自动评估
批准号:
6882741
负责人:
CHARLES S WATSON
金额:
$10.0万
依托单位国家:
美国
项目类别:
财政年份:
2004
资助国家:
美国
项目状态:
已结题
起止时间:
2004-09-24 至 2005-09-30

项目摘要

项目成果

CHARLES S WATSON的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
DESCRIPTION (provided by applicant): Tests of several different approaches to the automatic evaluation of the quality of speech segments are proposed. Previous systems for use in pronunciation training have typically employed either automatic speech-recognition (ASR) technology, or have used templates based on a limited number of utterances rated as excellent by L1 listeners (and sometimes also employing a second set of utterances containing a common pronunciation error). Here speech-processing technologies (HMM's and ANN's) will be developed specifically for use as evaluation systems (not recognition systems) to predict quality and locus-of-error judgments assigned by listeners. Termed the "evaluation-of-single-words" (ESW) approach, the special feature of these systems will derive from the training tokens employed in their development: multiple recordings of a single word made by groups of native and non-native talkers. Sixty talkers will be native speakers of Arabic, whose intelligibility in English ranges from poor to near-perfect, and 60 talkers will be native speakers of middle-American English. There will be twelve words divided between one, two, and three syllables. Ten productions of each word will be recorded by each talker, yielding 14,400 tokens. Each token will be rated by listening juries for pronunciation quality, and the tokens will also be categorized into perceptual clusters, using MDS and cluster-analysis techniques. At least two computer-based evaluation systems (HMM and ANN) will be trained for each individual word, with the goals of predicting overall pronunciation quality and identifying specific commonly occurring pronunciation errors. It is expected that these word-specific systems, each representing a discrete "evaluator" custom-built for an individual word, will approach the maximum accuracy that can be expected of this class of processors. If successful, the ESW approach may have a broad range of applications in pronunciation training, identification of a speaker's L1, foreign-language instruction, and other non-lexical applications. However, our specific goal is the development of systems that can provide informative feedback during automated pronunciation training. In ASR applications, the goal is to respond the same way to a word, no matter how it is pronounced. The goal of an ESW system is to respond differentially to pronunciation variants. This distinction between ASR and ESW is central to the development of successful evaluation systems as it dictates different modeling constraints.
期刊论文(1)
专著(0)
科研奖励(0)
会议论文
Multi-site study of speech perception training for hearing-aid users
Multi-site study of speech perception training for hearing-aid users
Multi-site study of speech perception training for hearing-aid users
Multi-site study of speech perception training for hearing-aid users
国内基金
海外基金
基于MFSD2A调控血迷路屏障跨细胞囊泡转运机制的噪声性听力损失防治研究
  • 批准号:
    82371144
  • 项目类别:
    面上项目
  • 资助金额:
    49.00万元
  • 批准年份:
    2023
  • 负责人:
    汪雪玲
  • 依托单位:
YTHDF1通过m6A修饰调控耳蜗毛细胞炎症反应在老年性聋中的作用机制研究
  • 批准号:
    82371140
  • 项目类别:
    面上项目
  • 资助金额:
    49.00万元
  • 批准年份:
    2023
  • 负责人:
    李姝娜
  • 依托单位:
TRIM21蛋白促进HIF1α的降解介导耳蜗血管纹缘细胞缺血再灌注致听力损伤的机制研究
  • 批准号:
    82371142
  • 项目类别:
    面上项目
  • 资助金额:
    49.00万元
  • 批准年份:
    2023
  • 负责人:
    刘君
  • 依托单位:
基于WHO-HEARING理论框架的老年人听力障碍社区康复模式构建与优化策略研究
  • 批准号:
    --
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    30万元
  • 批准年份:
    2022
  • 负责人:
    江帆
  • 依托单位: