课题基金 / 基金详情

RI: Small: A New Voice Source Model: From Glottal Areas to Better Speech Synthesis

RI: Small: A New Voice Source Model: From Glottal Areas to Better Speech Synthesis
RI:Small:一种新的语音源模型:从声门区域到更好的语音合成
批准号:
1018863
负责人:
Abeer Alwan
金额:
$45.0万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2010
资助国家:
美国
项目状态:
已结题
起止时间:
2010-09-01 至 2015-07-31
关键词:

项目摘要

项目成果

Abeer Alwan的其他基金

相似基金

相关文献

中文摘要
翻译
本研究的目的是在对30名成人说话者的声带进行生理观察的基础上,开发和评估一个新的声源模型。现有来源模型的缺点部分归因于它们的开发方式:基于少数说话者的有限数据,没有直接的生理观察,也没有感知验证。更大的数据集不仅有助于开发一个可以解释扬声器内部和跨扬声器的一系列语音质量的源模型,而且还有助于理解哪些模型参数是扬声器和/或性别特定的,以及哪些模型参数是特定的。模型开发将从最早阶段考虑模型参数的感知效应。更好的源模型还可以提高语音处理算法的性能,例如文本到语音合成(TTS)。通常在这类算法的开发中,重点一直放在与语音频谱包络相关的声学特征上。另一方面,声源的声学特性受到的关注较少。提出的工作包括:1)同时记录15名男性和15名女性说话者的声带振动的高速图像;2)从图像中提取声门区域功能以参数化新的声源模型;3)进行感知实验以揭示哪些模型参数在感知上显着;4)在TTS中使用新的声源模型。该项目的跨学科团队(拥有建模、合成、识别、语音学和心理语言学方面的专业知识)是唯一有资格进行这项变革性研究的人。
英文摘要
The goal of the proposed research is to develop and evaluate a new voicesource model based on physiological observations of the vocal folds of 30 adult speakers. Shortcomings of existing source models can be in part attributed to the way in which they were developed: based on limited data from a few speakers, without direct physiological observations, and without perceptual validation. A larger dataset would help in not only developing a source model that could account for a range of voice qualities within and across speakers, but also result in an understanding of how and which model parameter(s) are speaker and/or gender specific. Model development will consider the perceptual effects of the model's parameters from the earliest stages.A better source model might also improve the performance of speech processing algorithms such as text-to-speech synthesis (TTS). Typically in the development of such algorithms, the emphasis has been on acoustic features related to the speech spectral envelope. The acoustics of the voice source, on the other hand, have received less attention. The proposed work involves: 1) recording high-speed images of vocal foldvibrations with simultaneous audio recordings from 15 male and 15 female speakers, 2) extracting glottal area functions from the images to parameterize a new voice source model, 3) performing perception experiments to uncover which model parameters are perceptually salient, and 4) using the new voice source model in TTS. The project's interdisciplinary team (with expertise in modeling, synthesis, recognition, phonetics, and psycholinguistics) is uniquely qualified to conduct this transformative research.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Collaborative Research: Improving speech technology for better learning outcomes: the case of AAE child speakers
Collaborative Research: RI: Small: From Ultrasound and MRI to articulatory and acoustic models of child speech development
Workshop for Undergraduate and MS Female Students in Speech Science and Technology
NRI: INT: COLLAB: Development, Deployment and Evaluation of Personalized Learning Companion Robots for Early Literacy and Language Learning
国内基金
海外基金
昼夜节律性small RNA在血斑形成时间推断中的法医学应用研究
  • 批准号:
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2024
  • 负责人:
  • 依托单位:
tRNA-derived small RNA上调YBX1/CCL5通路参与硼替佐米诱导慢性疼痛的机制研究
  • 批准号:
  • 项目类别:
    省市级项目
  • 资助金额:
    10.0万元
  • 批准年份:
    2022
  • 负责人:
    张祥忠
  • 依托单位:
Small RNA调控I-F型CRISPR-Cas适应性免疫性的应答及分子机制
Small RNAs调控解淀粉芽胞杆菌FZB42生防功能的机制研究
  • 批准号:
    31972324
  • 项目类别:
    面上项目
  • 资助金额:
    58.0万元
  • 批准年份:
    2019
  • 负责人:
    高学文
  • 依托单位: