课题基金 / 基金详情

Collaborative Research: RI: Medium: Flexible Deep Speech Synthesis through Gestural Modeling

Collaborative Research: RI: Medium: Flexible Deep Speech Synthesis through Gestural Modeling
合作研究:RI:Medium:通过手势建模进行灵活的深度语音合成
批准号:
2106930
负责人:
Louis Goldstein
金额:
$40.0万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2021
资助国家:
美国
项目状态:
已结题
起止时间:
2021-10-01 至 2024-09-30

项目摘要

项目成果

Louis Goldstein的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Voice based interactions have become the norm everywhere from cars, to mobile phones to digital home assistants. As speech based machine interaction becomes more pervasive, there is increased demand and expectation of human-like performance and personality from these systems. It is important for the machine to deliver responses about the weather on a pleasant sunny day or an impending hurricane in an appropriate manner. Machines need to be able to respond sympathetically or emphatically depending on the context of their use. Critically, when machines fail, they should do so in human understandable ways, so that there are no unintended consequences of technology. This project aims to create more natural and flexible speech synthesis technology that is inspired by human strategies and mechanisms for speech production. Bringing together the science of speech production and current state-of-the-art engineering speech systems, this project aims to impart explainability, naturalness and flexibility to speech technologies. This project has the potential to impact all systems that use speech output like automated tutoring, interactive voice response, speech translation in commercial and military settings, digital assistants, robotics and rehabilitative healthcare applications like Brain-Computer Interfaces. Current speech synthesis techniques are focused on end-to-end systems, avoiding explicit modeling of internal structure of the speech signal. Consequently, such systems may have good results but fail to allow any generalization beyond their recorded databases. This project concentrates on incorporating aspects of human speech production into computer speech synthesis. Using data-driven techniques and vocal tract imaging datasets, the project aims to discover and model compositional aspects of the speech signal as described by Articulatory Phonology. Novel deep-learning based approaches will be developed for joint optimization of diverse speech representations such as acoustic, phonological and physiological data within an analysis-by-synthesis framework. New strategies will be developed for incorporating grounded representations into text-to-speech training and evaluated in a range of applications in flexible speech synthesis.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(5)
专著(0)
科研奖励(0)
会议论文
Deep Neural Convolutive Matrix Factorization for Articulatory Representation Decomposition
用于发音表示分解的深度神经卷积矩阵分解
DOI: 10.21437/interspeech.2022-11233
发表时间: 2022
期刊: Interspeech 2022
影响因子: --
作者: [Lian, Jiachen, Black, Alan W, Goldstein, Louis, Anumanchipalli, Gopala Krishna]
通讯作者: Anumanchipalli, Gopala Krishna
DOI: 10.21437/interspeech.2023-2316
发表时间: 2023-07
期刊:
影响因子: --
作者: [Peter Wu;Tingle Li;Yijingxiu Lu;Yubin Zhang;Jiachen Lian;A. Black;L. Goldstein;Shinji Watanabe;G. Anumanchipalli]
通讯作者: Peter Wu;Tingle Li;Yijingxiu Lu;Yubin Zhang;Jiachen Lian;A. Black;L. Goldstein;Shinji Watanabe;G. Anumanchipalli
DOI: 10.21437/interspeech.2022-10892
发表时间: 2022-09
期刊:
影响因子: --
作者: [Peter Wu;Shinji Watanabe;L. Goldstein;A. Black;G. Anumanchipalli]
通讯作者: Peter Wu;Shinji Watanabe;L. Goldstein;A. Black;G. Anumanchipalli
DOI: 10.1109/icassp49357.2023.10096401
发表时间: 2022-10
期刊: ICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
影响因子: --
作者: [Jiachen Lian;A. Black;Yijingxiu Lu;L. Goldstein;Shinji Watanabe;G. Anumanchipalli]
通讯作者: Jiachen Lian;A. Black;Yijingxiu Lu;L. Goldstein;Shinji Watanabe;G. Anumanchipalli
Collaborative Research: Prosodic Structure: An Integrated Empirical and Modeling Investigation
  • 批准号:
    1551695
  • 项目类别:
    Standard Grant
  • 资助金额:
    $11.11万
  • 财政年份:
    2016
  • 负责人:
    Louis Goldstein
  • 依托单位:
Collaborative Research: Landmark-based Robust Speech Recognition Using Prosody-guided models of speech variability
  • 批准号:
    0703048
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $0.0万
  • 财政年份:
    2007
  • 负责人:
    Louis Goldstein
  • 依托单位:
Laboratory Phonology Conference, Yale University, June, 2002
  • 批准号:
    0132005
  • 项目类别:
    Standard Grant
  • 资助金额:
    $2.6万
  • 财政年份:
    2002
  • 负责人:
    Louis Goldstein
  • 依托单位:
Modeling Phonetic Structure Using Articulatory Dynamics
  • 批准号:
    9514730
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $37.14万
  • 财政年份:
    1996
  • 负责人:
    Louis Goldstein
  • 依托单位:
国内基金
海外基金
Research on Quantum Field Theory without a Lagrangian Description
  • 批准号:
    24ZR1403900
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2024
  • 负责人:
    SATOSHI NAWATA
  • 依托单位:
Cell Research
Cell Research
Cell Research (细胞研究)