Collaborative Research: Landmark-based Robust Speech Recognition Using Prosody-guided models of speech variability
Collaborative Research: Landmark-based Robust Speech Recognition Using Prosody-guided models of speech variability
批准号:
0703048
负责人:
Louis Goldstein
金额:
$0.0万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2007
资助国家:
美国
项目状态:
已结题
起止时间:
2007-06-01 至 2011-05-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Proposal ID 0703859 Date 04/11/2007 Despite great strides in the development of automatic speech recognition technology, we do not yet have a system with performance comparable to humans in automatically transcribing unrestricted conversational speech, representing many speakers and dialects, and embedded in adverse acoustic environments. This approach applies new high-dimensional machine learning techniques, constrained by empirical and theoretical studies of speech production and perception, to learn from data the information structures that human listeners extract from speech. To do this, we will develop large-vocabulary psychologically realistic models of speech acoustics, pronunciation variability, prosody, and syntax by deriving knowledge representations that reflect those proposed for human speech production and speech perception, using machine learning techniques to adjust the parameters of all knowledge representations simultaneously in order to minimize the structural risk of the recognizer. The team will develop nonlinear acoustic landmark detectors and pattern classifiers that integrate auditory-based signal processing and acoustic phonetic processing, are invariant to noise, change in speaker characteristics and reverberation, and can be learned in a semi-supervised fashion from labeled and unlabeled data. In addition, they will use variable frame rate analysis, which will allow for multi-resolution analysis, as well as implement lexical access based on gesture, using a variety of training data. The work will improve communication and collaboration between people and machines and also improve understanding of how human produce and perceive speech. The work brings together a team of experts in speech processing, acoustic phonetics, prosody, gestural phonology, statistical pattern matching, language modeling, and speech perception, with faculty across engineering, computer science and linguistics. Support and engagement of students and postdoctoral fellows are part of the project, engaging in speech modeling and algorithm development. Finally, the proposed work will result in a set of databases and tools that will be disseminated to serve the research and education community at large.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Collaborative Research: RI: Medium: Flexible Deep Speech Synthesis through Gestural Modeling
-
批准号:2106930
-
项目类别:Standard Grant
-
资助金额:$40.0万
-
财政年份:2021
-
负责人:Louis Goldstein
-
依托单位:
Collaborative Research: Prosodic Structure: An Integrated Empirical and Modeling Investigation
-
批准号:1551695
-
项目类别:Standard Grant
-
资助金额:$11.11万
-
财政年份:2016
-
负责人:Louis Goldstein
-
依托单位:
Laboratory Phonology Conference, Yale University, June, 2002
-
批准号:0132005
-
项目类别:Standard Grant
-
资助金额:$2.6万
-
财政年份:2002
-
负责人:Louis Goldstein
-
依托单位:
Modeling Phonetic Structure Using Articulatory Dynamics
-
批准号:9514730
-
项目类别:Continuing Grant
-
资助金额:$37.14万
-
财政年份:1996
-
负责人:Louis Goldstein
-
依托单位:
国内基金
海外基金
登录
查看更多内容
Research on Quantum Field Theory without a Lagrangian Description
-
批准号:24ZR1403900
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:SATOSHI NAWATA
-
依托单位:
Cell Research
-
批准号:31224802
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2012
-
负责人:程磊
-
依托单位:
Cell Research
-
批准号:31024804
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2010
-
负责人:程磊
-
依托单位:
Cell Research (细胞研究)
-
批准号:30824808
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2008
-
负责人:张爱兰
-
依托单位:
Research on the Rapid Growth Mechanism of KDP Crystal
-
批准号:10774081
-
项目类别:面上项目
-
资助金额:45.0万元
-
批准年份:2007
-
负责人:滕冰
-
依托单位: