CAREER: Breaking the phonetic code: novel acoustic-lexical modeling techniques for robust automatic speech recognition
CAREER: Breaking the phonetic code: novel acoustic-lexical modeling techniques for robust automatic speech recognition
批准号:
0643901
负责人:
Eric Fosler-Lussier
金额:
$50.3万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2006
资助国家:
美国
项目状态:
已结题
起止时间:
2006-12-15 至 2012-11-30
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Spontaneous speech, accented speech, and speech in noise continue to provide automatic speech recognition (ASR) technology with significant challenges; error rates of ASR systems are still unacceptably high for these types of speech. This project establishes a consistent framework that seeks to cope with all of these conditions. The novel approach to phonetic variability investigated here views the problem as one of phonetic information underspecification: some subset of information that the listener receives will be missing or uncertain. Lexical access is thus a phonetic code-breaking problem --- how can a system accumulate phonetic cues in each of these conditions to recognize words on the basis of incomplete evidence? The research program of this project takes a multidisciplinary approach to integrating linguistic theory with speech recognition technology; discriminative statistical models of linguistic features are employed to model nonlinear, overlapping phonological effects observed in speech. The framework allows derivation of new linguistic insights through analysis of trained systems. The educational program fosters interdisciplinary research (with cross-disciplinary graduate seminars) and increases participation of underrepresented students in Computer Science by introducing language technology topics early into the undergraduate curriculum and encouraging undergraduate research. Apart from cultivating a new way of thinking about pronunciation variation for ASR, the broader impacts of this research are to provide collaborative resources for the ASR and linguistics communities to discuss in tutorial and workshop settings. Addressing noise, accent, and speaking style in a consistent framework will also improve ASR technology for many who are underserved by current systems.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Deep Learning Based Complex Spectral Mapping for Multi-Channel Speaker Separation and Speech Enhancement
-
批准号:2125074
-
项目类别:Standard Grant
-
资助金额:$39.06万
-
财政年份:2021
-
负责人:Eric Fosler-Lussier
-
依托单位:
RI: Small: Early Elementary Reading Verification in Challenging Acoustic Environments
-
批准号:2008043
-
项目类别:Standard Grant
-
资助金额:$45.0万
-
财政年份:2020
-
负责人:Eric Fosler-Lussier
-
依托单位:
RI: Medium: Deep Neural Networks for Robust Speech Recognition through Integrated Acoustic Modeling and Separation
-
批准号:1409431
-
项目类别:Continuing Grant
-
资助金额:$79.81万
-
财政年份:2014
-
负责人:Eric Fosler-Lussier
-
依托单位:
CI-ADDO-NEW: Collaborative Research: The Speech Recognition Virtual Kitchen
-
批准号:1305319
-
项目类别:Standard Grant
-
资助金额:$38.21万
-
财政年份:2013
-
负责人:Eric Fosler-Lussier
-
依托单位:
CI-P:Collaborative Research:The Speech Recognition Virtual Kitchen
-
批准号:1205424
-
项目类别:Standard Grant
-
资助金额:$4.85万
-
财政年份:2012
-
负责人:Eric Fosler-Lussier
-
依托单位:
RI: Medium: Collaborative Research: Explicit Articulatory Models of Spoken Language, with Application to Automatic Speech Recognition
-
批准号:0905420
-
项目类别:Standard Grant
-
资助金额:$33.45万
-
财政年份:2009
-
负责人:Eric Fosler-Lussier
-
依托单位:
Workshop: Student Research in Computational Linguistics, at the HLT/NAACL 2004 Conference
-
批准号:0422841
-
项目类别:Standard Grant
-
资助金额:$2.02万
-
财政年份:2004
-
负责人:Eric Fosler-Lussier
-
依托单位:
海外基金