RI: Medium: Collaborative Research: Explicit Articulatory Models of Spoken Language, with Application to Automatic Speech Recognition
RI: Medium: Collaborative Research: Explicit Articulatory Models of Spoken Language, with Application to Automatic Speech Recognition
批准号:
0905633
负责人:
Karen Livescu
金额:
$43.88万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2009
资助国家:
美国
项目状态:
已结题
起止时间:
2009-07-01 至 2013-06-30
中文摘要
点击翻译按钮获取中文摘要
英文摘要
This award is funded under the American Recovery and Reinvestment Act of 2009 (Public Law 111-5).One of the main challenges in automatic speech recognition is variability in speaking style, including speaking rate changes and coarticulation. Models of the articulators (such as the lips and tongue) can succinctly represent much of this variability. Most previous work on articulatory models has focused on the relationship between acoustics and articulation, but more significant improvements require models of the hidden articulatory state structure. This work has both a technological goal of improving recognition and a scientific goal of better understanding articulatory phenomena.The project considers larger model classes than previously studied. In particular, the project develops graphical models, including dynamic Bayesian networks and conditional random fields, designed to take advantage of articulatory knowledge. A new framework for hybrid directed and undirected graphical models is being developed, in recognition of the benefits of both directed and undirected models, and of both generative and discriminative training. The project activities include major extension of earlier articulatory models with context modeling, asynchrony structures, and specialized training; development of factored conditional random field models of articulatory variables; and discriminative training to alleviate word confusability.The scientific goal addresses questions about the ways in which articulatory trajectories vary in different contexts. Existing databases are used, and initial work in manual articulatory annotation is being extended. In addition, the project uses articulatory models to perform forced transcription of larger data sets, providing an additional resource for the research community. Other broad impacts include new models and techniques with applicability to other time-series modeling problems. Extending the applicability of speech recognition will help it fulfill its promise of enabling more efficient storage of and access to spoken information, and equalizing the technological playing field for those with hearing or motor disabilities.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
RI: Small: From acoustics to semantics: Embedding speech for a hierarchy of tasks
-
批准号:1816627
-
项目类别:Continuing Grant
-
资助金额:$45.0万
-
财政年份:2018
-
负责人:Karen Livescu
-
依托单位:
EAGER: Discovery of Segmental Sub-Word Structure in Speech
-
批准号:1433485
-
项目类别:Standard Grant
-
资助金额:$9.99万
-
财政年份:2014
-
负责人:Karen Livescu
-
依托单位:
RI: Medium: Collaborative Research: Models of Handshape Articulatory Phonology for Recognition and Analysis of American Sign Language
-
批准号:1409837
-
项目类别:Standard Grant
-
资助金额:$85.41万
-
财政年份:2014
-
负责人:Karen Livescu
-
依托单位:
RI: Small: Multi-View Learning of Acoustic Features for Speech Recognition Using Articulatory Measurements
-
批准号:1321015
-
项目类别:Continuing Grant
-
资助金额:$44.49万
-
财政年份:2013
-
负责人:Karen Livescu
-
依托单位:
海外基金