EAGER: Collaborative Research: Towards Modeling Human Speech Confusions in Noise
EAGER: Collaborative Research: Towards Modeling Human Speech Confusions in Noise
批准号:
1247809
负责人:
Abeer Alwan
金额:
$10.0万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2012
资助国家:
美国
项目状态:
已结题
起止时间:
2012-08-01 至 2015-07-31
中文摘要
这项早期概念探索性研究补助金(AGER)支持一项探索性研究,以评估在噪声存在的情况下预测人类语音识别的模型组件。这样的模型有可能预测在不同水平的背景噪音和不同的语速下细微的语音区别之间的混淆。这项研究利用了现代生理学结果,这些结果表明,初级听觉皮质执行谱-时间滤波;也就是说,在每个听觉频率上,有对特定谱-时间调制敏感的细胞。在这个项目中,对于一个以两种不同语速记录的CVC音节数据库,在存在平稳和非平稳加性噪声以及不同信噪比的情况下进行的知觉实验产生了混淆统计数据。然后,将这些统计数据与由结合了这些谱-时间过滤器的元素增强的听觉模型产生的统计数据进行比较。这项研究的成功结果将建议对目前的听力模型进行增强,并最终在本研究机构作为试点的更广泛的研究之后,促进对人类言语感知的理解。背景噪声对包括助听器和自动语音识别(ASR)系统在内的各种语音和听力设备来说是一个具有挑战性的问题。由于听力正常的人类听者非常擅长在噪音中感知语音,这种对人类模型的理解的改进可能会导致更好的人工语音处理系统。为这项研究开发的数据库和工具将分发给研究界。
英文摘要
This EArly-concept Grant for Exploratory Research (EAGER) supports an exploratory study to evaluate model components for prediction of human speech recognition in the presence of noise. Such a model has the potential to predict confusions between fine phonetic distinctions in different levels of background noise and at different speaking rates. The study takes advantage of modern physiological results that indicate that the primary auditory cortex performs spectro-temporal filtering; that is, that there are cells that are sensitive to particular spectro-temporal modulations at each auditory frequency. In this project, perceptual experiments in the presence of both stationary and non-stationary additive noise and at different signal-to-noise ratios for a database of CVC syllables recorded at 2 different speaking rates yield confusion statistics. These statistics are then compared to those resulting from an auditory model enhanced by elements incorporating these spectro-temporal filters. Successful results from this study will suggest enhancements to current hearing models and ultimately, after a broader study for which this EAGER is a pilot, advance the understanding of human speech perception. Background noise presents a challenging problem for a variety of speech and hearing devices including hearing aids and automatic speech recognition (ASR) systems. Since normal-hearing human listeners are extremely adept at perceiving speech in noise, this improved understanding of human models could lead to better artificial systems for speech processing. The databases and tools developed for this study will be disseminated to the research community.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Collaborative Research: Improving speech technology for better learning outcomes: the case of AAE child speakers
-
批准号:2202585
-
项目类别:Standard Grant
-
资助金额:$31.89万
-
财政年份:2022
-
负责人:Abeer Alwan
-
依托单位:
Collaborative Research: RI: Small: From Ultrasound and MRI to articulatory and acoustic models of child speech development
-
批准号:2006979
-
项目类别:Standard Grant
-
资助金额:$23.0万
-
财政年份:2020
-
负责人:Abeer Alwan
-
依托单位:
Workshop for Undergraduate and MS Female Students in Speech Science and Technology
-
批准号:1745166
-
项目类别:Standard Grant
-
资助金额:$2.5万
-
财政年份:2017
-
负责人:Abeer Alwan
-
依托单位:
NRI: INT: COLLAB: Development, Deployment and Evaluation of Personalized Learning Companion Robots for Early Literacy and Language Learning
-
批准号:1734380
-
项目类别:Standard Grant
-
资助金额:$61.56万
-
财政年份:2017
-
负责人:Abeer Alwan
-
依托单位:
RI: Medium: Collaborative Research: Variance and Invariance in Voice Quality: Implications for Machine and Human Speaker Identification
-
批准号:1704167
-
项目类别:Continuing Grant
-
资助金额:$85.16万
-
财政年份:2017
-
负责人:Abeer Alwan
-
依托单位:
A Workshop for Junior Female Researchers in Speech Science and Technology
-
批准号:1637240
-
项目类别:Standard Grant
-
资助金额:$3.0万
-
财政年份:2016
-
负责人:Abeer Alwan
-
依托单位:
The Role of Speech Science in Developing Robust Speech Technology Applications
-
批准号:1543522
-
项目类别:Standard Grant
-
资助金额:$3.5万
-
财政年份:2015
-
负责人:Abeer Alwan
-
依托单位:
EAGER: Collaborative Research: Models of Child Speech
-
批准号:1551113
-
项目类别:Standard Grant
-
资助金额:$14.0万
-
财政年份:2015
-
负责人:Abeer Alwan
-
依托单位:
EAGER: Variance and Invariance in Voice Quality
-
批准号:1450992
-
项目类别:Standard Grant
-
资助金额:$20.0万
-
财政年份:2014
-
负责人:Abeer Alwan
-
依托单位:
RI: Small: A New Voice Source Model: From Glottal Areas to Better Speech Synthesis
-
批准号:1018863
-
项目类别:Continuing Grant
-
资助金额:$45.0万
-
财政年份:2010
-
负责人:Abeer Alwan
-
依托单位:
RI: Medium: Collaborative Research: The Effect of Subglottal Resonances on Machine and Human Speaker Normalization
-
批准号:0905381
-
项目类别:Standard Grant
-
资助金额:$63.97万
-
财政年份:2009
-
负责人:Abeer Alwan
-
依托单位:
Collaborative Research: IDBR: VoxNet--A deployable bioacoustic sensor network
-
批准号:0936454
-
项目类别:Continuing Grant
-
资助金额:$4.25万
-
财政年份:2008
-
负责人:Abeer Alwan
-
依托单位:
Collaborative Research: IDBR: VoxNet--A deployable bioacoustic sensor network
-
批准号:0754120
-
项目类别:Continuing Grant
-
资助金额:$0.0万
-
财政年份:2008
-
负责人:Abeer Alwan
-
依托单位:
Collaborative Research: Landmark-based Robust Speech Recognition using Prosody-guided Models of Speech Variability
-
批准号:0703805
-
项目类别:Continuing Grant
-
资助金额:$0.0万
-
财政年份:2007
-
负责人:Abeer Alwan
-
依托单位:
ITR-Collaborative Research: Development and Evaluation of a Hybrid Concatenative/Rule-Based Visual Speech Synthesis System
-
批准号:0312810
-
项目类别:Standard Grant
-
资助金额:$18.32万
-
财政年份:2003
-
负责人:Abeer Alwan
-
依托单位:
IERI Collaborative Research: Automating Early Assesment of Academic Standards for Very Young Native and Non-Native Speakers of American English
-
批准号:0326214
-
项目类别:Standard Grant
-
资助金额:$0.0万
-
财政年份:2003
-
负责人:Abeer Alwan
-
依托单位:
CAREER: From Imaging and Acoustic Data to Articulatory Synthesis
-
批准号:9503089
-
项目类别:Continuing Grant
-
资助金额:$13.94万
-
财政年份:1995
-
负责人:Abeer Alwan
-
依托单位:
RIA: A Model of Speech Perception in Noise
-
批准号:9309418
-
项目类别:Standard Grant
-
资助金额:$10.0万
-
财政年份:1993
-
负责人:Abeer Alwan
-
依托单位:
海外基金