Audiovisual Description and Recognition of Dysarthric Speech
Audiovisual Description and Recognition of Dysarthric Speech
批准号:
7230076
负责人:
MARK ALLAN HASEGAWA-JOHNSON
金额:
$17.38万
依托单位国家:
美国
项目类别:
财政年份:
2006
资助国家:
美国
项目状态:
已结题
起止时间:
2006-04-01 至 2009-03-31
中文摘要
描述(由申请人提供):本提案的主要目的是描述和自动识别由痉挛性构音障碍患者产生的语音对比的可听和可见相关物。为了实现主要目标,本研究将:- 招募总共16名患有构音障碍的受试者和16名对照受试者,-使用我们的八个麦克风和四个摄像机的AVICAR阵列记录每个受试者的语音平衡和实用有用的语音材料的产生,-测量发音的辅音位置的声学和可见相关性,包括共振峰轨迹、摩擦频谱,唇孔面积和下颌高度,-开发自动的仅音频和视听孤立词识别算法,以及-记录每个受试者参与视听语音识别、仅音频语音识别和打字作为人机接口的文本输入方法的客观比较。非专业语言摘要:受试者的神经运动缺陷排除或阻碍他们使用键盘,但可能保留一些控制语音发音。我们的初步数据表明,受试者与19-30%的可懂度(由人类听众评定),但可以达到90-100%的识别准确率在自动孤立的数字识别任务。我们已经发现,在自动语音识别中使用视频提高了没有构音障碍的说话者的单词识别准确性;我们建议扩展我们的工作,为构音障碍的说话者寻求同样的收益。
英文摘要
DESCRIPTION (provided by applicant): The primary aim of this proposal is to describe and automatically recognize the audible and visible correlates of phonological contrasts as produced by persons with spastic dysarthria. In order to meet the primary aim, this investigation will: - enroll a total of 16 subjects with dysarthria and 16 control subjects, - record each subject's production of phonetically balanced and pragmatically useful speech material using our AVICAR array of eight microphones and four video cameras, - measure the acoustic and visible correlates of consonant place of articulation, including formant locus, frication spectrum, lip aperture area, and jaw height, - develop automatic audio-only and audiovisual isolated word recognition algorithms, and - record each subject's participation in an objective comparison of audiovisual speech recognition, audio-only speech recognition, and typing as text input methods for human computer interface. LAY LANGUAGE SUMMARY: Subjects whose neuromotor deficit precludes or hinders their use of a keyboard may nevertheless retain some control over speech articulators. Our preliminary data demonstrate that subjects with 19-30% intelligibility (as rated by human listeners) may nevertheless achieve 90-100% recognition accuracy in an automatic isolated digit recognition task. We have found that the use of video in automatic speech recognition improves word recognition accuracy for talkers without dysarthria; we propose to extend our work to seek the same gains for talkers with dysarthria.
期刊论文(3)
专著(0)
科研奖励(0)
会议论文
LANDMARK-BASED SPEECH RECOGNITION: REPORT OF THE 2004 JOHNS HOPKINS SUMMER WORKSHOP.
基于地标的语音识别:2004 年约翰·霍普金斯大学夏季研讨会报告。
DOI:
10.1109/icassp.2005.1415088
发表时间:
2005
期刊:
Proceedings of the ... IEEE International Conference on Acoustics, Speech, and Signal Processing. ICASSP (Conference)
影响因子:
--
作者:
[Hasegawa-Johnson,Mark, Baker,James, Borys,Sarah, Chen,Ken, Coogan,Emily, Greenberg,Steven, Juneja,Amit, Kirchhoff,Katrin, Livescu,Karen, Mohan,Srividya, Muller,Jennifer, Sonmez,Kemal, Wang,Tianyu]
通讯作者:
Wang,Tianyu
Validation of a Virtual Still Face Procedure and Deep Learning Algorithms to Assess Infant Emotion Regulation and Infant-Caregiver Interactions in the Wild
-
批准号:10777825
-
项目类别:
-
资助金额:$62.28万
-
财政年份:2023
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
Audiovisual Description and Recognition of Dysarthric Speech
-
批准号:7075053
-
项目类别:
-
资助金额:$21.49万
-
财政年份:2006
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
FACTOR ANALYSIS OF MRI DERIVED ARTICULATOR SHAPES
-
批准号:2872124
-
项目类别:
-
资助金额:$2.22万
-
财政年份:1999
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
FACTOR ANALYSIS OF MRI DERIVED ARTICULATOR SHAPES
-
批准号:2522259
-
项目类别:
-
资助金额:$2.62万
-
财政年份:1998
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
国内基金
海外基金
Research on Quantum Field Theory without a Lagrangian Description
-
批准号:24ZR1403900
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:SATOSHI NAWATA
-
依托单位: