Audiovisual Description and Recognition of Dysarthric Speech
Audiovisual Description and Recognition of Dysarthric Speech
批准号:
7230076
负责人:
MARK ALLAN HASEGAWA-JOHNSON
金额:
$17.38万
依托单位国家:
美国
项目类别:
财政年份:
2006
资助国家:
美国
项目状态:
已结题
起止时间:
2006-04-01 至 2009-03-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
DESCRIPTION (provided by applicant): The primary aim of this proposal is to describe and automatically recognize the audible and visible correlates of phonological contrasts as produced by persons with spastic dysarthria. In order to meet the primary aim, this investigation will: - enroll a total of 16 subjects with dysarthria and 16 control subjects, - record each subject's production of phonetically balanced and pragmatically useful speech material using our AVICAR array of eight microphones and four video cameras, - measure the acoustic and visible correlates of consonant place of articulation, including formant locus, frication spectrum, lip aperture area, and jaw height, - develop automatic audio-only and audiovisual isolated word recognition algorithms, and - record each subject's participation in an objective comparison of audiovisual speech recognition, audio-only speech recognition, and typing as text input methods for human computer interface. LAY LANGUAGE SUMMARY: Subjects whose neuromotor deficit precludes or hinders their use of a keyboard may nevertheless retain some control over speech articulators. Our preliminary data demonstrate that subjects with 19-30% intelligibility (as rated by human listeners) may nevertheless achieve 90-100% recognition accuracy in an automatic isolated digit recognition task. We have found that the use of video in automatic speech recognition improves word recognition accuracy for talkers without dysarthria; we propose to extend our work to seek the same gains for talkers with dysarthria.
期刊论文(3)
专著(0)
科研奖励(0)
会议论文
LANDMARK-BASED SPEECH RECOGNITION: REPORT OF THE 2004 JOHNS HOPKINS SUMMER WORKSHOP.
基于地标的语音识别:2004 年约翰·霍普金斯大学夏季研讨会报告。
DOI:
10.1109/icassp.2005.1415088
发表时间:
2005
期刊:
Proceedings of the ... IEEE International Conference on Acoustics, Speech, and Signal Processing. ICASSP (Conference)
影响因子:
--
作者:
[Hasegawa-Johnson,Mark, Baker,James, Borys,Sarah, Chen,Ken, Coogan,Emily, Greenberg,Steven, Juneja,Amit, Kirchhoff,Katrin, Livescu,Karen, Mohan,Srividya, Muller,Jennifer, Sonmez,Kemal, Wang,Tianyu]
通讯作者:
Wang,Tianyu
Validation of a Virtual Still Face Procedure and Deep Learning Algorithms to Assess Infant Emotion Regulation and Infant-Caregiver Interactions in the Wild
-
批准号:10777825
-
项目类别:
-
资助金额:$62.28万
-
财政年份:2023
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
Audiovisual Description and Recognition of Dysarthric Speech
-
批准号:7075053
-
项目类别:
-
资助金额:$21.49万
-
财政年份:2006
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
FACTOR ANALYSIS OF MRI DERIVED ARTICULATOR SHAPES
-
批准号:2872124
-
项目类别:
-
资助金额:$2.22万
-
财政年份:1999
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
FACTOR ANALYSIS OF MRI DERIVED ARTICULATOR SHAPES
-
批准号:2522259
-
项目类别:
-
资助金额:$2.62万
-
财政年份:1998
-
负责人:MARK ALLAN HASEGAWA-JOHNSON
-
依托单位:
国内基金
海外基金
Research on Quantum Field Theory without a Lagrangian Description
-
批准号:24ZR1403900
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:SATOSHI NAWATA
-
依托单位: