An audio-visual corpus for speech perception and automatic speech recognition (L)

An audio-visual corpus for speech perception and automatic speech recognition (L)
复制标题

DOI:
10.1121/1.2229005
复制
发表时间:
2006-11-01
影响因子:
2.4
通讯作者:
Shao, Xu
Shao, Xu
中科院分区:
物理与天体物理3区
文献类型:
--
作者:
Cooke, Martin;Barker, Jon;Shao, Xu

文献摘要

被引文献

相似文献

一个视听语料库已收集到支持使用的共同材料在语音感知和自动语音识别研究。该语料库由34个谈话者中的每个人所说的1000个句子的高质量音频和视频记录组成。句子都是简单的,句法上相同的短语,如“现在把绿色放在B 4。“使用音频信号进行的清晰度测试表明,这种材料在安静和低水平的静态噪音中很容易识别。注释的语料库可在网上供研究使用。(c)2006年,美国声学学会。
An audio-visual corpus has been collected to support the use of common material in speech perception and automatic speech recognition studies. The corpus consists of high-quality audio and video recordings of 1000 sentences spoken by each of 34 talkers. Sentences are simple, syntactically identical phrases such as "place green at B 4 now." Intelligibility tests using the audio signals suggest that the-material is easily identifiable in quiet and low levels of stationary noise. The annotated corpus is available on the web for research use. (c) 2006 Acoustical Society of America.