An audio-visual corpus for speech perception and automatic speech recognition (L)
An audio-visual corpus for speech perception and automatic speech recognition (L)
复制标题
DOI:
10.1121/1.2229005
复制
发表时间:
2006-11-01
影响因子:
2.4
通讯作者:
Shao, Xu
中科院分区:
文献类型:
--
作者:
Cooke, Martin;Barker, Jon;Shao, Xu
An audio-visual corpus has been collected to support the use of common material in speech perception and automatic speech recognition studies. The corpus consists of high-quality audio and video recordings of 1000 sentences spoken by each of 34 talkers. Sentences are simple, syntactically identical phrases such as "place green at B 4 now." Intelligibility tests using the audio signals suggest that the-material is easily identifiable in quiet and low levels of stationary noise. The annotated corpus is available on the web for research use. (c) 2006 Acoustical Society of America.