Longitudinal study of ASR performance on ageing voices

Longitudinal study of ASR performance on ageing voices
复制标题

ASR 对老化声音性能的纵向研究

DOI:
--
复制
发表时间:
2008
期刊:
Interspeech
影响因子:
--
通讯作者:
Joe Frankel
Joe Frankel
中科院分区:
--
文献类型:
--
作者:
Ravichander Vipperla;S. Renals;Joe Frankel

文献摘要

被引文献

相似文献

本文介绍了一个纵向研究的结果ASR性能老化的声音。实验是在美国最高法院(SCOTUS)的诉讼录音上进行的。结果表明,老年人声音的自动语音识别(ASR)单词错误率(WERs)显著高于成人声音。随着年龄的增长,老年人的词汇错误率逐渐增加。使用最大似然线性回归(MLLR)为基础的扬声器适应老化的声音提高了WER,但性能仍然是相当低的成人的声音相比。然而,扬声器适应减少WER随着年龄的增长在老年。索引词:老龄化的声音,纵向研究,SCOTUS语料库,MLLR
This paper presents the results of a longitudinal study of ASR performance on ageing voices. Experiments were conducted on the audio recordings of the proceedings of the Supreme Court Of The United States (SCOTUS). Results show that the Automatic Speech Recognition (ASR) Word Error Rates (WERs) for elderly voices are significantly higher than those of adult voices. The word error rate increases gradually as the age of the elderly speakers increase. Use of maximum likelihood linear regression (MLLR) based speaker adaptation on ageing voices improves the WER though the performance is still considerably lower compared to adult voices. Speaker adaptation however reduces the increase in WER with age during old age. IndexTerms: Ageing Voices, longitudinal study, SCOTUScorpus, MLLR