Voice search of structured media data

Voice search of structured media data
复制标题

结构化媒体数据的语音搜索

DOI:
10.1109/icassp.2009.4960490
复制
发表时间:
2009
期刊:
2009 IEEE International Conference on Acoustics, Speech and Signal Processing
影响因子:
--
通讯作者:
A. Acero
A. Acero
中科院分区:
--
文献类型:
--
作者:
Young;Ye;Y. Ju;M. Seltzer;I. Tashev;A. Acero

文献摘要

被引文献

相似文献

本文解决了语音搜索应用中使用非结构化查询来搜索结构化数据库的问题。通过将结构化信息整合到音乐元数据中,文本查询的端到端搜索错误率降低了15%,语音查询的错误率降低了11%。在此基础上,与基线系统相比,HMM顺序评分模型将文本查询的错误率降低了28%,将语音查询的错误率降低了23%。此外,引入语音相似度模型对语音识别误差进行补偿,使得端到端搜索精度在不同的语音识别精度水平上得到了一致的提高。
This paper addresses the problem of using unstructured queries to search a structured database in voice search applications. By incorporating structural information in music metadata, the end-to-end search error has been reduced by 15% on text queries and up to 11% on spoken queries. Based on that, an HMM sequential rescoring model has reduced the error rate by 28% on text queries and up to 23% on spoken queries compared to the baseline system. Furthermore, a phonetic similarity model has been introduced to compensate speech recognition errors, which has improved the end-to-end search accuracy consistently across different levels of speech recognition accuracy.