Subword unit representations for spoken document retrieval
Subword unit representations for spoken document retrieval
复制标题
用于语音文档检索的子词单元表示
DOI:
10.21437/eurospeech.1997-460
复制
发表时间:
1997
期刊:
影响因子:
--
通讯作者:
V. Zue
中科院分区:
文献类型:
--
作者:
Kenney Ng;V. Zue
This paper investigates the feasibility of using subword unit representations for spoken document retrieval as an alternative to using words generated by either keyword spotting or word recognition. Our investigation is motivated by the observation that word-based retrieval approaches face the problem of either having to know the keywords to search for a priori, or requiring a very large recognition vocabulary in order to cover the contents of growing and diverse message collections. In this study, we examine a range of subword units of varying complexity derived from phonetic transcriptions. The basic underlying unit is the phone; more and less complex units are derived by varying the level of detail and the length of sequences of the phonetic units. We measure the ability of the di(cid:11)erent subword units to e(cid:11)ectively index and retrieve a large collection of recorded speech messages. We also compare their performance when the underlying phonetic transcriptions are perfect and when they contain phonetic recognition errors.