Constructing Japanese test collections for spoken term detection

Constructing Japanese test collections for spoken term detection
复制标题

DOI:
10.21437/interspeech.2010-258
复制
发表时间:
2010
期刊:
--
影响因子:
--
通讯作者:
Y. Itoh;H. Nishizaki;Xinhui Hu;H. Nanjo;T. Akiba;Tatsuya Kawahara;S. Nakagawa;T. Matsui;Y. Yamashita
Y. Itoh;H. Nishizaki;Xinhui Hu;H. Nanjo;T. Akiba;Tatsuya Kawahara;S. Nakagawa;T. Matsui;Y. Yamashita
中科院分区:
其他
文献类型:
--
作者:
Y. Itoh;H. Nishizaki;Xinhui Hu;H. Nanjo;T. Akiba;Tatsuya Kawahara;S. Nakagawa;T. Matsui;Y. Yamashita

文献摘要

相似文献

口语文档检索(SDR)和口语词检测(STD)是当前口语文档处理研究中最热门的两个课题,由文本检索会议(TREC)和美国国家标准技术研究院(NIST)建立的SDR和STD测试集是这两个领域的研究热点。由于日本的口语文档处理研究人员也需要这样的测试集合SDR和STD,我们已经建立了一个工作组,以开发这些集合在特殊兴趣小组-口语语言处理(SIG-SLP)的信息处理学会的日本。该工作组已经构建并提供了一个SDR测试集合,现在正在构建新的STD测试集合,将向研究人员开放。本文介绍了新的测试集的政策,大纲和时间表。然后,将新的测试集合与NIST STD测试集合进行比较。索引词:口语词检测,测试收集
Spoken Document Retrieval (SDR) and Spoken Term Detection (STD) have been two of the most intensively investigated topics in spoken document processing research according to the establishment of the SDR and STD test collections by the Text REtrieval Conference (TREC) and NIST. Because Japanese spoken document processing researchers also requires such test collections for SDR and STD, we have established a working group to develop these collections in Special Interest Group -Spoken Language Processing (SIG-SLP) of the Information Processing Society of Japan. The working group has constructed and made available a test collection for SDR, and is now constructing new test collections for STD that will be open to researchers. The present paper introduces the policies, outline, and schedule of the new test collections. Then, the new test collections are compared with the NIST STD test collections. Index Terms: spoken term detection, test collection