Robust speech recognition over IP networks

Robust speech recognition over IP networks
复制标题

IP 网络上的强大语音识别

DOI:
10.1109/icassp.2000.862101
复制
发表时间:
2000
期刊:
2000 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No.00CH37100)
影响因子:
--
通讯作者:
S. Semnani
S. Semnani
中科院分区:
--
文献类型:
--
作者:
B. Milner;S. Semnani

文献摘要

被引文献

相似文献

这项工作着眼于在基于数据包的网络(如IP网络)上执行鲁棒语音识别所涉及的问题。这涉及到将鲁棒语音识别与通过IP网络发送语音数据的可靠方法相结合。考虑了语音在网络上发送的格式,结果表明直接传输前端特征比用编解码器编码语音具有更好的鲁棒性。为了解决丢包问题,提出了一种新的丢失帧检测和估计方案。这样可以将50%的数据包丢失从33%恢复到90%,仅比无丢失情况低3%。
This work looks at the issues involved in performing robust speech recognition over a packet-based network such as the IP network. This involves the combination of robust speech recognition together with a reliable method of sending speech data over the IP network. The format in which the speech is sent over the network is considered and results show that much better robustness is achieved when the front-end features are transmitted directly rather than encoding the speech with a codec. The problem of packet loss is addressed and a novel detection and estimation scheme for missing frames is introduced to overcome this problem. This is shown to recover performance with 50% packet loss from 33% to 90% which is only 3% below the no loss case.