Segment quantization for very-low-rate speech coding

Segment quantization for very-low-rate speech coding
复制标题

极低速率语音编码的分段量化

DOI:
10.1109/icassp.1982.1171472
复制
发表时间:
1982
期刊:
IEEE International Conference on Acoustics, Speech, and Signal Processing
影响因子:
--
通讯作者:
J. Makhoul
J. Makhoul
中科院分区:
--
文献类型:
--
作者:
Salim Roukos;R. Schwartz;J. Makhoul

文献摘要

被引文献

相似文献

提出了一种新的超低码率语音编码方法,将输入语音作为可变长度片段序列。一段是由一组帧组成的频谱,其中每一帧由频谱、基音和增益表示。我们使用一种自动分割算法来获得平均持续时间与音素相当的片段。段被量化为单个块。用于量化的距离测量结合了两段的适当时间对齐。我们采用一种计算效率高的度量,不使用通常的动态规划时间翘曲。使用上述分组量化方法的两个基本声码器已被用于以200b /s的速度传输可理解的语音。
We introduce a new method for very-low-rate vocoding that the input speech as a sequence of variable-length segments. A segment is a by a spectrum of frames, where each frame is represented by a spectrum, pitch and gain. We use an automatic segmentation algorithm to obtain segments with an average duration comparable to that of a phoneme. A segment is quantized as a single block. The distance measure used for quantization incooporates the appropriate time alignment of two segments. We employ a computationally efficient metric that does not use the usual dynamic programming time warping. Two basic vocoders using the above approach of block quantization have been used to transmit intelligible speech at 200 b/s.