A novel multichannel audio signal compression method based on tensor representation and decomposition

A novel multichannel audio signal compression method based on tensor representation and decomposition
复制标题

DOI:
10.1109/cc.2014.6825261
复制
发表时间:
2014-06
影响因子:
4.1
通讯作者:
Wang Jing;Xie Xiang;Kuang Jingming
Wang Jing;Xie Xiang;Kuang Jingming
中科院分区:
计算机科学3区
文献类型:
--
作者:
Wang Jing;Xie Xiang;Kuang Jingming

文献摘要

相似文献

多声道音频信号比单声道和立体声音频信号更难压缩。提出了一种基于张量表示和张量分解的多通道音频信号压缩方法。该方法将多声道音频信号用三阶张量空间表示,并将其分解为具有声道、时间和频率三个因子矩阵的核心张量。仅传输截断的核心张量,其将乘以预训练的因子矩阵以重建原始张量空间。客观和主观的实验已经做了一个非常明显的压缩能力与可接受的输出质量。所提出的压缩方法的新奇在于,它能够实现高压缩能力和向后兼容性,并且对听觉具有有限的信号失真。
Multichannel audio signal is more difficult to be compressed than mono and stereo ones. A novel multichannel audio signal compression method based on tensor representation and decomposition is proposed in this paper. The multichannel audio is represented with 3-order tensor space and is decomposed into core tensor with three factor matrices in the way of channel, time and frequency. Only the truncated core tensor is transmitted which will be multiplied by the pre-trained factor matrices to reconstruct the original tensor space. Objective and subjective experiments have been done to show a very noticeable compression capability with an acceptable output quality. The novelty of the proposed compression method is that it enables both high compression capability and backward compatibility with limited signal distortion to the hearing.