Perceptual Objective Listening Quality Assessment ( POLQA ) , The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II – Perceptual Model

Perceptual Objective Listening Quality Assessment ( POLQA ) , The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II – Perceptual Model
复制标题

感知客观听力质量评估 (POLQA),第三代 ITU-T 端到端语音质量测量标准第二部分 – 感知模型

DOI:
--
复制
发表时间:
2013
期刊:
影响因子:
--
通讯作者:
M. Keyhl
M. Keyhl
中科院分区:
--
文献类型:
--
作者:
J. Beerends;Christian Schmidmer;J. Berger;Matthias Obermann;Raphael Ullmann;Joachim Pomy;M. Keyhl

文献摘要

被引文献

相似文献

在两篇密切相关的论文中,我们介绍了POLQA(感知客观听力质量评估),这是第三代感知客观语音质量测量算法,由国际电信联盟(ITU-T)于2011年作为建议P.863标准化。该测量算法模拟使用五点意见量表在听力测试中对语音片段的质量进行评级的受试者。与第二代客观语音质量测量方法PESQ(语音质量感知评估)相比,新标准在预测主观语音质量方面提供了显著改进的性能。新的POLQA算法允许在广泛的失真范围内预测语音质量,从“高清”超宽带语音(HD语音,音频带宽高达14 kHz)到极度失真的窄带电话语音(音频带宽低至2 kHz),使用48至8 kHz之间的采样率。POLQA适用于PESQ范围之外的失真,例如线性频率响应失真、IP语音中的时间拉伸/压缩、某些类型的编解码器失真、混响和播放音量的影响。POLQA在评估任何类型的降级方面都优于PESQ,使其成为当今和未来移动的和基于IP的网络中所有语音质量测量的理想工具。本文(第二部分)概述了底层感知模型的核心要素,并给出了最终结果。
In two closely related papers we present POLQA (Perceptual Objective Listening Quality Assessment), the third generation perceptual objective speech quality measurement algorithm, standardized by the International Telecommunication Union (ITU-T) as Recommendation P.863 in 2011. This measurement algorithm simulates subjects that rate the quality of a speech fragment in a listening test using a five-point opinion scale. The new standard provides a significantly improved performance in predicting the subjective speech quality in terms of Mean Opinion Scores when compared to PESQ (Perceptual Evaluation of Speech Quality), the second generation of objective speech quality measurements. The new POLQA algorithm allows for predicting speech quality over a wide range of distortions, from “High Definition” super-wideband speech (HD Voice, audio bandwidth up to 14 kHz) to extremely distorted narrowband telephony speech (audio bandwidth down to 2 kHz), using sample rates between 48 and 8 kHz. POLQA is suited for distortions that are outside the scope of PESQ such as linear frequency response distortions, time stretching/compression as found in Voice-over-IP, certain types of codec distortions, reverberations, and the impact of playback volume. POLQA outperforms PESQ in assessing any kind of degradation making it an ideal tool for all speech quality measurements in today’s and future mobile and IP based networks. This paper (Part II) outlines the core elements of the underlying perceptual model and presents the final results.