J. G. Beerends et al., “Perceptual Objective Listening Quality Assessment (POLQA), The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II—Perceptual Model,” J. Audio Eng. Soc., vol. 61, no. 6, pp. 385–402, Jun. 2013.
Beerends JG, Schmidmer C, Berger J, Obermann M, Ullmann R, Pomy J, Keyhl M. Perceptual Objective Listening Quality Assessment (POLQA), The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II—Perceptual Model. J Audio Eng Soc. 2013;61(6):385-402. Available from: https://aes.org/publications/elibrary-page/?id=16830
@article{Beerends2013_16830,
author = {Beerends, John G. and Schmidmer, Christian and Berger, Jens and Obermann, Matthias and Ullmann, Raphael and Pomy, Joachim and Keyhl, Michael},
title = {{Perceptual Objective Listening Quality Assessment (POLQA), The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II—Perceptual Model}},
journal = {Journal of the Audio Engineering Society},
volume = {61},
number = {6},
pages = {385--402},
year = {2013},
month = jun,
publisher = {Audio Engineering Society},
url = {https://aes.org/publications/elibrary-page/?id=16830}
}
TY - JOUR
TI - Perceptual Objective Listening Quality Assessment (POLQA), The Third Generation ITU-T Standard for End-to-End Speech Quality Measurement Part II—Perceptual Model
AU - Beerends, John G.
AU - Schmidmer, Christian
AU - Berger, Jens
AU - Obermann, Matthias
AU - Ullmann, Raphael
AU - Pomy, Joachim
AU - Keyhl, Michael
T2 - Journal of the Audio Engineering Society
J2 - J. Audio Eng. Soc.
VL - 61
IS - 6
SP - 385
EP - 402
PY - 2013
DA - 2013/06/06
UR - https://aes.org/publications/elibrary-page/?id=16830
PB - Audio Engineering Society
LA - en
AB - In this and the companion paper Part I, the authors present the Perceptual Objective Listening Quality Assessment (POLQA), the third-generation speech quality measurement algorithm, standardized by the International Telecommunication Union in 2011 as Recommendation P.863. This paper describes the newly developed perceptual model of this standard, allowing to assess speech quality over a wide range of distortions, from “High Definition” super-wideband speech (HD Voice, audio bandwidth up to 14 kHz) to extremely distorted narrowband telephony speech (audio bandwidth down to 2 kHz), using sample rates between 48 and 8 kHz. POLQA is suited for distortions that are outside the scope of PESQ, such as linear frequency response distortions, super-wideband degradations, time stretching/compression as found in Voice-over-IP, certain types of codec distortions, reverberations, and the impact of playback volume. Part II outlines the core elements of the underlying perceptual model and presents the final results.
ER -