Opens in a new tab

AES E-Library

← Back to search

Conference Paper

Spatial Audio Compression with Adaptive Singular Value Decomposition Using Reconstructed Frames

Authors: Namazi, Mahmoud; Elshafiy, Ahmed; Rose, Kenneth

AES Conference: AES 2022 International Audio for Virtual and Augmented Reality Conference · Paper 28 · August 2022

Abstract

MPEG-H 3D Audio is the current standard for the compression of higher-order ambisonics data. It uses singular value decomposition (SVD) to spatially decorrelate higher-order ambisonics data, followed by the modified discrete cosine transform to exploit temporal decorrelation. Prominent and ambient sound components are then separately encoded (e.g., using the standard core audio codec) and sent to the decoder. Significant improvements in bitrate and audio quality have been gained in earlier work over MPEG-H by applying the SVD operation in the frequency domain rather than the ambisonics domain. In this work, we provide additional compression gains by adaptively calculating and extending the set of SVD basis vectors, at negligible increase in side information cost, using information attained from the previously reconstructed frame. Objective and subjective results provide evidence for higher compression gains when compared to existing methods.

Details

Published in
AES Conference: AES 2022 International Audio for Virtual and Augmented Reality Conference
Paper number
28
Publication date
August 6, 2022
Session subject
Paper
Affiliation
University of California, Santa Barbara, CA, USA (See document for exact affiliation information.)
Type
Conference Paper