Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Perceptual Audio Modeling Based on Total Least Squares Algorithms

Authors: Hermus, Kris; Verhelst, Werner; Wambacq, Patrick

AES Convention 112 · Paper 5571 · April 2002

Abstract

Total Least Squares (TLS) algorithms automatically decompose (audio) frames into a number of exponentially damped sinusoids. This can provide for more efficient modeling than plain sinusoidal modeling, especially in the case of transitional frames. Straightforward implementations of TLS optimize a SNR criterion. In our implementation we apply TLS in a subband scheme in which the number of damped sinusoids is both frame and subband dependent. This is made possible through the use of perceptual information provided by the MPEG-I psycho-acoustic model I. Experiments on different audio tracks provide proof of concept for our perceptual ESM, and illustrate the significant reduction in modeling components compared to a non-perceptual ESM.

Details

Published in
AES Convention 112
AES Convention
112
Paper number
5571
Publication date
April 6, 2002
Session subject
Low Bit-Rate Audio Coding
Affiliation
Katholieke Universiteit Leuven, dept. ESAT - div. PSI, Leuven, BELGIUM ; Vrije Universiteit Brussel, dept. ETRO - div. DSSP, Brussels, BELGIUM (See document for exact affiliation information.)
Type
Convention Paper