Opens in a new tab

AES E-Library

← Back to search

Convention Paper

A Fractal Self-Similarity Model for the Spectral Representation of Audio Signals

Authors: Ferreira, Anibal J. S.; Sen, Deep; Sinha, Deepen

AES Convention 118 · Paper 6467 · May 2005

Abstract

In the application of conventional audio compression algorithms to low bit rate audio coding one is faced with the unsatisfactory tradeoff between coarser quantization and audio bandwidth reduction. BandwidthExtension has therefore emerged as an important tool for the satisfactory performance of low bit rate audio codecs. In this paper we describe one of a newer class of Frequency Extension techniques which are applied directly to the high frequency resolution representation of the signal (e.g., MDCT). This particular technique is based on a Fractal Self-Similarity Model (FSSM) for the short-term frequency representation of the signal and takes advantage of the high frequency resolution of the MDCT, namely in terms of parameter estimation.. The FSSM model, which may include multiple dilation and translation terms, has been found to be effective for a wide variety of speech and music signals and provides a compact description for long term correlation that may exist in frequency domain.. The Structure of the FSSM model is presented, issues related to parameter estimation, and its application to audio coding for bit rates of 8-48 kbps are discussed.

Details

Published in
AES Convention 118
AES Convention
118
Paper number
6467
Publication date
May 6, 2005
Session subject
Low Bit Rate Audio Coding (Research)
Affiliation
ATC Labs; University of New South Wales/ATC Labs; University of Porto/ATC Labs (See document for exact affiliation information.)
Type
Convention Paper