Opens in a new tab

AES E-Library

← Back to search

Conference Paper

Optimising HRTFs to Improve Spatial Release from Masking

Authors: Nils Marggraf-Turley; Niels Pontoppidan; Lorenzo Picinali

AVARIG 2026: Audio for Virtual and Augmented Reality and Immersive Games · Paper 464 · June 2026

Abstract

Binaural hearing supports speech understanding by allowing listeners to perceptually segregate spatially separated sources, a phenomenon termed spatial release from masking (SRM). The spatial cues that facilitate SRM are determined by the listeners morphology and as such are described by the head-related transfer function (HRTF). Although individual HRTFs are generally considered important for accurate localisation, previous work suggests they might not necessarily maximise performance across all aspects of spatial perception, including SRM. This
motivates the concept of application-specific HRTFs. Here, we propose an HRTF augmentation method to increase model-predicted SRM in frontback cocktail-party configurations where SRM is limited. HRTF log-magnitude spectra are parameterised using principal component analysis (PCA), and PCA coefficients are optimised using a differentiable auditory-model-based objective to increase better-ear SNR while constraining broadband interaural distortions. The proposed method produces model-predicted SRM improvements of 49 dB for frontback targetmasker configurations while maintaining broadband interaural level difference (ILD) deviations below 1 dB.

Details

Published in
AVARIG 2026: Audio for Virtual and Augmented Reality and Immersive Games
Paper number
464
Publication date
June 30, 2026
Session subject
Binaural audio rendering / reproduction, Perception and subjective evaluation of spatial / immersive audio
Type
Conference Paper