Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Individualized HRTF-based Binaural Renderer for Higher-Order Ambisonics

Authors: Zhang, Mengfan; Guan, Tianyi; Chen, Lianwu; Fu, Tianxiao; Su, Dan; Qu, Tianshu

AES Convention 150 · Paper 10454 · May 2021

Abstract

Ambisonics is a promising spatial sound technique in augmented and virtual reality. In our previous study, we modeled the individual head-related transfer functions (HRTFs) using deep neural networks based on spatial principal component analysis. This paper proposes an individualized HRTF-based binaural renderer for the higher-order Ambisonics. The binaural renderer is implemented by filtering the virtual loudspeaker signals using individualized HRTFs. We perform subjective experiments to evaluate generic and individualized binaural renderers. Results show that the individualized binaural renderer has front-back confusion rates that are significantly lower than those of the generic binaural renderer. Therefore, we validate that using individualized HRTFs to convolve with those virtual loudspeaker signals to generate virtual sound at an arbitrary spatial direction still performs better than those using generic HRTFs. In addition, by measuring or modeling individual’s HRTFs in a small set of directions, our proposed binaural renderer system effectively predict individual’s HRTFs in arbitrary spatial directions.

Details

Published in
AES Convention 150
AES Convention
150
Paper number
10454
Publication date
May 6, 2021
Session subject
HRTF
Affiliation
Key Laboratory on Machine Perception (Ministry of Education), Speech and Hearing Research Center, Peking University, China; Tencent AI Lab, Shenzhen, China (See document for exact affiliation information.)
Type
Convention Paper