F. Brinkmann, H. Gamper, N. Raghuvanshi, and I. Tashev, “Towards encoding perceptually salient early reflections for parametric spatial audio rendering,” in Proc. AES Convention 148, May 2020, Paper 10380. [Online]. Available: https://aes.org/publications/elibrary-page/?id=20797
Brinkmann F, Gamper H, Raghuvanshi N, Tashev I. Towards encoding perceptually salient early reflections for parametric spatial audio rendering. In: AES Convention 148. Audio Engineering Society; 2020. Paper 10380. Available from: https://aes.org/publications/elibrary-page/?id=20797
@inproceedings{Brinkmann2020_20797,
author = {Brinkmann, Fabian and Gamper, Hannes and Raghuvanshi, Nikunj and Tashev, Ivan},
title = {{Towards encoding perceptually salient early reflections for parametric spatial audio rendering}},
booktitle = {AES Convention 148},
note = {Paper 10380},
year = {2020},
month = may,
publisher = {Audio Engineering Society},
url = {https://aes.org/publications/elibrary-page/?id=20797}
}
TY - CPAPER
TI - Towards encoding perceptually salient early reflections for parametric spatial audio rendering
AU - Brinkmann, Fabian
AU - Gamper, Hannes
AU - Raghuvanshi, Nikunj
AU - Tashev, Ivan
T2 - AES Convention 148
M1 - Paper 10380
PY - 2020
DA - 2020/05/06
UR - https://aes.org/publications/elibrary-page/?id=20797
PB - Audio Engineering Society
LA - en
AB - Parametric spatial audio rendering promises fast and perceptually convincing audio cues that remain playback-system agnostic and enable aesthetic modifications of the acoustic experience within games and virtual reality. We propose a parametric encoder for spatial room impulse responses that is tested with nine simulated rooms spanning a large range of sizes and reverberation times. A key component of the pipeline is a perceptually inspired model for determining a minimal set of salient early reflections to reduce computational complexity. The results of a listening study with 27 subjects suggest that rendering six early reflections is indiscernible from a fully-rendered reference for the tested speech content and frequency-independent room simulations based on the image source method. However, the proposed model requires further improvements with respect to detecting and selecting the most-salient early reflections.
ER -