E. Gallo and N. Tsingos, “Extracting and Re-Rendering Structured Auditory Scenes from Field Recordings,” in Proc. AES Conference: 30th International Conference: Intelligent Audio Environments, Mar. 2007, Paper 19. [Online]. Available: https://aes.org/publications/elibrary-page/?id=13917
Gallo E, Tsingos N. Extracting and Re-Rendering Structured Auditory Scenes from Field Recordings. In: AES Conference: 30th International Conference: Intelligent Audio Environments. Audio Engineering Society; 2007. Paper 19. Available from: https://aes.org/publications/elibrary-page/?id=13917
@inproceedings{Gallo2007_13917,
author = {Gallo, Emmanuel and Tsingos, Nicolas},
title = {{Extracting and Re-Rendering Structured Auditory Scenes from Field Recordings}},
booktitle = {AES Conference: 30th International Conference: Intelligent Audio Environments},
note = {Paper 19},
year = {2007},
month = mar,
publisher = {Audio Engineering Society},
url = {https://aes.org/publications/elibrary-page/?id=13917}
}
TY - CPAPER
TI - Extracting and Re-Rendering Structured Auditory Scenes from Field Recordings
AU - Gallo, Emmanuel
AU - Tsingos, Nicolas
T2 - AES Conference: 30th International Conference: Intelligent Audio Environments
M1 - Paper 19
PY - 2007
DA - 2007/03/06
UR - https://aes.org/publications/elibrary-page/?id=13917
PB - Audio Engineering Society
LA - en
AB - We present an approach to automatically extract and re-render a structured auditory scene from field recordings obtained with a small set of microphones, freely positioned in the environment. From the recordings and the calibrated position of the microphones, the 3D location of various auditory events can be estimated together with their corresponding content. This structured description is reproduction-setup independent. We propose solutions to classify foreground, well-localized sounds and more diffuse background ambiance and adapt our rendering strategy accordingly. Warping the original recordings during playback allows for simulating smooth changes in the listening point or position of sources. Comparisons to reference binaural and B-format recordings show that our approach achieves good spatial rendering while remaining independent of the reproduction setup and offering extended authoring capabilities.
ER -