P. Demonte, Y. Tang, R. J. Hughes, T. Cox, B. Fazenda, and B. Shirley, “Speech-To-Screen: Spatial Separation of Dialogue from Noise towards Improved Speech Intelligibility for the Small Screen,” in Proc. AES Convention 144, May 2018, Paper 10011. [Online]. Available: https://aes.org/publications/elibrary-page/?id=19407
Demonte P, Tang Y, Hughes RJ, Cox T, Fazenda B, Shirley B. Speech-To-Screen: Spatial Separation of Dialogue from Noise towards Improved Speech Intelligibility for the Small Screen. In: AES Convention 144. Audio Engineering Society; 2018. Paper 10011. Available from: https://aes.org/publications/elibrary-page/?id=19407
@inproceedings{Demonte2018_19407,
author = {Demonte, Philippa and Tang, Yan and Hughes, Richard J. and Cox, Trevor and Fazenda, Bruno and Shirley, Ben},
title = {{Speech-To-Screen: Spatial Separation of Dialogue from Noise towards Improved Speech Intelligibility for the Small Screen}},
booktitle = {AES Convention 144},
note = {Paper 10011},
year = {2018},
month = may,
publisher = {Audio Engineering Society},
url = {https://aes.org/publications/elibrary-page/?id=19407}
}
TY - CPAPER
TI - Speech-To-Screen: Spatial Separation of Dialogue from Noise towards Improved Speech Intelligibility for the Small Screen
AU - Demonte, Philippa
AU - Tang, Yan
AU - Hughes, Richard J.
AU - Cox, Trevor
AU - Fazenda, Bruno
AU - Shirley, Ben
T2 - AES Convention 144
M1 - Paper 10011
PY - 2018
DA - 2018/05/06
UR - https://aes.org/publications/elibrary-page/?id=19407
PB - Audio Engineering Society
LA - en
AB - Can externalizing dialogue when in the presence of stereo background noise improve speech intelligibility? This has been investigated for audio over headphones using head-tracking in order to explore potential future developments for small-screen devices. A quantitative listening experiment tasked participants with identifying target words in spoken sentences played in the presence of background noise via headphones. Sixteen different combinations of 3 independent variables were tested: speech and noise locations (internalized/externalized), video (on/off), and masking noise (stationary/fluctuating noise). The results revealed that the best improvements to speech intelligibility were generated by both the video-on condition and externalizing speech at the screen while retaining masking noise in the stereo mix.
ER -