N. Vryzas, R. Kotsakis, A. Liatsou, C. A. Dimoulas, and G. Kalliris, “Speech Emotion Recognition for Performance Interaction,” J. Audio Eng. Soc., vol. 66, no. 6, pp. 457–467, Jun. 2018, doi: 10.17743/jaes.2018.0036.
Vryzas N, Kotsakis R, Liatsou A, Dimoulas CA, Kalliris G. Speech Emotion Recognition for Performance Interaction. J Audio Eng Soc. 2018;66(6):457-467. doi:10.17743/jaes.2018.0036
@article{Vryzas2018_19585,
author = {Vryzas, Nikolaos and Kotsakis, Rigas and Liatsou, Aikaterini and Dimoulas, Charalampos A. and Kalliris, George},
title = {{Speech Emotion Recognition for Performance Interaction}},
journal = {Journal of the Audio Engineering Society},
volume = {66},
number = {6},
pages = {457--467},
year = {2018},
month = jun,
publisher = {Audio Engineering Society},
doi = {10.17743/jaes.2018.0036},
url = {https://doi.org/10.17743/jaes.2018.0036}
}
TY - JOUR
TI - Speech Emotion Recognition for Performance Interaction
AU - Vryzas, Nikolaos
AU - Kotsakis, Rigas
AU - Liatsou, Aikaterini
AU - Dimoulas, Charalampos A.
AU - Kalliris, George
T2 - Journal of the Audio Engineering Society
J2 - J. Audio Eng. Soc.
VL - 66
IS - 6
SP - 457
EP - 467
PY - 2018
DA - 2018/06/06
DO - 10.17743/jaes.2018.0036
UR - https://doi.org/10.17743/jaes.2018.0036
PB - Audio Engineering Society
LA - en
AB - This research explores the relevance of machine-driven Speech Emotion Recognition (SER) as a way to augment theatrical performances and interactions, such as controlling stage color/light, stimulating active audience engagement, actors’ interactive training, etc. It is well known that the meaning of a speech utterance arises from more than the linguistic content. Emotional affect can dramatically change meaning. As the basis for classification experiments, the authors developed the Acted Emotional Speech Dynamic Database (AESDD, which contains spoken utterances from 5 actors with 5 emotions. Several audio features and various classification techniques were implemented and evaluated using this database, as well comparing results with the Surrey Audio-Visual Expressed Emotion (SAVEE) database. The training classified was integrated into a novel application that performed live SER, fitting the needs of actor training.
ER -