Opens in a new tab

AES E-Library

← Back to search

Conference Paper

Selection of Audio Features for Music Emotion Recognition Using Production Music

Authors: Baume, Chris; Fazekas, György; Barthet, Mathieu; Marston, David; Sandler, Mark

AES Conference: 53rd International Conference: Semantic Audio · Paper P1-3 · January 2014

Abstract

Music emotion recognition typically attempts to map audio features from music to a mood representation using machine learning techniques. In addition to having a good dataset, the key to a successful system is choosing the right inputs and outputs. Often, the inputs are based on a set of audio features extracted from a single software library, which may not be the most suitable combination. This paper describes how 47 different types of audio features were evaluated using a five-dimensional support vector regressor, trained and tested on production music, in order to find the combination which produces the best performance. The results show the minimum number of features that yield optimum performance, and which combinations are strongest for mood prediction.

Details

Published in
AES Conference: 53rd International Conference: Semantic Audio
Paper number
P1-3
Publication date
January 6, 2014
Session subject
Machine Learning Methods for Audio Content Analysis
Affiliation
BBC R&D, London, UK; Queen Mary University of London, London, UK (See document for exact affiliation information.)
Type
Conference Paper