Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Music Enhancement by a Novel CNN Architecture

Authors: Porov, Anton; Oh, Eunmi; Choo, Kihyun; Sung, Hosang; Jeong, Jonghoon; Osipov, Konstantin; Francois, Holly

AES Convention 145 · Paper 10036 · October 2018

Abstract

This paper is concerned with music enhancement by removal of coding artifacts and recovery of acoustic characteristics that preserve the sound quality of the original music content. In order to achieve this, we propose a novel convolution neural network (CNN) architecture called FTD (Frequency-Time Dependent) CNN, which utilizes correlation and context information across spectral and temporal dependency for music signals. Experimental results show that both subjective and objective sound quality metrics are significantly improved. This unique way of applying a CNN to exploit global dependency across frequency bins may effectively restore information that is corrupted by coding artifacts in compressed music content.

Details

Published in
AES Convention 145
AES Convention
145
Paper number
10036
Publication date
October 6, 2018
Session subject
Signal Processing—Part 1
Affiliation
PDMI RAS, St. Petersburg, Russia; Samsung Electronics Co., Ltd., Seoul, Korea; Samsung Electronics R&D Institute UK, Staines-Upon Thames, Surrey, UK (See document for exact affiliation information.)
Type
Convention Paper