Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Investigation of an Encoder-Decoder LSTM Model on the Enhancement of Speech Intelligibility in Noise for Hearing Impaired Listeners

Authors: Thoidis, Iordanis; Vrysis, Lazaros; Pastiadis, Konstantinos; Markou, Konstantinos; Papanikolaou, George

AES Convention 146 · Paper 10206 · March 2019

Abstract

Hearing impaired (HI) listeners often struggle to follow conversations when exposed in a complex acoustic environment. This is partly due to the reduced ability in recovering the target speech Temporal Envelope (ENV) cues from Temporal Fine Structure (TFS). This study investigates the enhancement of speech intelligibility in HI listeners by processing the ENV of speech signals corrupted by real-world environmental noise. An Encoder-Decoder Long Short Term Memory (LSTM) model is exploited after perceptually motivated processing stages to compensate for the important ENV characteristics of comprehensible speech for hearing impairment. The computational model is evaluated using the Short-Time Objective Intelligibility (STOI) measure for speech intelligibility. Finally, results indicate a 6% improvement in the mean STOI measure across different SNR values.

Details

Published in
AES Convention 146
AES Convention
146
Paper number
10206
Publication date
March 6, 2019
Session subject
Poster Session 4
Affiliation
Aristotle University of Thessaloniki, Thessaloniki, Greece (See document for exact affiliation information.)
Type
Convention Paper