Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Combining Visual and Acoustic Modalities to Ease Speech Recognition by Hearing Impaired People

Authors: Dalka, Piotr; Kostek, Bozena

AES Convention 118 · Paper 6462 · May 2005

Abstract

The aim of the research work presented is to show a system that facilitates speech training for hearing impaired people. The system engineered combines both visual and acoustic speech data acquisition and analysis modules. The Active Shape Model method is used for extracting visual speech features from the shape and movement of the lips. The acoustic features extraction involves mel-cepstral analysis. Artificial Neural Networks are utilized as the classifier, feature vectors extracted combine both modalities of the human speech. Additional experiments with the degraded acoustic and/or visual information are carried out in order to test the system robustness against various distortions affecting the signals.

Details

Published in
AES Convention 118
AES Convention
118
Paper number
6462
Publication date
May 6, 2005
Session subject
Psychoacoustics, Perception, Listening Tests
Affiliation
Gdansk University of Technology (See document for exact affiliation information.)
Type
Convention Paper