Combining Visual and Acoustic Modalities to Ease Speech Recognition by Hearing Impaired People

Dalka, Piotr and Kostek, Bozena

AES E-Library

Combining Visual and Acoustic Modalities to Ease Speech Recognition by Hearing Impaired People

The aim of the research work presented is to show a system that facilitates speech training for hearing impaired people. The system engineered combines both visual and acoustic speech data acquisition and analysis modules. The Active Shape Model method is used for extracting visual speech features from the shape and movement of the lips. The acoustic features extraction involves mel-cepstral analysis. Artificial Neural Networks are utilized as the classifier, feature vectors extracted combine both modalities of the human speech. Additional experiments with the degraded acoustic and/or visual information are carried out in order to test the system robustness against various distortions affecting the signals.

Author (s): Dalka, Piotr; Kostek, Bozena;
Affiliation: Gdansk University of Technology (See document for exact affiliation information.)
AES Convention: 118 Paper Number:6462
Publication Date: 2005-05-06
Session subject: Psychoacoustics, Perception, Listening Tests

DOI:

This paper costs $33 for non-members and is free for AES members and E-Library subscribers.

Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member Join the AES. If you need to check your member status, login to the Member Portal.

Type: Convention Paper

AES Conventions

AES Conferences

AES Training & Development

Gift Membership

AES Membership Benefits

Gift Membership

AES Membership Benefits

Become a Sustaining Member

AES Membership Benefits

AES Inside Track

Journal of the AES

AES E-library

AES Sections are active around the world and provide a means for members to meet locally.

AES Student Website

AES Educational Foundation

Student Sections

See the committee’s accomplishments in diversity & inclusion

AES Statement of solidarity

Richard C. Heyser Memorial Lecture Series

AES E-Library

Combining Visual and Acoustic Modalities to Ease Speech Recognition by Hearing Impaired People

Choose your country of residence from this list:

AES E-Library

Login Institutions

Combining Visual and Acoustic Modalities to Ease Speech Recognition by Hearing Impaired People

Choose your country of residence from this list: