Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Esophageal Voice Enhancement by Modeling Radiated Pulses in Frequency Domain

Authors: Bonada, Jordi; Loscos, Alex

AES Convention 121 · Paper 6952 · October 2006

Abstract

Although esophageal speech has demonstrated to be the most popular voice recovering method after laryngectomy surgery, it is difficult to master and shows a poor degree of intelligibility. This article proposes a new method for esophageal voice enhancement using speech digital signal processing techniques based on modeling radiated voice pulses in frequency domain. The analysis-transformation-synthesis technique creates a non-pathological spectrum for those utterances featured as voiced and filters those unvoiced. Healthy spectrum generation implies transforming the original timbre, modeling harmonic phase coupling from the spectral shape envelope, and deriving pitch from frame energy analysis. Resynthesized speech aims to improve intelligibility, minimize artificial artifacts, and acquire resemblance to patient’s pre-surgery original voice.

Details

Published in
AES Convention 121
AES Convention
121
Paper number
6952
Publication date
October 6, 2006
Session subject
Signal Processing
Affiliation
Music Technology Group, Universitat Pompeu Fabra (See document for exact affiliation information.)
Type
Convention Paper