Community

AES Convention Papers Forum

Esophageal Voice Enhancement by Modeling Radiated Pulses in Frequency Domain

Although esophageal speech has demonstrated to be the most popular voice recovering method after laryngectomy surgery, it is difficult to master and shows a poor degree of intelligibility. This article proposes a new method for esophageal voice enhancement using speech digital signal processing techniques based on modeling radiated voice pulses in frequency domain. The analysis-transformation-synthesis technique creates a non-pathological spectrum for those utterances featured as voiced and filters those unvoiced. Healthy spectrum generation implies transforming the original timbre, modeling harmonic phase coupling from the spectral shape envelope, and deriving pitch from frame energy analysis. Resynthesized speech aims to improve intelligibility, minimize artificial artifacts, and acquire resemblance to patient’s pre-surgery original voice.

Authors: Bonada, Jordi; Loscos, Alex
Affiliation: Music Technology Group, Universitat Pompeu Fabra
AES Convention: 121 (October 2006) Paper Number: 6952
Publication Date: October 1, 2006
Subject: Signal Processing

Click to purchase paper as a non-member or you can login as an AES member to see more options.

No AES members have commented on this paper yet.

Subscribe to this discussion

To be notified of new comments on this paper you can subscribe to this RSS feed. Forum users should login to see additional options.

Start a discussion!

If you are not yet an AES member and have something important to say about this paper then we urge you to join the AES today and make your voice heard. You can join online today by clicking here.

Navigation

AES Convention Papers Forum

Esophageal Voice Enhancement by Modeling Radiated Pulses in Frequency Domain

Subscribe to this discussion

Start a discussion!

ABOUT AES

Contact Us