Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Feature Extraction for Voice-driven Synthesis

Authors: Janer, Jordi

AES Convention 118 · Paper 6414 · May 2005

Abstract

This paper explores the singing voice from an unusual perspective, not as a
musical instrument but as a musical controller. A set of spectral processing
algorithms extract features form the input voice. These features are categorized
in four groups: excitation, vocal tract, voice quality and context. The
extracted values are then transmitted as Open Sound Control (OSC) messages to be
used in an external synthesis engine. In this document, we provide first a
technical description of the algorithms, and in a second part, we detail the
components of the system. A practical example of voice-driven synthesis using
PureData (Pd) is also presented.

Details

Published in
AES Convention 118
AES Convention
118
Paper number
6414
Publication date
May 6, 2005
Session subject
Analysis and Synthesis of Sound
Affiliation
Universitat Pompeu Fabra (See document for exact affiliation information.)
Type
Convention Paper