Opens in a new tab

AES E-Library

← Back to search

Conference Paper

Speaker Recognition Method Combining FFT, Wavelet Functions and Neural Networks

Authors: Domitrovic, Hrvoje; Grubesa, Sanja; Grubesa, Tomislav

AES Conference: 26th International Conference: Audio Forensics in the Digital Age · Paper 2-2 · July 2005

Abstract

The method of speaker recognition based on wavelet functions and neural networks is presented in this paper. The wavelet functions are used to obtain the approximation function and the details of the speaker’s averaged spectrum in order to extract speaker’s voice characteristics from the frequency spectrum. The approximation function and the details are then used as input data for decision-making neural networks. In this recognition process, not only the decision on the speaker’s identity is made, but also the probability that the decision is correct can be provided.

Details

Published in
AES Conference: 26th International Conference: Audio Forensics in the Digital Age
Paper number
2-2
Publication date
July 6, 2005
Session subject
Audio Forensics in the Digital Age
Affiliation
Faculty of EE and Computing; RIZ Transmitters Co. (See document for exact affiliation information.)
Type
Conference Paper