Opens in a new tab

AES E-Library

← Back to search

Conference Paper

Efficient Synthesis of Violin Sounds Using a BiLSTM Network Based Source Filter Model

Authors: Dai, Yi-Ren; Yang, Hung-Chih; Su, Alvin W.Y.

AES Conference: 2020 AES International Conference on Audio for Virtual and Augmented Reality (August 2020) · Paper 10476 · August 2020

Abstract

The dynamic changes in playing skills generated from bow-string interaction make synthesizing bowed string instrument sounds a difficult task. Recently, a source filter model incorporating the LSTM predictor and the granular wavetables gives encouraging results. However, the prediction error is still large and the model hasn’t caught the nuance caused by the constantly changing characteristics of a playing violin. In this paper, the granular wavetable is represented of DCT coefficients and a new training strategy is proposed to reduce the predictor error. In addition, we analyze the difference between the original violin tone and the corresponding synthesis tone. A random pitch perturbation and a DCT coefficient shaping method are proposed to imitate the changing characteristics since results sound regular.

Details

Published in
AES Conference: 2020 AES International Conference on Audio for Virtual and Augmented Reality (August 2020)
Paper number
10476
Publication date
August 6, 2020
Session subject
Synthesis
Affiliation
National Cheng-Kung University Tainan, Taiwan (See document for exact affiliation information.)
Type
Conference Paper