Opens in a new tab

AES E-Library

← Back to search

Convention Paper

Efficient Synthesis of Violin Sounds Using a BiLSTM Network Based Source Filter Model

Authors: Dai, Yi-Ren; Yang, Hung-Chih; Su, Alvin W.Y.

AES Convention 150 · Paper 10476 · May 2021

Abstract

The dynamic changes in playing skills generated from bow-string interaction make synthesizing bowed string instrument sounds a difficult task. Recently, a source filter model incorporating the LSTM predictor and the granular wavetables gives encouraging results. However, the prediction error is still large and the model hasn’t caught the nuance caused by the constantly changing characteristics of a playing violin. In this paper, the granular wavetable is represented of DCT coefficients and a new training strategy is proposed to reduce the predictor error. In addition, we analyze the difference between the original violin tone and the corresponding synthesis tone. A random pitch perturbation and a DCT coefficient shaping method are proposed to imitate the changing characteristics since results sound regular.

Details

Published in
AES Convention 150
AES Convention
150
Paper number
10476
Publication date
May 6, 2021
Session subject
Synthesis
Affiliation
National Cheng-Kung University Tainan, Taiwan (See document for exact affiliation information.)
Type
Convention Paper