Real-time Percussive Technique Recognition and Embedding Learning for the Acoustic Guitar

07/13/2023
by   Andrea Martelloni, et al.
0

Real-time music information retrieval (RT-MIR) has much potential to augment the capabilities of traditional acoustic instruments. We develop RT-MIR techniques aimed at augmenting percussive fingerstyle, which blends acoustic guitar playing with guitar body percussion. We formulate several design objectives for RT-MIR systems for augmented instrument performance: (i) causal constraint, (ii) perceptually negligible action-to-sound latency, (iii) control intimacy support, (iv) synthesis control support. We present and evaluate real-time guitar body percussion recognition and embedding learning techniques based on convolutional neural networks (CNNs) and CNNs jointly trained with variational autoencoders (VAEs). We introduce a taxonomy of guitar body percussion based on hand part and location. We follow a cross-dataset evaluation approach by collecting three datasets labelled according to the taxonomy. The embedding quality of the models is assessed using KL-Divergence across distributions corresponding to different taxonomic classes. Results indicate that the networks are strong classifiers especially in a simplified 2-class recognition task, and the VAEs yield improved class separation compared to CNNs as evidenced by increased KL-Divergence across distributions. We argue that the VAE embedding quality could support control intimacy and rich interaction when the latent space's parameters are used to control an external synthesis engine. Further design challenges around generalisation to different datasets have been identified.

READ FULL TEXT
research
12/07/2020

Multi-Instrumentalist Net: Unsupervised Generation of Music from Body Movements

We propose a novel system that takes as an input body movements of a mus...
research
11/01/2018

Neural Music Synthesis for Flexible Timbre Control

The recent success of raw audio waveform synthesis models like WaveNet m...
research
05/22/2018

Generative timbre spaces with variational audio synthesis

Timbre spaces have been used in music perception to study the relationsh...
research
10/24/2019

Fast and High-Quality Singing Voice Synthesis System based on Convolutional Neural Networks

The present paper describes singing voice synthesis based on convolution...
research
10/01/2019

Domain Expansion in DNN-based Acoustic Models for Robust Speech Recognition

Training acoustic models with sequentially incoming data – while both le...
research
08/21/2022

A Survey of Augmented Piano Prototypes: Has Augmentation Improved Learning Experiences?

Humans have been developing and playing musical instruments for millenni...

Please sign up or login with your details

Forgot password? Click here to reset