Multi-Source Contrastive Learning from Musical Audio

02/14/2023
by   Christos Garoufis, et al.
0

Contrastive learning constitutes an emerging branch of self-supervised learning that leverages large amounts of unlabeled data, by learning a latent space, where pairs of different views of the same sample are associated. In this paper, we propose musical source association as a pair generation strategy in the context of contrastive music representation learning. To this end, we modify COLA, a widely used contrastive learning audio framework, to learn to associate a song excerpt with a stochastically selected and automatically extracted vocal or instrumental source. We further introduce a novel modification to the contrastive loss to incorporate information about the existence or absence of specific sources. Our experimental evaluation in three different downstream tasks (music auto-tagging, instrument classification and music genre classification) using the publicly available Magna-Tag-A-Tune (MTAT) as a source dataset yields competitive results to existing literature methods, as well as faster network convergence. The results also show that this pre-training method can be steered towards specific features, according to the selected musical source, while also being dependent on the quality of the separated sources.

READ FULL TEXT

page 2

page 6

research
04/24/2023

Pre-Training Strategies Using Contrastive Learning and Playlist Information for Music Classification and Similarity

In this work, we investigate an approach that relies on contrastive lear...
research
08/03/2021

Improving Music Performance Assessment with Contrastive Learning

Several automatic approaches for objective music performance assessment ...
research
03/17/2021

Contrastive Learning of Musical Representations

While supervised learning has enabled great advances in many areas of mu...
research
10/28/2022

On the Role of Visual Context in Enriching Music Representations

Human perception and experience of music is highly context-dependent. Co...
research
07/07/2022

Self-Supervised Learning of Music-Dance Representation through Explicit-Implicit Rhythm Synchronization

Although audio-visual representation has been proved to be applicable in...
research
09/18/2023

HumTrans: A Novel Open-Source Dataset for Humming Melody Transcription and Beyond

This paper introduces the HumTrans dataset, which is publicly available ...
research
07/18/2019

Leveraging Knowledge Bases And Parallel Annotations For Music Genre Translation

Prevalent efforts have been put in automatically inferring genres of mus...

Please sign up or login with your details

Forgot password? Click here to reset