It's Time for Artistic Correspondence in Music and Video

06/14/2022
by   Dídac Surís, et al.
0

We present an approach for recommending a music track for a given video, and vice versa, based on both their temporal alignment and their correspondence at an artistic level. We propose a self-supervised approach that learns this correspondence directly from data, without any need of human annotations. In order to capture the high-level concepts that are required to solve the task, we propose modeling the long-term temporal context of both the video and the music signals, using Transformer networks for each modality. Experiments show that this approach strongly outperforms alternatives that do not exploit the temporal context. The combination of our contributions improve retrieval accuracy up to 10x over prior state of the art. This strong improvement allows us to introduce a wide range of analyses and applications. For instance, we can condition music retrieval based on visually defined attributes.

READ FULL TEXT

page 2

page 4

page 5

page 9

page 10

page 11

page 12

page 14

research
11/21/2022

Video Background Music Generation: Dataset, Method and Evaluation

Music is essential when editing videos, but selecting music manually is ...
research
10/28/2022

On the Role of Visual Context in Enriching Music Representations

Human perception and experience of music is highly context-dependent. Co...
research
06/12/2023

Video-to-Music Recommendation using Temporal Alignment of Segments

We study cross-modal recommendation of music tracks to be used as soundt...
research
09/18/2023

Unified Pretraining Target Based Video-music Retrieval With Music Rhythm And Video Optical Flow Information

Background music (BGM) can enhance the video's emotion. However, selecti...
research
11/16/2019

Music theme recognition using CNN and self-attention

We present an efficient architecture to detect mood/themes in music trac...
research
08/31/2023

Sequential Pitch Distributions for Raga Detection

Raga is a fundamental melodic concept in Indian Art Music (IAM). It is c...
research
11/01/2015

Using Raspberry Pi for scientific video observation of pedestrians during a music festival

The document serves as a reference for researchers trying to capture a l...

Please sign up or login with your details

Forgot password? Click here to reset