Controlling Perceived Emotion in Symbolic Music Generation with Monte Carlo Tree Search

08/10/2022
by   Lucas N. Ferreira, et al.
3

This paper presents a new approach for controlling emotion in symbolic music generation with Monte Carlo Tree Search. We use Monte Carlo Tree Search as a decoding mechanism to steer the probability distribution learned by a language model towards a given emotion. At every step of the decoding process, we use Predictor Upper Confidence for Trees (PUCT) to search for sequences that maximize the average values of emotion and quality as given by an emotion classifier and a discriminator, respectively. We use a language model as PUCT's policy and a combination of the emotion classifier and the discriminator as its value function. To decode the next token in a piece of music, we sample from the distribution of node visits created during the search. We evaluate the quality of the generated samples with respect to human-composed pieces using a set of objective metrics computed directly from the generated samples. We also perform a user study to evaluate how human subjects perceive the generated samples' quality and emotion. We compare PUCT against Stochastic Bi-Objective Beam Search (SBBS) and Conditional Sampling (CS). Results suggest that PUCT outperforms SBBS and CS in almost all metrics of music quality and emotion.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/16/2020

Computer-Generated Music for Tabletop Role-Playing Games

In this paper we present Bardo Composer, a system to generate background...
research
04/25/2022

Which Discriminator for Cooperative Text Generation?

Language models generate texts by successively predicting probability di...
research
12/31/2021

Evaluating Deep Music Generation Methods Using Data Augmentation

Despite advances in deep algorithmic music generation, evaluation of gen...
research
01/14/2023

An Order-Complexity Model for Aesthetic Quality Assessment of Symbolic Homophony Music Scores

Computational aesthetics evaluation has made great achievements in the f...
research
09/28/2021

Generating texts under constraint through discriminator-guided MCTS

Large pre-trained language models (LM) based on Transformers allow to ge...
research
09/15/2023

Stack-and-Delay: a new codebook pattern for music generation

In language modeling based music generation, a generated waveform is rep...

Please sign up or login with your details

Forgot password? Click here to reset