Modelling Emotion Dynamics in Song Lyrics with State Space Models

10/17/2022
by   Yingjin Song, et al.
0

Most previous work in music emotion recognition assumes a single or a few song-level labels for the whole song. While it is known that different emotions can vary in intensity within a song, annotated data for this setup is scarce and difficult to obtain. In this work, we propose a method to predict emotion dynamics in song lyrics without song-level supervision. We frame each song as a time series and employ a State Space Model (SSM), combining a sentence-level emotion predictor with an Expectation-Maximization (EM) procedure to generate the full emotion dynamics. Our experiments show that applying our method consistently improves the performance of sentence-level baselines without requiring any annotated songs, making it ideal for limited training data scenarios. Further analysis through case studies shows the benefits of our method while also indicating the limitations and pointing to future directions.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/15/2022

The Emotion is Not One-hot Encoding: Learning with Grayscale Label for Emotion Recognition in Conversation

In emotion recognition in conversation (ERC), the emotion of the current...
research
08/03/2021

EMOPIA: A Multi-Modal Pop Piano Dataset For Emotion Recognition and Emotion-based Music Generation

While there are many music datasets with emotion labels in the literatur...
research
06/16/2021

SEOVER: Sentence-level Emotion Orientation Vector based Conversation Emotion Recognition Model

For the task of conversation emotion recognition, recent works focus on ...
research
03/04/2021

Morphset:Augmenting categorical emotion datasets with dimensional affect labels using face morphing

Emotion recognition and understanding is a vital componentin human-machi...
research
10/11/2021

Cross Domain Emotion Recognition using Few Shot Knowledge Transfer

Emotion recognition from text is a challenging task due to diverse emoti...
research
01/26/2022

Self-attention fusion for audiovisual emotion recognition with incomplete data

In this paper, we consider the problem of multimodal data analysis with ...

Please sign up or login with your details

Forgot password? Click here to reset