Gaussian-smoothed Imbalance Data Improves Speech Emotion Recognition

02/17/2023
by   Xuefeng Liang, et al.
0

In speech emotion recognition tasks, models learn emotional representations from datasets. We find the data distribution in the IEMOCAP dataset is very imbalanced, which may harm models to learn a better representation. To address this issue, we propose a novel Pairwise-emotion Data Distribution Smoothing (PDDS) method. PDDS considers that the distribution of emotional data should be smooth in reality, then applies Gaussian smoothing to emotion-pairs for constructing a new training set with a smoother distribution. The required new data are complemented using the mixup augmentation. As PDDS is model and modality agnostic, it is evaluated with three SOTA models on the IEMOCAP dataset. The experimental results show that these models are improved by 0.2% - 4.8% and 1.5% - 5.9% in terms of WA and UA. In addition, an ablation study demonstrates that the key advantage of PDDS is the reasonable data distribution rather than a simple data augmentation.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/09/2023

Learning Emotional Representations from Imbalanced Speech Data for Speech Emotion Recognition and Emotional Text-to-Speech

Effective speech emotional representations play a key role in Speech Emo...
research
10/27/2020

CopyPaste: An Augmentation Method for Speech Emotion Recognition

Data augmentation is a widely used strategy for training robust machine ...
research
09/19/2023

Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition

In this paper, we explored how to boost speech emotion recognition (SER)...
research
03/15/2023

Reevaluating Data Partitioning for Emotion Detection in EmoWOZ

This paper focuses on the EmoWoz dataset, an extension of MultiWOZ that ...
research
03/03/2023

Unproportional mosaicing

Data shift is a gap between data distribution used for training and data...
research
06/18/2018

On Enhancing Speech Emotion Recognition using Generative Adversarial Networks

Generative Adversarial Networks (GANs) have gained a lot of attention fr...
research
06/05/2023

Synthesizing Affective Neurophysiological Signals Using Generative Models: A Review Paper

The integration of emotional intelligence in machines is an important st...

Please sign up or login with your details

Forgot password? Click here to reset