Utilizing ChatGPT Generated Data to Retrieve Depression Symptoms from Social Media

07/05/2023
by   Ana-Maria Bucur, et al.
0

In this work, we present the contribution of the BLUE team in the eRisk Lab task on searching for symptoms of depression. The task consists of retrieving and ranking Reddit social media sentences that convey symptoms of depression from the BDI-II questionnaire. Given that synthetic data provided by LLMs have been proven to be a reliable method for augmenting data and fine-tuning downstream models, we chose to generate synthetic data using ChatGPT for each of the symptoms of the BDI-II questionnaire. We designed a prompt such that the generated data contains more richness and semantic diversity than the BDI-II responses for each question and, at the same time, contains emotional and anecdotal experiences that are specific to the more intimate way of sharing experiences on Reddit. We perform semantic search and rank the sentences' relevance to the BDI-II symptoms by cosine similarity. We used two state-of-the-art transformer-based models (MentalRoBERTa and a variant of MPNet) for embedding the social media posts, the original and generated responses of the BDI-II. Our results show that using sentence embeddings from a model designed for semantic search outperforms the approach using embeddings from a model pre-trained on mental health data. Furthermore, the generated synthetic data were proved too specific for this task, the approach simply relying on the BDI-II responses had the best performance.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/09/2023

Detection of depression on social networks using transformers and ensembles

As the impact of technology on our lives is increasing, we witness incre...
research
04/20/2023

"Can We Detect Substance Use Disorder?": Knowledge and Time Aware Classification on Social Media from Darkweb

Opioid and substance misuse is rampant in the United States today, with ...
research
06/29/2023

Harnessing the Power of Hugging Face Transformers for Predicting Mental Health Disorders in Social Networks

Early diagnosis of mental disorders and intervention can facilitate the ...
research
05/21/2018

Multi-Perspective Relevance Matching with Hierarchical ConvNets for Social Media Search

Despite substantial interest in applications of neural networks to infor...
research
12/09/2022

Incorporating Emotions into Health Mention Classification Task on Social Media

The health mention classification (HMC) task is the process of identifyi...
research
11/14/2022

Semantic Similarity Models for Depression Severity Estimation

Depressive disorders constitute a severe public health issue worldwide. ...
research
11/10/2022

Assistive Completion of Agrammatic Aphasic Sentences: A Transfer Learning Approach using Neurolinguistics-based Synthetic Dataset

Damage to the inferior frontal gyrus (Broca's area) can cause agrammatic...

Please sign up or login with your details

Forgot password? Click here to reset