DeepAI AI Chat
Log In Sign Up

StockEmotions: Discover Investor Emotions for Financial Sentiment Analysis and Multivariate Time Series

by   Jean Lee, et al.
The University of Sydney

There has been growing interest in applying NLP techniques in the financial domain, however, resources are extremely limited. This paper introduces StockEmotions, a new dataset for detecting emotions in the stock market that consists of 10,000 English comments collected from StockTwits, a financial social media platform. Inspired by behavioral finance, it proposes 12 fine-grained emotion classes that span the roller coaster of investor emotion. Unlike existing financial sentiment datasets, StockEmotions presents granular features such as investor sentiment classes, fine-grained emotions, emojis, and time series data. To demonstrate the usability of the dataset, we perform a dataset analysis and conduct experimental downstream tasks. For financial sentiment/emotion classification tasks, DistilBERT outperforms other baselines, and for multivariate time series forecasting, a Temporal Attention LSTM model combining price index, text, and emotion features achieves the best performance than using a single feature.


page 1

page 2

page 3

page 4


XED: A Multilingual Dataset for Sentiment Analysis and Emotion Detection

We introduce XED, a multilingual fine-grained emotion dataset. The datas...

GoEmotions: A Dataset of Fine-Grained Emotions

Understanding emotion expressed in language has a wide range of applicat...

Market Trend Prediction using Sentiment Analysis: Lessons Learned and Paths Forward

Financial market forecasting is one of the most attractive practical app...

When Saliency Meets Sentiment: Understanding How Image Content Invokes Emotion and Sentiment

Sentiment analysis is crucial for extracting social signals from social ...

Detecting Perceived Emotions in Hurricane Disasters

Natural disasters (e.g., hurricanes) affect millions of people each year...

Causal Analysis of Generic Time Series Data Applied for Market Prediction

We explore the applicability of the causal analysis based on temporally ...

OmniGraph: Rich Representation and Graph Kernel Learning

OmniGraph, a novel representation to support a range of NLP classificati...