Topic-Guided Sampling For Data-Efficient Multi-Domain Stance Detection

by   Erik Arakelyan, et al.

Stance Detection is concerned with identifying the attitudes expressed by an author towards a target of interest. This task spans a variety of domains ranging from social media opinion identification to detecting the stance for a legal claim. However, the framing of the task varies within these domains, in terms of the data collection protocol, the label dictionary and the number of available annotations. Furthermore, these stance annotations are significantly imbalanced on a per-topic and inter-topic basis. These make multi-domain stance detection a challenging task, requiring standardization and domain adaptation. To overcome this challenge, we propose Topic Efficient StancE Detection (TESTED), consisting of a topic-guided diversity sampling technique and a contrastive objective that is used for fine-tuning a stance classifier. We evaluate the method on an existing benchmark of 16 datasets with in-domain, i.e. all topics seen and out-of-domain, i.e. unseen topics, experiments. The results show that our method outperforms the state-of-the-art with an average of 3.5 F1 points increase in-domain, and is more generalizable with an averaged increase of 10.2 F1 on out-of-domain evaluation while using ≤10% of the training data. We show that our sampling technique mitigates both inter- and per-topic class imbalances. Finally, our analysis demonstrates that the contrastive learning objective allows the model a more pronounced segmentation of samples with varying labels.


page 1

page 2

page 3

page 4


IMHO Fine-Tuning Improves Claim Detection

Claims are the central component of an argument. Detecting claims across...

Cross-Domain Label-Adaptive Stance Detection

Stance detection concerns the classification of a writer's viewpoint tow...

Contrastive Domain Adaptation for Early Misinformation Detection: A Case Study on COVID-19

Despite recent progress in improving the performance of misinformation d...

Mitigating Data Sparsity for Short Text Topic Modeling by Topic-Semantic Contrastive Learning

To overcome the data sparsity issue in short text topic modeling, existi...

Mention Annotations Alone Enable Efficient Domain Adaptation for Coreference Resolution

Although, recent advances in neural network models for coreference resol...

Bridging the gap between supervised classification and unsupervised topic modelling for social-media assisted crisis management

Social media such as Twitter provide valuable information to crisis mana...

Please sign up or login with your details

Forgot password? Click here to reset