Dataset of Propaganda Techniques of the State-Sponsored Information Operation of the People's Republic of China

by   Rong-Ching Chang, et al.

The digital media, identified as computational propaganda provides a pathway for propaganda to expand its reach without limit. State-backed propaganda aims to shape the audiences' cognition toward entities in favor of a certain political party or authority. Furthermore, it has become part of modern information warfare used in order to gain an advantage over opponents. Most of the current studies focus on using machine learning, quantitative, and qualitative methods to distinguish if a certain piece of information on social media is propaganda. Mainly conducted on English content, but very little research addresses Chinese Mandarin content. From propaganda detection, we want to go one step further to provide more fine-grained information on propaganda techniques that are applied. In this research, we aim to bridge the information gap by providing a multi-labeled propaganda techniques dataset in Mandarin based on a state-backed information operation dataset provided by Twitter. In addition to presenting the dataset, we apply a multi-label text classification using fine-tuned BERT. Potentially this could help future research in detecting state-backed propaganda online especially in a cross-lingual context and cross platforms identity consolidation.


page 1

page 2

page 3

page 4


Characterising User Content on a Multi-lingual Social Network

Social media has been on the vanguard of political information diffusion...

CL-UZH at SemEval-2023 Task 10: Sexism Detection through Incremental Fine-Tuning and Multi-Task Learning with Label Descriptions

The widespread popularity of social media has led to an increase in hate...

Detecting and Reasoning of Deleted Tweets before they are Posted

Social media platforms empower us in several ways, from information diss...

Volta at SemEval-2021 Task 6: Towards Detecting Persuasive Texts and Images using Textual and Multimodal Ensemble

Memes are one of the most popular types of content used to spread inform...

Understanding the Communist Party of China's Information Operations

The Communist Party of China is known to engage in Information Operation...

Learning Cross-lingual Embeddings from Twitter via Distant Supervision

Cross-lingual embeddings represent the meaning of words from different l...

PACO: Provocation Involving Action, Culture, and Oppression

In India, people identify with a particular group based on certain attri...

Please sign up or login with your details

Forgot password? Click here to reset