Deception detection in text and its relation to the cultural dimension of individualism/collectivism

05/26/2021
by   Katerina Papantoniou, et al.
9

Deception detection is a task with many applications both in direct physical and in computer-mediated communication. Our focus is on automatic deception detection in text across cultures. We view culture through the prism of the individualism/collectivism dimension and we approximate culture by using country as a proxy. Having as a starting point recent conclusions drawn from the social psychology discipline, we explore if differences in the usage of specific linguistic features of deception across cultures can be confirmed and attributed to norms in respect to the individualism/collectivism divide. We also investigate if a universal feature set for cross-cultural text deception detection tasks exists. We evaluate the predictive power of different feature sets and approaches. We create culture/language-aware classifiers by experimenting with a wide range of n-gram features based on phonology, morphology and syntax, other linguistic cues like word and phoneme counts, pronouns use, etc., and token embeddings. We conducted our experiments over 11 datasets from 5 languages i.e., English, Dutch, Russian, Spanish and Romanian, from six countries (US, Belgium, India, Russia, Mexico and Romania), and we applied two classification methods i.e, logistic regression and fine-tuned BERT models. The results showed that our task is fairly complex and demanding. There are indications that some linguistic cues of deception have cultural origins, and are consistent in the context of diverse domains and dataset settings for the same language. This is more evident for the usage of pronouns and the expression of sentiment in deceptive language. The results of this work show that the automatic deception detection across cultures and languages cannot be handled in a unified manner, and that such approaches should be augmented with knowledge about cultural differences and the domains of interest.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/04/2019

Studying Cultural Differences in Emoji Usage across the East and the West

Global acceptance of Emojis suggests a cross-cultural, normative use of ...
research
03/31/2023

Cross-Cultural Transfer Learning for Chinese Offensive Language Detection

Detecting offensive language is a challenging task. Generalizing across ...
research
05/25/2023

Multi-lingual and Multi-cultural Figurative Language Understanding

Figurative language permeates human communication, but at the same time ...
research
03/07/2023

ChatGPT: Beginning of an End of Manual Linguistic Data Annotation? Use Case of Automatic Genre Identification

ChatGPT has shown strong capabilities in natural language generation tas...
research
03/28/2022

EnCBP: A New Benchmark Dataset for Finer-Grained Cultural Background Prediction in English

While cultural backgrounds have been shown to affect linguistic expressi...
research
06/28/2018

Cross-Discourse and Multilingual Exploration of Textual Corpora with the DualNeighbors Algorithm

Word choice is dependent on the cultural context of writers and their su...
research
02/03/2022

Quantifying knowledge synchronisation in the 21st century

Humans acquire and accumulate knowledge through language usage and eager...

Please sign up or login with your details

Forgot password? Click here to reset