Towards Improved Model Design for Authorship Identification: A Survey on Writing Style Understanding

09/30/2020
by   Weicheng Ma, et al.
0

Authorship identification tasks, which rely heavily on linguistic styles, have always been an important part of Natural Language Understanding (NLU) research. While other tasks based on linguistic style understanding benefit from deep learning methods, these methods have not behaved as well as traditional machine learning methods in many authorship-based tasks. With these tasks becoming more and more challenging, however, traditional machine learning methods based on handcrafted feature sets are already approaching their performance limits. Thus, in order to inspire future applications of deep learning methods in authorship-based tasks in ways that benefit the extraction of stylistic features, we survey authorship-based tasks and other tasks related to writing style understanding. We first describe our survey results on the current state of research in both sets of tasks and summarize existing achievements and problems in authorship-related tasks. We then describe outstanding methods in style-related tasks in general and analyze how they are used in combination in the top-performing models. We are optimistic about the applicability of these models to authorship-based tasks and hope our survey will help advance research in this field.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/02/2020

Natural Language Processing Advancements By Deep Learning: A Survey

Natural Language Processing (NLP) helps empower intelligent machines by ...
research
09/18/2019

A Lexical, Syntactic, and Semantic Perspective for Understanding Style in Text

With a growing interest in modeling inherent subjectivity in natural lan...
research
07/13/2017

Is writing style predictive of scientific fraud?

The problem of detecting scientific fraud using machine learning was rec...
research
11/09/2019

xSLUE: A Benchmark and Analysis Platform for Cross-Style Language Understanding and Evaluation

Every natural text is written in some style. The style is formed by a co...
research
11/01/2020

Deep Learning for Text Attribute Transfer: A Survey

Driven by the increasingly larger deep learning models, neural language ...
research
10/16/2018

Creating a New Persian Poet Based on Machine Learning

In this article we describe an application of Machine Learning (ML) and ...
research
08/10/2023

Optical Script Identification for multi-lingual Indic-script

Script identification and text recognition are some of the major domains...

Please sign up or login with your details

Forgot password? Click here to reset