Deep Investigation of Cross-Language Plagiarism Detection Methods

05/24/2017
by   Jeremy Ferrero, et al.
0

This paper is a deep investigation of cross-language plagiarism detection methods on a new recently introduced open dataset, which contains parallel and comparable collections of documents with multiple characteristics (different genres, languages and sizes of texts). We investigate cross-language plagiarism detection methods for 6 language pairs on 2 granularities of text units in order to draw robust conclusions on the best methods while deeply analyzing correlations across document styles and languages.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/19/2022

Detecting Text Formality: A Study of Text Classification Approaches

Formality is an important characteristic of text documents. The automati...
research
02/10/2017

UsingWord Embedding for Cross-Language Plagiarism Detection

This paper proposes to use distributed representation of words (word emb...
research
11/16/2020

Robust Facial Landmark Detection by Cross-order Cross-semantic Deep Network

Recently, convolutional neural networks (CNNs)-based facial landmark det...
research
11/18/2021

Detecting Cross-Language Plagiarism using Open Knowledge Graphs

Identifying cross-language plagiarism is challenging, especially for dis...
research
03/15/2019

An Exploration of State-of-the-art Methods for Offensive Language Detection

We provide a comprehensive investigation of different custom and off-the...
research
03/15/2019

SemEval 2019 Task 6: An exploration of state-of-the-art methods for offensive language detection

We provide a comprehensive investigation of different custom and off-the...
research
09/15/2022

Open Challenges in Synthetic Speech Detection

In this paper the current status and open challenges of synthetic speech...

Please sign up or login with your details

Forgot password? Click here to reset