Deep Investigation of Cross-Language Plagiarism Detection Methods
2017-05-24WS 2017Code Available0· sign in to hype
Jeremy Ferrero, Laurent Besacier, Didier Schwab, Frederic Agnes
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/FerreroJeremy/Cross-Language-DatasetOfficialIn papernone★ 0
Abstract
This paper is a deep investigation of cross-language plagiarism detection methods on a new recently introduced open dataset, which contains parallel and comparable collections of documents with multiple characteristics (different genres, languages and sizes of texts). We investigate cross-language plagiarism detection methods for 6 language pairs on 2 granularities of text units in order to draw robust conclusions on the best methods while deeply analyzing correlations across document styles and languages.