1 dataset found
  1. PAN14 Originality: Source Retrieval

    • zenodo.org
    Updated Sep 15, 2014
  2. Not seeing a result you expected?
    Learn how you can add new datasets to our index.

Share
Facebook
Twitter
Email
Click to copy link
Link copied

PAN14 Originality: Source Retrieval

Dataset updated Sep 15, 2014
Dataset provided by
Bauhaus University, Weimarhttp://www.uni-weimar.de/
Martin-Luther-University Halle-Wittenberghttp://www.uni-halle.de/
Leipzig Universityhttp://www.uni-leipzig.de/
Authors
Potthast, Martin; Hagen, Matthias; Beyer, Anne; Busse, Matthias; Tippmann, Martin; Rosso, Paolo; Stein, Benno
Description

We provide you with a training corpus that consists of suspicious documents. Each suspicious document is about a specific topic and may consist of plagiarized passages obtained from web pages on that topic found in the ClueWeb09 corpus.

Search
Clear search
Close search
Google apps
Main menu