Saved datasets
1 dataset found
  1. z

    PAN14 Originality: Source Retrieval

    • zenodo.org
    Updated Sep 15, 2014
  2. Not seeing a result you expected?
    Learn how you can add new datasets to our index.

Share
FacebookFacebook
TwitterTwitter
Email
Click to copy link
Link copied
Close
Cite
Potthast, Martin; Hagen, Matthias; Beyer, Anne; Busse, Matthias; Tippmann, Martin; Rosso, Paolo; Stein, Benno (2014). PAN14 Originality: Source Retrieval [Dataset]. http://doi.org/10.5281/zenodo.3716010

PAN14 Originality: Source Retrieval

Explore at:
Dataset updated
Sep 15, 2014
Dataset provided by
Universität Leipzig
Martin-Luther-Universität Halle-Wittenberg
Bauhaus-Universität Weimar
Authors
Potthast, Martin; Hagen, Matthias; Beyer, Anne; Busse, Matthias; Tippmann, Martin; Rosso, Paolo; Stein, Benno
Description

We provide you with a training corpus that consists of suspicious documents. Each suspicious document is about a specific topic and may consist of plagiarized passages obtained from web pages on that topic found in the ClueWeb09 corpus.

Search
Clear search
Close search
Google apps
Main menu