Saved datasets
1 dataset found
  1. PAN12 Originality: Source Retrieval

    • zenodo.org
    Updated Sep 17, 2012
  2. Not seeing a result you expected?
    Learn how you can add new datasets to our index.

Share
FacebookFacebook
TwitterTwitter
Email
Click to copy link
Link copied
Close
Cite
Potthast, Martin; Gollub, Tim; Hagen, Matthias; Graßegger, Jan; Kiesel, Johannes; Michel, Maximilian; Oberländer, Arnd; Tippmann, Martin; Barrón-Cedeño, Alberto; Gupta, Parth; Rosso, Paolo; Stein, Benno (2012). PAN12 Originality: Source Retrieval [Dataset]. http://doi.org/10.5281/zenodo.3713288
Organization logoOrganization logo

PAN12 Originality: Source Retrieval

Dataset updated Sep 17, 2012
Dataset provided by
Bauhaus-Universität Weimarhttps://www.uni-weimar.de/
Martin-Luther-University Halle-Wittenberghttp://www.uni-halle.de/
Universität Leipzig
Authors
Potthast, Martin; Gollub, Tim; Hagen, Matthias; Graßegger, Jan; Kiesel, Johannes; Michel, Maximilian; Oberländer, Arnd; Tippmann, Martin; Barrón-Cedeño, Alberto; Gupta, Parth; Rosso, Paolo; Stein, Benno
Description

We provide you with a training corpus that consists of suspicious documents. Each suspicious document is about a specific topic and may consist of plagiarized passages obtained from web pages on that topic found in the ClueWeb09 corpus.

Search
Clear search
Close search
Google apps
Main menu