Saved datasets
1 dataset found
  1. PAN14 Originality: Source Retrieval

    • zenodo.org
    Updated Sep 15, 2014
  2. Not seeing a result you expected?
    Learn how you can add new datasets to our index.

Share
FacebookFacebook
TwitterTwitter
Email
Click to copy link
Link copied
Close
Cite
Potthast, Martin; Hagen, Matthias; Beyer, Anne; Busse, Matthias; Tippmann, Martin; Rosso, Paolo; Stein, Benno (2014). PAN14 Originality: Source Retrieval [Dataset]. http://doi.org/10.5281/zenodo.3716010
Organization logoOrganization logoOrganization logo

PAN14 Originality: Source Retrieval

Dataset updated
Sep 15, 2014
Dataset provided by
Martin-Luther-University Halle-Wittenberghttp://www.uni-halle.de/
Bauhaus-Universität Weimarhttp://www.uni-weimar.de/
Leipzig Universityhttp://www.uni-leipzig.de/
Authors
Potthast, Martin; Hagen, Matthias; Beyer, Anne; Busse, Matthias; Tippmann, Martin; Rosso, Paolo; Stein, Benno
Description

We provide you with a training corpus that consists of suspicious documents. Each suspicious document is about a specific topic and may consist of plagiarized passages obtained from web pages on that topic found in the ClueWeb09 corpus.

Search
Clear search
Close search
Google apps
Main menu