arXiv · 1703.03923
A German Corpus for Text Similarity Detection Tasks
Abstract
Text similarity detection aims at measuring the degree of similarity between a pair of texts. Corpora available for text similarity detection are designed to evaluate the algorithms to assess the paraphrase level among documents. In this paper we present a textual German corpus for similarity detection. The purpose of this corpus is to automatically assess the similarity between a pair of texts and to evaluate different similarity measures, both for whole documents or for individual sentences. Therefore we have calculated several simple measures on our corpus based on a library of similarity functions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Juan-Manuel Torres-Moreno, Gerardo Sierra, Peter Peinl. 2017-03-11. A German Corpus for Text Similarity Detection Tasks. https://arxiv.org/abs/1703.03923
Cite the original work for its findings. Save a collection to share your selection of sources.