Locked History Actions

Diff for "Scwad/CDSCorpus"

Differences between revisions 14 and 15
Revision 14 as of 2017-08-09 09:18:05
Size: 1233
Comment:
Revision 15 as of 2017-08-09 09:18:42
Size: 1233
Comment:
Deletions are marked like this. Additions are marked like this.
Line 7: Line 7:
Polish CDSCorpus consists of 10K Polish sentence pairs which are human-annotated for semantic relatedness and entailment. The dataset may be used for the evaluation of compositional distributional semantics models of Polish. For more details, please refer to the paper describing the dataset [[http://www.aclweb.org/anthology/P/P17/P17-1073.pdf|(Wróblewska and Krasnowska-Kieraś, 2017)]]. Polish CDSCorpus consists of 10K Polish sentence pairs which are human-annotated for semantic relatedness and entailment. The dataset may be used for the evaluation of compositional distributional semantics models of Polish. For more details, please refer to the [[http://www.aclweb.org/anthology/P/P17/P17-1073.pdf|paper]] describing the dataset (Wróblewska and Krasnowska-Kieraś, 2017).

Polish CDSCorpus

Polish CDSCorpus consists of 10K Polish sentence pairs which are human-annotated for semantic relatedness and entailment. The dataset may be used for the evaluation of compositional distributional semantics models of Polish. For more details, please refer to the paper describing the dataset (Wróblewska and Krasnowska-Kieraś, 2017).

Download

You can have a look at a part of CDSCorpus (1K annotated sentence pairs). If you wish to get the entire CDSCorpus (10K annotated sentence pairs) please contact alina <at> ipipan.waw.pl (replace <at> with @).

Publication

Alina Wróblewska and Katarzyna Krasnowska-Kieraś (2017) Polish evaluation dataset for compositional distributional semantics models. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 784–792, DOI: doi.org/10.18653/v1/P17-1073.

Contact

For contacting Alina Wróblewska, please write to the email alina <at> ipipan.waw.pl.