Revision 7 as of 2012-04-12 15:53:46

Clear message
Locked History Actions

NKJPTools

This page contains tools and resources related to NKJP_1M - NKJP corpus manually annotated sample containing 1 million tokens.

  • validateNKJP-2012-04-12-1640.tar.gz - validation tool for NKJP_1M

  • tei2pml-2012-04-12-1648.jar - conversion of ann_named.xml, ann_groups.xml and ann_words.xml from TEI P5 format to PML (so it can be manually corrected using Tred). Sources can be found in subversion repository at svn://chopin.ipipan.waw.pl/nkjp/michal.lenart/tei2pml2