The Electronic Corpus of 17th- and 18th-century Polish Texts.

The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of sampl...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 56; no. 1; pp. 309 - 333
Autores principales: Gruszczyński, Włodzimierz, Adamiec, Dorota, Bronikowska, Renata, Kieraś, Witold, Modrzejewski, Emanuel, Wieczorek, Aleksandra, Woliński, Marcin
Formato: Artículo
Publicado: Springer Nature Mar2022
Materias:
Acceso en línea:Ver este registro en EBSCOhost
Descripción
Sumario:The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of samples extracted from over seven hundred texts written and published between 1601 and 1772, summing up to a total size of 13.5 million tokens which makes it one of the largest historical corpora for a Slavic language.