The Electronic Corpus of 17th- and 18th-century Polish Texts.
The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of sampl...
| Published in: | Language Resources & Evaluation Vol. 56; no. 1; pp. 309 - 333 |
|---|---|
| Main Authors: | , , , , , , |
| Format: | Article |
| Published: |
Springer Nature
Mar2022
|
| Subjects: | |
| Online Access: | View this record in EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=155080016&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 155080016 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 1574020X 179V jtl: Language Resources & Evaluation issn: 1574020X maglogo: N pubinfo: dt: Mar2022 vid: 56 iid: 1 pid: 237 pub: Springer Nature artinfo: ui: 155080016 10.1007/s10579-021-09549-1 ppf: 309 ppct: 24 formats: fmt: – @attributes: type: T – @attributes: type: P size: 1.1MB tig: atl: The Electronic Corpus of 17th- and 18th-century Polish Texts. aug: au: Gruszczyński, Włodzimierz Adamiec, Dorota Bronikowska, Renata Kieraś, Witold Modrzejewski, Emanuel Wieczorek, Aleksandra Woliński, Marcin affil: Institute of Polish Language, Polish Academy of Sciences, Cracow, Poland Institute of Computer Science, Polish Academy of Sciences, Warsaw, Poland su: Corpora Polish language Slavic languages sug: subj: Corpora Polish language Slavic languages keyword: Corpus annotation Corpus construction Historical corpora Middle Polish ab: The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of samples extracted from over seven hundred texts written and published between 1601 and 1772, summing up to a total size of 13.5 million tokens which makes it one of the largest historical corpora for a Slavic language. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Language Resources & Evaluation is a copyright of Springer, 2022. All Rights Reserved. item: Language Resources & Evaluation holder: Springer Nature dt: @attributes: year: 2022 holdings: @attributes: islocal: N |
|---|