The Electronic Corpus of 17th- and 18th-century Polish Texts.

The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of sampl...

Full description

Bibliographic Details
Published in:Language Resources & Evaluation Vol. 56; no. 1; pp. 309 - 333
Main Authors: Gruszczyński, Włodzimierz, Adamiec, Dorota, Bronikowska, Renata, Kieraś, Witold, Modrzejewski, Emanuel, Wieczorek, Aleksandra, Woliński, Marcin
Format: Article
Published: Springer Nature Mar2022
Subjects:
Online Access:View this record in EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=155080016&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 155080016
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        1574020X
        179V
      jtl: Language Resources & Evaluation
      issn: 1574020X
      maglogo: N
    pubinfo:
      dt: Mar2022
      vid: 56
      iid: 1
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        155080016
        10.1007/s10579-021-09549-1
      ppf: 309
      ppct: 24
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 1.1MB
      tig:
        atl: The Electronic Corpus of 17th- and 18th-century Polish Texts.
      aug:
        au:
          Gruszczyński, Włodzimierz
          Adamiec, Dorota
          Bronikowska, Renata
          Kieraś, Witold
          Modrzejewski, Emanuel
          Wieczorek, Aleksandra
          Woliński, Marcin
        affil:
          Institute of Polish Language, Polish Academy of Sciences, Cracow, Poland
          Institute of Computer Science, Polish Academy of Sciences, Warsaw, Poland
      su:
        Corpora
        Polish language
        Slavic languages
      sug:
        subj:
          Corpora
          Polish language
          Slavic languages
      keyword:
        Corpus annotation
        Corpus construction
        Historical corpora
        Middle Polish
      ab: The paper describes the process of building the electronic corpus of 17th- and 18th-century Polish texts, a relatively large, balanced, structurally and morphologically annotated resource of the Middle Polish language, available for searching at https://www.korba.edu.pl. The corpus consists of samples extracted from over seven hundred texts written and published between 1601 and 1772, summing up to a total size of 13.5 million tokens which makes it one of the largest historical corpora for a Slavic language.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Language Resources & Evaluation is a copyright of Springer, 2022. All Rights Reserved.
      item: Language Resources & Evaluation
      holder: Springer Nature
      dt:
        @attributes:
          year: 2022
    holdings:
      @attributes:
        islocal: N