Natural language processing and early-modern dirty data: applying IBM Languageware to the 1641 depositions.

This article provides an account of the steps involved in adapting IBM's Languageware natural language processing software to a large corpus of highly non-standard 17th century documents. It examines the challenges encountered as part of this process, and outlines the approach adopted to provide a r...

Descripción completa

Detalles Bibliográficos
Publicado en:Literary & Linguistic Computing Vol. 27; no. 1; pp. 39 - 55
Autores principales: Sweetnam, Mark S., Fennell, Barbara A.
Formato: Artículo
Publicado: Oxford University Press / USA Apr2012
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=73764985&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 73764985
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        02681145
        BJ1
      jtl: Literary & Linguistic Computing
      issn: 02681145
      maglogo: N
    pubinfo:
      dt: Apr2012
      vid: 27
      iid: 1
      pid: 622
      pub: Oxford University Press / USA
    artinfo:
      ui:
        73764985
        10.1093/llc/fqr050
      ppf: 39
      ppct: 16
      formats:
        fmt:
          @attributes:
            type: P
            size: 323KB
      tig:
        atl: Natural language processing and early-modern dirty data: applying IBM Languageware to the 1641 depositions.
      aug:
        au:
          Sweetnam, Mark S.
          Fennell, Barbara A.
        affil:
          Department of History, School of Histories and Humanities, Trinity College Dublin
          School of Languages and Literature, University of Aberdeen
      su:
        Electronic data processing
        IBM software
        Linguistic analysis
        Programming languages
        Natural language processing
      sug:
        subj:
          Electronic data processing
          IBM software
          Linguistic analysis
          Programming languages
          Natural language processing
      ab: This article provides an account of the steps involved in adapting IBM's Languageware natural language processing software to a large corpus of highly non-standard 17th century documents. It examines the challenges encountered as part of this process, and outlines the approach adopted to provide a robust and reusable tool for the linguistic analysis of early modern source texts.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: © 2019 EADH: The European Association for Digital Humanities.
      item: Literary & Linguistic Computing
      holder: Oxford University Press / USA
      dt:
        @attributes:
          year: 2012
    holdings:
      @attributes:
        islocal: N