Spotting Words in Medieval Manuscripts.

This article discusses the technology of handwritten text recognition (HTR) as a tool for the analysis of historical handwritten documents. We give a broad overview of this field of research, but the focus is on the use of a method called ‘word spotting’ for finding words directly and automatically...

Descripción completa

Detalles Bibliográficos
Publicado en:Studia Neophilologica Vol. 86; pp. 171 - 187
Autores principales: Wahlberg, Fredrik, Dahllöf, Mats, Mårtensson, Lasse, Brun, Anders
Formato: Artículo
Publicado: Taylor & Francis Ltd Jun2014 Supplement
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=95976787&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 95976787
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        00393274
        9DT
      jtl: Studia Neophilologica
      issn: 00393274
      maglogo: N
    pubinfo:
      dt: Jun2014 Supplement
      vid: 86
      pid: 377
      pub: Taylor & Francis Ltd
    artinfo:
      ui:
        95976787
        10.1080/00393274.2013.871975
      ppf: 171
      ppct: 16
      formats:
      tig:
        atl: Spotting Words in Medieval Manuscripts.
      aug:
        au:
          Wahlberg, Fredrik
          Dahllöf, Mats
          Mårtensson, Lasse
          Brun, Anders
      su:
        Nonbook materials
        Medieval manuscripts
        Archival materials
        Imaging systems
        Books
      sug:
        subj:
          Nonbook materials
          Medieval manuscripts
          Archival materials
          Imaging systems
          Books
      ab: This article discusses the technology of handwritten text recognition (HTR) as a tool for the analysis of historical handwritten documents. We give a broad overview of this field of research, but the focus is on the use of a method called ‘word spotting’ for finding words directly and automatically in scanned images of manuscript pages. We illustrate and evaluate this method by applying it to a medieval manuscript. Word spotting uses digital image analysis to represent stretches of writing as sequences of numerical features. These are intended to capture the linguistically significant aspects of the visual shape of the writing. Two potential words can then be compared mathematically and their degree of similarity assigned a value. Our version of this method gives a false positive rate of about 30%, when the true positive rate is close to 100%, for an application where we search for very frequent short words in a 16th-Century Old Swedish cursiva recentior manuscript. Word spotting would be of use e.g. to researchers who want to explore the content of manuscripts when editions or other transcriptions are unavailable.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      dt:
        @attributes:
          year: 2014
    holdings:
      @attributes:
        islocal: N