Understanding models understanding language.

Landgrebe and Smith (Synthese 198(March):2061–2081, 2021) present an unflattering diagnosis of recent advances in what they call language-centric artificial intelligence—perhaps more widely known as natural language processing: The models that are currently employed do not have sufficient expressivi...

Descripción completa

Detalles Bibliográficos
Publicado en:Synthese Vol. 200; no. 6; pp. 1 - 17
Autor principal: Søgaard, Anders
Formato: Artículo
Publicado: Springer Nature Dec2022
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=159943479&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 159943479
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        00397857
        4LI
      jtl: Synthese
      issn: 00397857
      maglogo: N
    pubinfo:
      dt: Dec2022
      vid: 200
      iid: 6
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        159943479
        10.1007/s11229-022-03931-4
      ppf: 1
      ppct: 16
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 349KB
      tig:
        atl: Understanding models understanding language.
      aug:
        au: Søgaard, Anders
        affil: Department of Computer Science, Pioneer Centre for Artificial Intelligence, and Department of Philosophy, University of Copenhagen, Lyngbyvej 2, 2100, Copenhagen, Denmark
      sug:
      keyword:
        Artificial intelligence
        Language
        Mind
      ab: Landgrebe and Smith (Synthese 198(March):2061–2081, 2021) present an unflattering diagnosis of recent advances in what they call language-centric artificial intelligence—perhaps more widely known as natural language processing: The models that are currently employed do not have sufficient expressivity, will not generalize, and are fundamentally unable to induce linguistic semantics, they say. The diagnosis is mainly derived from an analysis of the widely used Transformer architecture. Here I address a number of misunderstandings in their analysis, and present what I take to be a more adequate analysis of the ability of Transformer models to learn natural language semantics. To avoid confusion, I distinguish between inferential and referential semantics. Landgrebe and Smith (2021)’s analysis of the Transformer architecture’s expressivity and generalization concerns inferential semantics. This part of their diagnosis is shown to rely on misunderstandings of technical properties of Transformers. Landgrebe and Smith (2021) also claim that referential semantics is unobtainable for Transformer models. In response, I present a non-technical discussion of techniques for grounding Transformer models, giving them referential semantics, even in the absence of supervision. I also present a simple thought experiment to highlight the mechanisms that would lead to referential semantics, and discuss in what sense models that are grounded in this way, can be said to understand language. Finally, I discuss the approach Landgrebe and Smith (2021) advocate for, namely manual specification of formal grammars that associate linguistic expressions with logical form.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Synthese is a copyright of Springer, 2022. All Rights Reserved.
      item: Synthese
      holder: Springer Nature
      dt:
        @attributes:
          year: 2022
    holdings:
      @attributes:
        islocal: N