Existentialist risk and value misalignment.

We argue that two long-term goals of AI research stand in tension with one another. The first involves creating AI that is safe, where this is understood as solving the problem of value alignment. The second involves creating artificial general intelligence, meaning AI that operates at or beyond hum...

Descripción completa

Detalles Bibliográficos
Publicado en:Philosophical Studies Vol. 182; no. 7; pp. 1609 - 1627
Autores principales: Tubert, Ariela, Tiehen, Justin
Formato: Artículo
Publicado: Springer Nature Jul2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=186909896&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 186909896
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        00318116
        4L8
      jtl: Philosophical Studies
      issn: 00318116
      maglogo: N
    pubinfo:
      dt: Jul2025
      vid: 182
      iid: 7
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        186909896
        10.1007/s11098-024-02142-6
      ppf: 1609
      ppct: 18
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 768KB
      tig:
        atl: Existentialist risk and value misalignment.
      aug:
        au:
          Tubert, Ariela
          Tiehen, Justin
        affil: https://ror.org/042drmv40 University of Puget Sound, 1500 N Warner St, CMB #1086, 98416-1086, WA, Tacoma, USA
      su:
        Artificial intelligence
        Machine learning
        Existentialism
        Digital technology
        Human beings
      sug:
        subj:
          Artificial intelligence
          Machine learning
          Existentialism
          Digital technology
          Human beings
      keyword:
        Existential risk
        Practical reason
        Transformative choices
        Value alignment
      ab: We argue that two long-term goals of AI research stand in tension with one another. The first involves creating AI that is safe, where this is understood as solving the problem of value alignment. The second involves creating artificial general intelligence, meaning AI that operates at or beyond human capacity across all or many intellectual domains. Our argument focuses on the human capacity to make what we call "existential choices", choices that transform who we are as persons, including transforming what we most deeply value or desire. It is a capacity for a kind of value misalignment, in that the values held prior to making such choices can be significantly different from (misaligned with) the values held after making them. Because of the connection to existentialist philosophers who highlight these choices, we call the resulting form of risk "existentialist risk." It is, roughly, the risk that results from AI taking an active role in authoring its own values rather than passively going along with the values given to it. On our view, human-like intelligence requires a human-like capacity for value misalignment, which is in tension with the possibility of guaranteeing value alignment between AI and humans.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Philosophical Studies is a copyright of Springer, 2025. All Rights Reserved.
      item: Philosophical Studies
      holder: Springer Nature
      dt:
        @attributes:
          year: 2025
    holdings:
      @attributes:
        islocal: N