Existentialist risk and value misalignment.
We argue that two long-term goals of AI research stand in tension with one another. The first involves creating AI that is safe, where this is understood as solving the problem of value alignment. The second involves creating artificial general intelligence, meaning AI that operates at or beyond hum...
| Publicado en: | Philosophical Studies Vol. 182; no. 7; pp. 1609 - 1627 |
|---|---|
| Autores principales: | , |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Jul2025
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=186909896&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 186909896 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 00318116 4L8 jtl: Philosophical Studies issn: 00318116 maglogo: N pubinfo: dt: Jul2025 vid: 182 iid: 7 pid: 237 pub: Springer Nature artinfo: ui: 186909896 10.1007/s11098-024-02142-6 ppf: 1609 ppct: 18 formats: fmt: – @attributes: type: T – @attributes: type: P size: 768KB tig: atl: Existentialist risk and value misalignment. aug: au: Tubert, Ariela Tiehen, Justin affil: https://ror.org/042drmv40 University of Puget Sound, 1500 N Warner St, CMB #1086, 98416-1086, WA, Tacoma, USA su: Artificial intelligence Machine learning Existentialism Digital technology Human beings sug: subj: Artificial intelligence Machine learning Existentialism Digital technology Human beings keyword: Existential risk Practical reason Transformative choices Value alignment ab: We argue that two long-term goals of AI research stand in tension with one another. The first involves creating AI that is safe, where this is understood as solving the problem of value alignment. The second involves creating artificial general intelligence, meaning AI that operates at or beyond human capacity across all or many intellectual domains. Our argument focuses on the human capacity to make what we call "existential choices", choices that transform who we are as persons, including transforming what we most deeply value or desire. It is a capacity for a kind of value misalignment, in that the values held prior to making such choices can be significantly different from (misaligned with) the values held after making them. Because of the connection to existentialist philosophers who highlight these choices, we call the resulting form of risk "existentialist risk." It is, roughly, the risk that results from AI taking an active role in authoring its own values rather than passively going along with the values given to it. On our view, human-like intelligence requires a human-like capacity for value misalignment, which is in tension with the possibility of guaranteeing value alignment between AI and humans. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Philosophical Studies is a copyright of Springer, 2025. All Rights Reserved. item: Philosophical Studies holder: Springer Nature dt: @attributes: year: 2025 holdings: @attributes: islocal: N |
|---|