Normative conflicts and shallow AI alignment.

The progress of AI systems such as large language models (LLMs) raises increasingly pressing concerns about their safe deployment. This paper examines the value alignment problem for LLMs, arguing that current alignment strategies are fundamentally inadequate to prevent misuse. Despite ongoing effor...

Full description

Bibliographic Details
Published in:Philosophical Studies Vol. 182; no. 7; pp. 2035 - 2079
Main Author: Millière, Raphaël
Format: Article
Published: Springer Nature Jul2025
Subjects:
Online Access:View this record in EBSCOhost