Normative conflicts and shallow AI alignment.

The progress of AI systems such as large language models (LLMs) raises increasingly pressing concerns about their safe deployment. This paper examines the value alignment problem for LLMs, arguing that current alignment strategies are fundamentally inadequate to prevent misuse. Despite ongoing effor...

Descripción completa

Detalles Bibliográficos
Publicado en:Philosophical Studies Vol. 182; no. 7; pp. 2035 - 2079
Autor principal: Millière, Raphaël
Formato: Artículo
Publicado: Springer Nature Jul2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost