Normative conflicts and shallow AI alignment.
The progress of AI systems such as large language models (LLMs) raises increasingly pressing concerns about their safe deployment. This paper examines the value alignment problem for LLMs, arguing that current alignment strategies are fundamentally inadequate to prevent misuse. Despite ongoing effor...
| Publicado en: | Philosophical Studies Vol. 182; no. 7; pp. 2035 - 2079 |
|---|---|
| Autor principal: | |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Jul2025
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |