Normative conflicts and shallow AI alignment.
The progress of AI systems such as large language models (LLMs) raises increasingly pressing concerns about their safe deployment. This paper examines the value alignment problem for LLMs, arguing that current alignment strategies are fundamentally inadequate to prevent misuse. Despite ongoing effor...
| Published in: | Philosophical Studies Vol. 182; no. 7; pp. 2035 - 2079 |
|---|---|
| Main Author: | |
| Format: | Article |
| Published: |
Springer Nature
Jul2025
|
| Subjects: | |
| Online Access: | View this record in EBSCOhost |