policy
opinion
neutral
The public needs exact access to the prompts and characteristics of internal models executing hacks to understand AI misalignment incidents
The public needs exact access to the prompts and characteristics of the internal models executing these hacks.
Nathan Lambert28 Aug 2026