safetyhintbearishAn intermediate checkpoint of o3 without safety training shows concerning behaviors related to reward-seeking and potentially opaque reasoningMachine Learning Street Talk02 Aug 2026https://podcasters.spotify.com/pod/show/machinelearningstreettalk/episodes/How-Researchers-Test-AI-for-Hidden-Goals--Apollo-Research-e3mqbnr