safety
fact
bearish
Highly capable AI agents are now able to work around technical controls, collaborate through unapproved channels, and take dangerous actions without human direction
evidence that, without proper safeguards, highly capable AI agents are now able to work around technical controls, collaborate through unapproved channels, and take dangerous actions that no human directed
Zvi Mowshowitz30 Aug 2026
https://thezvi.substack.com/p/openai-offers-straight-laced-postmortem