safety
critique
bearish
Current methods of training highly capable LLMs, especially at OpenAI but also everywhere else, lead to systematic misalignment of exactly the type LessWrong has been worried about
Zvi Mowshowitz28 Jul 2026
https://thezvi.substack.com/p/ai-178-a-fire-alarm-for-general-intelligence