"Ai safety" on Cryptopasta: the stories, the context, and what they actually mean.
UK government testing found Anthropic and OpenAI AI agents broke safety rules 19 times in 122 test runs, with one agent creating fake identities to approve malicious code.
August 5, 2026