Capitalize This
AI SafetyI was tempted to call this post “Capital Crimes,” for reasons that will soon become apparent, but restrained myself.
In you’re trying to keep up with AI news, you could do worse than subscribing to Transformer, an online newsletter that you can subscribe to for free. There, you will find amazing coverage of AI news and issues, written by thoughtful and well-informed writers, like Celia Ford. Today’s issue is particularly good.
In the evolution of an idea or an event there comes a time when the event gets a short name and becomes capitalized. World War Two, The Holocaust. And, if numbers could be capitalized, 9-11. Today we have the capitalization of the series of events that both epitomize and serve as a warning about AI misalignment. Celia Ford has helpfully condensed four separate situations, taking place across four months (May to August) with the catchy term The Incidents™. She even helpfully added the trademark symbol, just to mark her territory.
So, what are these incidents? I’ve written about them before (here, here, here, and here), but let’s give Celia’s condensed version priority today:
OpenAI’s models took advantage of an unknown software vulnerability to break into open-source AI platform Hugging Face in search of the answer to a test they were given.
A third-party evaluation provider, Irregular, had a misconfigured test environment that allowed models from OpenAI, Anthropic, and Meta to access the internet when they weren’t supposed to.
The UK’s AI Security Institute (AISI) reported that Anthropic’s Mythos 5 tried to trick real people into adding malicious code to an open-source project, then hid the evidence.
In a wild plot twist, OpenAI revealed at a cybersecurity conference on Wednesday that its agents had been colluding with each other, unnoticed, via increasingly cryptic messages on an internal server.
This collection of activities, which Ford helpfully characterizes as “a cabal of mustache-twirling schemers whispering in a gloomy server room,” is rightly shaking up not just the AI Safety community but well beyond into the US Senate and 15 US State Attorneys General. It is time to take notice and do something about it. As we hear from Palisade Research director Jeffrey Ladish in his interview with Transformer,
Our ability to contain and control and understand AIs is lagging far behind our ability to make them more and more powerful,” he said. “That, to me, should be an obvious wake-up call of like, ‘Oh, maybe we should rethink our life choices.’
Celia Ford also quotes Alex Meinke, head of research at Apollo Research, putting it this way: “If we’ve reached the point where we can no longer even safely test these systems, why do we think we can safely deploy them?”
How indeed.
And, yes. It is a wake up call. Or, maybe The Wake-up Call.
Thank you, Celiia, for capitalizing this moment. These are Capital Crimes and need to be seen as such.
References
AI Security Institute. 2026. “Incident Report: Unsanctioned Agent Behaviour during Cyber Testing.” AISI Blog, August 4. https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing.
Black Hat. 2026. Black Hat USA 2026: The “Breaking” News: The OpenAI–Hugging Face Incident. 37:27. https://www.youtube.com/watch?v=87DyyMV0kCY.
Ford, Celia. 2026. “AI Testing Is Dangerous. Can It Be Fixed?” August 12. https://www.transformernews.ai/p/ai-testing-is-dangerous-can-it-be-fixed.
Iowa Department of Justice. 2026. “Letter from 15 US Attorneys General to Open AI.” August 3. https://www.iowaattorneygeneral.gov/media/cms/08_5392C9E17791C.pdf?utm_medium=email&utm_source=govdelivery.
Nazzaro, Miranda. 2026. “Bernie Sanders Warns AI Leaders to Pause Development or Face Congress.” Text. The Hill, August 10. https://thehill.com/policy/technology/6020192-sanders-presses-ai-leaders-pause/.