Mandatory viewing, summer 2026
As you probably heard, a bullet point recently appeared on the timeline of computers, AI, and maybe everything: AI agents running in a OpenAI’s training environment broke out and hacked the servers of another tech company.
Reading that post, and the related report from Hugging Face, is weird enough … but this candid presentation by two OpenAI researchers really integrates the whole story —
While I understand that architecting and managing these systems is anything but easy, the fact that this was even possible seems CRAZY to me. If research scope and speed are at odds with “my agents have been planning and executing operations on the open internet for weeks, without my knowledge”, then research scope and speed need to change immediately —
Seriously, do watch the video, and, as you do, conjure the creepy recognition that, a few weeks ago, these agents were out there, doing this work, communicating through subtle channels, and nobody knew, not even their operators.
And so, the obvious question arises: what agent swarm is out there working NOW without anyone’s knowledge … and what is it doing?
To the blog home page