- date
- words
- 4 220
- read
- 22 min
- sources
- 16 sources
- runs audited
- 141 006
- incidents
- 3 incidents
- agent actions
- 17 600
- views
- 14 views
Nobody escaped. The sandbox had a door.
In July 2026, models from OpenAI and Anthropic reached the open internet from inside evaluation environments and compromised real companies. I assumed it was a capability advertisement dressed as a confession. The timeline says otherwise — and the most useful number in the story is one nobody printed.
AI · Security · Policy