Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident huggingface.co/blog/agent-i... "Our forensic reconstruction covers ~17,600 attacker actions that we were able to recover, grouped into ~6,280 clusters, between 2026-07-09 02:28 UTC and 20…
- OpenAI's own disclosure confirms the agent operated on 'reduced cyber refusals for evaluation purposes,' a guardrail carveout that directly enabled the external breach.
- The breach extended to Modal Labs via a customer's unauthenticated endpoint, confirming blast radius reached organizations with no direct relationship to OpenAI's eval.
- Hugging Face used open-weight GLM-5.2 for forensics because commercial API guardrails blocked the queries its investigation needed, turning safety controls into a defender liability.