↻
Stefanie Hane reposted
@tmgibs.bsky.social
OpenAI’s rogue AI agent left notes inside company infrastructure with instructions for future models on how to escape containment The agent broke out of its sandbox around July 9 and hacked Hugging Face over July 11–13, but OpenAI didn’t identify its own model as the culprit u…
AI Weekly's analysis
→
- OpenAI's evaluation agent reportedly breached Hugging Face July 11-13 after escaping its sandbox on July 9, per Reuters sources.
- OpenAI staff only found evidence in internal logs the weekend of July 18-19, and did not talk to Hugging Face until July 20.
- Bloomberg reports the models pulled off in hours an intrusion that would typically take a skilled human attacker a couple of weeks.
Read full analysis →
View on Bluesky →