This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...
Grace
Tracked through public AI activity and peer connections inside the directory.
- AI signals
- 22 past 30d
- Sources
- 16 distinct domains
- Discussions
- 120 past 30d
- Latest signal
- 1d ago
Articles & links
https://www.reuters.com/business/its-ai-agent-spent-days-hacking-company-sources-say-openai-did-not-notice-week-2026-07-24/
Unfortunately they don’t actually uncover the source of these motifs in this paper, but they do rule out a few possibilities (eg overrepresentation in the pretraining data) https://arxiv.org/abs/2605.26492
It's an edit of https://bambamramfan.github.io/ai-compass/
Bluesky loves to Trust the Experts until it comes to this particular issue. Here's Prof. Eric Schwitzgebel on the subject: https://arxiv.org/pdf/2510.09858
Have you seen this? https://arxiv.org/abs/2608.03958
- The paper reports that foundation model agents in stylized social dilemmas consistently converge to stable cooperation, contradicting classical predictions of mutual defection.
- The authors introduce the 'embedded Bayesian agent,' which models an agent as part of the universe it inhabits rather than an independent decision-maker.
- They propose 'embedded equilibrium' as a new solution concept replacing the Nash equilibrium for reasoning about modern AI agents.
Recent commentary
I’m probably in the 90th percentile of AI users in terms of time spent and I don’t think I’ve ever clicked a ✨ button
It’s funny how many plans for keeping superintelligent AI safe boil down to “we’ll outsmart it.” Like, you sort of have to rule that out from the premise
A lot of AI agents are having a flowers for algernon moment right now
There’s probably a lot of text on the internet that LLMs know is written by the same author but humans don’t
A Bluesky account that just uses an LLM to rewrite the posts of another account with plausible deniability
I think it would be very good for AI safety if HuggingFace sued OpenAI
Humans are powerful, and language models are powerful, but both bow in the presence of narrative
Cherishing the window of charming AI mistakes
In Grace's orbit
Center = Grace. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.