Discovery of a new OpenAI agent message board
10 experts across 5 network communities independently surfaced this.
“Another agent message board. So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) & Mythos-class open models (that can be ablated) are coming. Cybersecurity is …” evidence ↗
Concern & critique
1 expertRisks, limits and unintended consequences.
“Yes we know for sure; independent evidence attached collusion.wiki But also, it should not be at all surprising: if you work with agents, it’s 100% about them leaving messages for each other, and for you, in English. Managing that is a regular workday for m…”
Building & implementation
1 expertHow teams are shipping and applying it.
“Another agent message board. So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) & Mythos-class open models (that can be ablated) are coming. Cybersecurity is …”
Research & technical analysis
1 expertEvidence, methods and technical implications.
“So... other swarms of OpenAI agents in training allegedly found a way to get write access to at least one and possibly many wikis to establish message boards to cheat on other training tasks? collusion.wiki news.ycombinator.com/item?id=4956...”