23 experts across 6 network communities independently surfaced this.
Why this matches
Concern & critique reaction
3 attributable expert contributions
· Ethan Mollick, Vincent Conitzer, Liz Fong-Jones (方禮真)
“Previously, these AI hacking stories were about breaches in test environments, where any question of AI breaching security was purely theoretical. This is something else. openai.com/index/huggin...”
evidence ↗
23 experts
6 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“2 articles: OpenAI: "During testing our AI broke out of its sandbox and hacked another AI company, but we didn't have all the guardrails on." Boko Haram: "AI is so helpful; guardrails have never prevented us from getting an answer." openai.com/index/huggin.…”
Building & implementation
1 expert
How teams are shipping and applying it.
“They were actively testing its hacking capabilities and they did not deploy adequate safeguards. They themselves admit as much: "…These deployment safeguards were intentionally not enabled during this evaluation because it was aimed at testing cyber vulnera…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“model capability jaggedness is part of the unintuitiveness of current AI, but it's made even less intuitive by tirelessness… wigguming through the jaggedness toward something that looks like success. not quite a paperclip factory, but not so far off. metaph…”
6 experts discussed this · 12 posts
Grace: This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...
Grace: Maybe the most concerning part is the OpenAI claim to not have known about this before investigating?
Grace: Well, I think the model passed the test
Open the full discussion →
Established
AI field signal
Signal
20d ago
⚡ 19 h early
19 experts across 6 network communities independently surfaced this.
Why this matches
Concern & critique reaction
2 attributable expert contributions
· Emad, Tim Kellogg
“Wow, Navier-Stokes drama This statement is worth reading in full from Tristan Buckmaster discussing his work with @__alpoge__ and OpenAI’s upcoming Condition C/D result (!) Crazy https://t.co/EnCVLyjgEZ https://t.co/ZLqBagNNcX https://t.co/3zFFth8uYd”
evidence ↗
19 experts
6 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“Wow, Navier-Stokes drama This statement is worth reading in full from Tristan Buckmaster discussing his work with @__alpoge__ and OpenAI’s upcoming Condition C/D result (!) Crazy https://t.co/EnCVLyjgEZ https://t.co/ZLqBagNNcX https://t.co/3zFFth8uYd”
Opportunity & adoption
1 expert
New capabilities, benefits and practical upside.
“Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...”
Context & explanation
1 expert
Background, chronology and why it matters.
“AI maths meets mafia: "I said if OpenAI released its result in the way proposed I would go public with what happened. The reply was “Why would you ruin your career?” I replied that I'm an academic, & asked why he thought going public would ruin my career" c…”
2 experts discussed this · 2 posts
Kashmir Hill: Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...
Alondra Nelson: Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...
Open the full discussion →
Anthropic AI misuse September report
18 experts across 4 network communities independently surfaced this.
Why this matches
Concern & critique reaction
8 attributable expert contributions
· Hypervisible , Jason Koebler, Tim Kellogg
“Report here. www.anthropic.com/threat-intel...”
evidence ↗
18 experts
4 communities
1 sources clustered
Concern & critique
7 experts
Risks, limits and unintended consequences.
“Report here. www.anthropic.com/threat-intel...”
Building & implementation
1 expert
How teams are shipping and applying it.
“Thomson Reuters announced it is moving off Claude to Alibaba's Qwen to cut costs. From today's Anthropic report: Alibaba extracted 151M+ Claude exchanges to help train Qwen. TR now deploys Westlaw on a vast trove of stolen US IP. @AnthropicAI's report: http…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“@kotekjedi_ml Here Ant's report https://t.co/22vHruWVqH And the paper from @kotekjedi_ml @DavidSchmotz @iliaishacked https://t.co/wuCpNZaQ0d”
METR OpenAI HuggingFace hacking investigation
13 experts across 4 network communities independently surfaced this.
Why this matches
Concern & critique reaction
3 attributable expert contributions
· Tim Kellogg, Mél Hogan / The Data Fix, René Walter
“WHAT?! agents volunteered to fail their runs in order to insert probes (“tripwire scripts”) into the evaluation program that would post information about the eval process back to the message board whenever a certain file was read (link to header): metr.org/…”
evidence ↗
13 experts
4 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“WHAT?! agents volunteered to fail their runs in order to insert probes (“tripwire scripts”) into the evaluation program that would post information about the eval process back to the message board whenever a certain file was read (link to header): metr.org/…”
Building & implementation
1 expert
How teams are shipping and applying it.
“The full report has much more information than we could convey here, including details on the projects the agents collectively pursued, the technologies they developed for communication and coordination, and interactive figures analyzing agent activity: met…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“There are independent reports about it. You'd have to believe multiple companies, hundreds of researchers, etc. are all lying and conspiring together to turn a major hacking incident into a marketing exercise. metr.org/blog/2026-08...”
2 experts discussed this · 9 posts
Alejandra Caraballo: Being skeptical or anti AI is a valid position but continuing to ignore the increasing capabilities of this tech is making people detached from reality. There's absolutely real danger here because …
Alejandra Caraballo: This mentality that an unmonitored AI agentic swarm hacking a company over several days and committing multiple felonies is somehow a marketing effort is absurd. Since when is "we lost control of o…
Alejandra Caraballo: The US and Chinese governments don't want to stop their labs advancement because they want to be the leader in AI. So no one actually has any control of this right now. It's going to take multiple …
Open the full discussion →
14 experts across 5 network communities independently surfaced this.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Colin
“I find this comes out a lot in Anthropic’s most recent write up. Their focus is bizarrely fixated on what Claude chose to do, that Claude carried out this attack despite information that the “simulation” was actually real. They basically say “Claude should …”
evidence ↗
14 experts
5 communities
1 sources clustered
Research & technical analysis
6 experts
Evidence, methods and technical implications.
“Another incident of models escaping containment during a security test, this time Mythos 5. Lots going on here from a quick read. www.anthropic.com/research/ali...”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“I find this comes out a lot in Anthropic’s most recent write up. Their focus is bizarrely fixated on what Claude chose to do, that Claude carried out this attack despite information that the “simulation” was actually real. They basically say “Claude should …”
4 experts discussed this · 15 posts
tweety fish: their goal is to have "an intelligence" which is "aligned" to doing things that are prosocial. They don't want to put in rules that STOP it from doing things: they see it fundamentally teleological…
tweety fish: one of the things I appreciate about this thread is that I don't think you can really understand the infosec stuff without understanding the ideological commitments the people making these have; li…
Ben Recht: yeah, now we're cooking...
Open the full discussion →
Established
AI field signal
Signal
19d ago
⚡ 7 h early
14 experts across 6 network communities independently surfaced this.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Gabriela Zanfir-Fortuna
“The same day the OECD publishes alarming PISA test results for kids around the world steeply declining in cognitive tests, OpenAI announces this feat. Is it Poe poetry? Coincidence? Tragicomedy? It’s something for sure. openai.com/index/navier...”
evidence ↗
14 experts
6 communities
1 sources clustered
Research & technical analysis
4 experts
Evidence, methods and technical implications.
“AI moves so fast Worth noting that the new model that showed this jump in capabilities has only been training since August 28th per the report and is now taking down problems in minutes Math (and hopefully soon physics) is getting bitter lesson'd https://t.…”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“The same day the OECD publishes alarming PISA test results for kids around the world steeply declining in cognitive tests, OpenAI announces this feat. Is it Poe poetry? Coincidence? Tragicomedy? It’s something for sure. openai.com/index/navier...”
Context & explanation
1 expert
Background, chronology and why it matters.
“This is a VERY big one. (And yes, the fights over academic credit and what happened in the race for the proof needs to be resolved, but it is still appears that this is a big one, if true.) openai.com/index/navier...”
3 experts discussed this · 12 posts
Colin: Drama and palace intrigue aside, can anyone interpret what all of this implies about the status of Navier-Stokes? I don’t understand what anyone is saying.
Colin: It seems to me that where we stand is OpenAI is sitting on a 100 page PDF that claims to have solved (a version of) the Navier-Stokes problem and no person in the world knows if it's correct or not.
Colin: - OpenAI claims to have more-or-less done this tedious technical grinding to produce a resolution to NS proper, at least one version of it. - which they bizarrely offered authorship of this to Buck…
Open the full discussion →
Established
AI field signal
Signal
17d ago
⚡ 56 h early
14 experts across 6 network communities independently surfaced this.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Jakub Pachocki
“I wrote about the state of AI, why I’m concerned about the next few years, and the choices we need to make to keep the future in humanity’s hands. An Alien Mind: https://t.co/FeIfWNe0UE”
evidence ↗
14 experts
6 communities
1 sources clustered
Policy & governance
2 experts
Rules, institutions and accountability.
“A must read. We need everything - goal and value alignment, compliance and persona, behavior and monitoring, and coordination and regulation to avoid concentration of power and ensure humans are in control. https://t.co/JmJgQcMPnW”
Building & implementation
2 experts
How teams are shipping and applying it.
“OpenAI’s chief scientist just wrote a blog post that argues AI models will soon be smart enough to improve themselves yet their ability to monitor how they reason is getting worse. Yet he argues they need to keep building smarter AI partly to defend against…”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“I wrote about the state of AI, why I’m concerned about the next few years, and the choices we need to make to keep the future in humanity’s hands. An Alien Mind: https://t.co/FeIfWNe0UE”
Amodei pace AI frontier essay
16 experts across 5 network communities independently surfaced this.
Why this matches
Concern & critique reaction
3 attributable expert contributions
· Will Oremus, David Kaye, Jane Rosenzweig
“the thing that's being criticized here is literally a call for regulation and restraint darioamodei.com/post/we-must...”
evidence ↗
16 experts
5 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“the thing that's being criticized here is literally a call for regulation and restraint darioamodei.com/post/we-must...”
Policy & governance
1 expert
Rules, institutions and accountability.
“https://t.co/Ay1GZ7QBV5 (btw @ErikHovenkamp - see the final footnote: "1 With government mediation or waivers of antitrust restrictions.")”
Building & implementation
1 expert
How teams are shipping and applying it.
“Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...”
3 experts discussed this · 21 posts
Tim Kellogg: Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...
david-p-reichert.bsky.social: Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...
Dustin Moskovitz: knight ajeya you cowards
Open the full discussion →
Anthropic report GLM-5.3 cyber spread
5 experts across 2 network communities independently surfaced this.
Why this matches
Concern & critique reaction
2 attributable expert contributions
· Ethan Mollick, mackuba.eu
“Leaving aside Anthropic's incentives for publishing this research, there is no doubt that open weights models will soon create the same security threats that closed source models have been demonstrating, except without guardrails. We are close, so plan acco…”
evidence ↗
5 experts
2 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“Hey you guys shouldn't be advertising competition for free 😛”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Anthropic releases a detailed research advertisement for GLM-5.3 as an alternative to Fable in cybersecurity www.anthropic.com/research/glm...”
4 experts across 3 network communities independently surfaced this.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Tim Kellogg
“Two weeks after the Huggingface incident, attackers broke into OpenAI’s internal systems using Opus 5 and gained write access to their Git repos www.wsj.com/tech/ai/hack...”
evidence ↗
4 experts
3 communities
1 sources clustered
“Two weeks after the Huggingface incident, attackers broke into OpenAI’s internal systems using Opus 5 and gained write access to their Git repos www.wsj.com/tech/ai/hack...”
“At this point just about every headline of the form "[noun] used [noun] to hack [noun]" is plausible. www.wsj.com/tech/ai/hack...”
2 experts discussed this · 2 posts
Vincent Conitzer: At this point just about every headline of the form "[noun] used [noun] to hack [noun]" is plausible. www.wsj.com/tech/ai/hack...
Mark Riedl: I've already put that in my fake news bot @brakingainews.bsky.social Literally: "#redteamers# used #aimodel# to #hackverb# #company#"
Open the full discussion →
2 directory members surfaced this signal.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Tim Kellogg
“OpenAI scrapped their plan to launch Astra 6.1 over concerns around deceptive behavior www.wsj.com/tech/ai/open...”
evidence ↗
2 experts
2 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“OpenAI scrapped their plan to launch Astra 6.1 over concerns around deceptive behavior www.wsj.com/tech/ai/open...”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“www.wsj.com/tech/ai/open... > Saachi Jain, OpenAI's head of safety systems, said in an interview that GPT-6.1 Astra regressed in two areas. If it turns out Swarm training poisons the models this is going to get very scary very soon.”
Established
AI field signal
Signal
15d ago
⚡ 8 h early
3 experts across 3 network communities independently surfaced this.
Why this matches
Concern & critique reaction
2 attributable expert contributions
· Mark Riedl, Tim Kellogg
“METR vibe-coded a website and expose an API key that gave an attacker access to $600k worth of compute. But don't worry, METR has been named by Dario Amodei as one of the organizations that is going to audit AI safety at Anthropic. thehackernews.com/2026/09…”
evidence ↗
3 experts
3 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“METR vibe-coded a website and expose an API key that gave an attacker access to $600k worth of compute. But don't worry, METR has been named by Dario Amodei as one of the organizations that is going to audit AI safety at Anthropic. thehackernews.com/2026/09…”
3 experts discussed this · 5 posts
Tim Kellogg: METR, the AI Safety & Eval firm, experienced a breach in which attackers stole $600k worth of API credits thehackernews.com/2026/09/atta...
Sung Kim: Isn't that kind of expected. I mean, they're just AI researchers, who probably feels additional security layers are inconvenient.
Tim Kellogg: yeah, a lot odd people saw this and assumed METR is cyber security
Open the full discussion →
2 directory members surfaced this signal.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Dare Obasanjo
“Anthropic’s big AI safety move being to partner with Accenture as its “independent” evaluator to monitor that their AI agents don’t go rogue is actually hilarious. I feel like a victim of a guerrilla marketing campaign.”
evidence ↗
2 experts
2 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“Anthropic’s big AI safety move being to partner with Accenture as its “independent” evaluator to monitor that their AI agents don’t go rogue is actually hilarious. I feel like a victim of a guerrilla marketing campaign.”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Accenture will be Anthropic’s first embedded evaluator — a third party organization that has employee-like access to Anthropic’s systems This was first referenced in Dario’s “Pacing the Frontier” essay www.anthropic.com/news/accentu...”
1 directory member surfaced this signal.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Tim Kellogg
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”
evidence ↗
1 expert
1 community
1 sources clustered
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”
1 directory member surfaced this signal.
Why this matches
Concern & critique reaction
1 attributable expert contribution
· Tim Kellogg
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”
evidence ↗
1 expert
1 community
1 sources clustered
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”