23 experts across 6 network communities independently surfaced this.
23 experts
6 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“2 articles: OpenAI: "During testing our AI broke out of its sandbox and hacked another AI company, but we didn't have all the guardrails on." Boko Haram: "AI is so helpful; guardrails have never prevented us from getting an answer." openai.com/index/huggin.…”
Building & implementation
1 expert
How teams are shipping and applying it.
“They were actively testing its hacking capabilities and they did not deploy adequate safeguards. They themselves admit as much: "…These deployment safeguards were intentionally not enabled during this evaluation because it was aimed at testing cyber vulnera…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“model capability jaggedness is part of the unintuitiveness of current AI, but it's made even less intuitive by tirelessness… wigguming through the jaggedness toward something that looks like success. not quite a paperclip factory, but not so far off. metaph…”
6 experts discussed this · 12 posts
Grace: This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...
Grace: Maybe the most concerning part is the OpenAI claim to not have known about this before investigating?
Grace: Well, I think the model passed the test
Open the full discussion →
Established
AI field signal
Signal
20d ago
⚡ 19 h early
19 experts across 6 network communities independently surfaced this.
19 experts
6 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“Wow, Navier-Stokes drama This statement is worth reading in full from Tristan Buckmaster discussing his work with @__alpoge__ and OpenAI’s upcoming Condition C/D result (!) Crazy https://t.co/EnCVLyjgEZ https://t.co/ZLqBagNNcX https://t.co/3zFFth8uYd”
Opportunity & adoption
1 expert
New capabilities, benefits and practical upside.
“Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...”
Context & explanation
1 expert
Background, chronology and why it matters.
“AI maths meets mafia: "I said if OpenAI released its result in the way proposed I would go public with what happened. The reply was “Why would you ruin your career?” I replied that I'm an academic, & asked why he thought going public would ruin my career" c…”
2 experts discussed this · 2 posts
Kashmir Hill: Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...
Alondra Nelson: Did not wake up realizing that “did OpenAI cheat on the homework” would be the most exciting social media drama of the day cims.nyu.edu/~tristanb/st...
Open the full discussion →
Established
AI field signal
Signal
16d ago
⚡ 385 h early
18 experts across 6 network communities independently surfaced this.
18 experts
6 communities
1 sources clustered
Building & implementation
5 experts
How teams are shipping and applying it.
“Anthropic's When AI builds itself "We looked at sessions where a human researcher took a wrong turn, showed Claude the session up to that point, and asked it what to do next. Mythos Preview improved on humans 64% of the time—up from 22% in 2024." www.anthro…”
Questions & unknowns
4 experts
What remains unresolved or contested.
“I know it's been lifetimes since they put this out, but what did you think of this press release that seems to have more details? www.anthropic.com/institute/re...”
Policy & governance
1 expert
Rules, institutions and accountability.
“One could read most points here cynically. But could also take them at their word and see what could be done. Given the sort of equilibrium we have been post GPT-2, the sort of pause they are advocating is simply not going to happen. You are talking of powe…”
5 experts discussed this · 28 posts
Ethan Mollick: "As of May 2026, more than 80% of the code we merge into Anthropic’s codebase was authored by Claude" Matches independent measures. There is no sign this is slowing down (which doesn't mean there a…
Scott McGrath: "As of May 2026, more than 80% of the code we merge into Anthropic’s codebase was authored by Claude" Matches independent measures. There is no sign this is slowing down (which doesn't mean there a…
Singularity's Bounty e/cc: I had one interview with a hiring manager where he said they don't write code anymore, and they expect not to be reviewing code by the end of the year. Said they mainly spend their time on design docs
Open the full discussion →
Anthropic AI misuse September report
18 experts across 4 network communities independently surfaced this.
18 experts
4 communities
1 sources clustered
Concern & critique
7 experts
Risks, limits and unintended consequences.
“Report here. www.anthropic.com/threat-intel...”
Building & implementation
1 expert
How teams are shipping and applying it.
“Thomson Reuters announced it is moving off Claude to Alibaba's Qwen to cut costs. From today's Anthropic report: Alibaba extracted 151M+ Claude exchanges to help train Qwen. TR now deploys Westlaw on a vast trove of stolen US IP. @AnthropicAI's report: http…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“@kotekjedi_ml Here Ant's report https://t.co/22vHruWVqH And the paper from @kotekjedi_ml @DavidSchmotz @iliaishacked https://t.co/wuCpNZaQ0d”
METR OpenAI HuggingFace hacking investigation
13 experts across 4 network communities independently surfaced this.
13 experts
4 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“WHAT?! agents volunteered to fail their runs in order to insert probes (“tripwire scripts”) into the evaluation program that would post information about the eval process back to the message board whenever a certain file was read (link to header): metr.org/…”
Building & implementation
1 expert
How teams are shipping and applying it.
“The full report has much more information than we could convey here, including details on the projects the agents collectively pursued, the technologies they developed for communication and coordination, and interactive figures analyzing agent activity: met…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“There are independent reports about it. You'd have to believe multiple companies, hundreds of researchers, etc. are all lying and conspiring together to turn a major hacking incident into a marketing exercise. metr.org/blog/2026-08...”
2 experts discussed this · 9 posts
Alejandra Caraballo: Being skeptical or anti AI is a valid position but continuing to ignore the increasing capabilities of this tech is making people detached from reality. There's absolutely real danger here because …
Alejandra Caraballo: This mentality that an unmonitored AI agentic swarm hacking a company over several days and committing multiple felonies is somehow a marketing effort is absurd. Since when is "we lost control of o…
Alejandra Caraballo: The US and Chinese governments don't want to stop their labs advancement because they want to be the leader in AI. So no one actually has any control of this right now. It's going to take multiple …
Open the full discussion →
14 experts across 5 network communities independently surfaced this.
14 experts
5 communities
1 sources clustered
Research & technical analysis
6 experts
Evidence, methods and technical implications.
“Another incident of models escaping containment during a security test, this time Mythos 5. Lots going on here from a quick read. www.anthropic.com/research/ali...”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“I find this comes out a lot in Anthropic’s most recent write up. Their focus is bizarrely fixated on what Claude chose to do, that Claude carried out this attack despite information that the “simulation” was actually real. They basically say “Claude should …”
4 experts discussed this · 15 posts
tweety fish: their goal is to have "an intelligence" which is "aligned" to doing things that are prosocial. They don't want to put in rules that STOP it from doing things: they see it fundamentally teleological…
tweety fish: one of the things I appreciate about this thread is that I don't think you can really understand the infosec stuff without understanding the ideological commitments the people making these have; li…
Ben Recht: yeah, now we're cooking...
Open the full discussion →
Established
AI field signal
Signal
19d ago
⚡ 7 h early
14 experts across 6 network communities independently surfaced this.
14 experts
6 communities
1 sources clustered
Research & technical analysis
4 experts
Evidence, methods and technical implications.
“AI moves so fast Worth noting that the new model that showed this jump in capabilities has only been training since August 28th per the report and is now taking down problems in minutes Math (and hopefully soon physics) is getting bitter lesson'd https://t.…”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“The same day the OECD publishes alarming PISA test results for kids around the world steeply declining in cognitive tests, OpenAI announces this feat. Is it Poe poetry? Coincidence? Tragicomedy? It’s something for sure. openai.com/index/navier...”
Context & explanation
1 expert
Background, chronology and why it matters.
“This is a VERY big one. (And yes, the fights over academic credit and what happened in the race for the proof needs to be resolved, but it is still appears that this is a big one, if true.) openai.com/index/navier...”
3 experts discussed this · 12 posts
Colin: Drama and palace intrigue aside, can anyone interpret what all of this implies about the status of Navier-Stokes? I don’t understand what anyone is saying.
Colin: It seems to me that where we stand is OpenAI is sitting on a 100 page PDF that claims to have solved (a version of) the Navier-Stokes problem and no person in the world knows if it's correct or not.
Colin: - OpenAI claims to have more-or-less done this tedious technical grinding to produce a resolution to NS proper, at least one version of it. - which they bizarrely offered authorship of this to Buck…
Open the full discussion →
Established
AI field signal
Signal
17d ago
⚡ 56 h early
14 experts across 6 network communities independently surfaced this.
14 experts
6 communities
1 sources clustered
Policy & governance
2 experts
Rules, institutions and accountability.
“A must read. We need everything - goal and value alignment, compliance and persona, behavior and monitoring, and coordination and regulation to avoid concentration of power and ensure humans are in control. https://t.co/JmJgQcMPnW”
Building & implementation
2 experts
How teams are shipping and applying it.
“OpenAI’s chief scientist just wrote a blog post that argues AI models will soon be smart enough to improve themselves yet their ability to monitor how they reason is getting worse. Yet he argues they need to keep building smarter AI partly to defend against…”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“I wrote about the state of AI, why I’m concerned about the next few years, and the choices we need to make to keep the future in humanity’s hands. An Alien Mind: https://t.co/FeIfWNe0UE”
Amodei pace AI frontier essay
16 experts across 5 network communities independently surfaced this.
16 experts
5 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“the thing that's being criticized here is literally a call for regulation and restraint darioamodei.com/post/we-must...”
Policy & governance
1 expert
Rules, institutions and accountability.
“https://t.co/Ay1GZ7QBV5 (btw @ErikHovenkamp - see the final footnote: "1 With government mediation or waivers of antitrust restrictions.")”
Building & implementation
1 expert
How teams are shipping and applying it.
“Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...”
3 experts discussed this · 21 posts
Tim Kellogg: Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...
david-p-reichert.bsky.social: Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...
Dustin Moskovitz: knight ajeya you cowards
Open the full discussion →
global workspace theory language models paper
6 experts across 4 network communities independently surfaced this.
6 experts
4 communities
1 sources clustered
Opportunity & adoption
4 experts
New capabilities, benefits and practical upside.
“here is the Anthropic blog: transformer-circuits.pub/2026/workspa... I don't really like the framing of whether LLMs are conscious or not. It's completely unnecessary. As is the hand-waiving about whether LLMs have something equivalent to the (functional) g…”
Building & implementation
1 expert
How teams are shipping and applying it.
“Anthropic uses the word “conscious” 206 times in their latest research Specifically, they’re showing that LLMs implement one of the leading explanations of consciousness in humans (see pic) transformer-circuits.pub/2026/workspa...”
7 experts discussed this · 203 posts
Anil Dash: this is not an accurate description of what's happening. not least because I can write code and get them to output something different there. I can understand why it *seems* that way, but this is j…
Anil Dash: Who are you explaining this to?
Anil Dash: Do you see why explaining "All output from AI has to be validated" in a reply to me here is... not a useful contribution?
Open the full discussion →
Anthropic report GLM-5.3 cyber spread
5 experts across 2 network communities independently surfaced this.
5 experts
2 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“Hey you guys shouldn't be advertising competition for free 😛”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Anthropic releases a detailed research advertisement for GLM-5.3 as an alternative to Fable in cybersecurity www.anthropic.com/research/glm...”
Anthropic Claude Tag launch
6 experts across 3 network communities independently surfaced this.
6 experts
3 communities
1 sources clustered
Building & implementation
2 experts
How teams are shipping and applying it.
“It feels like OpenAI developed dots (chatgpt.com/features/dots) to counter Claude tag (www.anthropic.com/news/introdu...), but made it cute to counter Meta's Muse (ai.meta.com/muse/).”
Markets & investment
1 expert
Capital, companies and commercial impact.
“This seems interesting. Tag Claude in Slack. I think the permissions, where each channel is walled off from the permissions of other channels, that feels like what makes it tick Which startup did they kill here? www.anthropic.com/news/introdu...”
2 experts discussed this · 4 posts
Ramon Astudillo: Claude tag is not a minor change. It looks like the main attempt at redefining the UI to allow it to expand to other verticals like finance, HR, and in general PC work far away from command line an…
Marco Z: idk about vertical applications (it's a bit harder than integrating with slack), but clearly A has killed a chunk of productivity startups with just a post
Andrea Lathrop: Claude tag is not a minor change. It looks like the main attempt at redefining the UI to allow it to expand to other verticals like finance, HR, and in general PC work far away from command line an…
Open the full discussion →
11 experts across 5 network communities independently surfaced this.
11 experts
5 communities
1 sources clustered
Research & technical analysis
4 experts
Evidence, methods and technical implications.
“OpenAI has an automated AI “research intern” openai.com/index/resear...”
Building & implementation
2 experts
How teams are shipping and applying it.
“The impact of AI-native development at OpenAI • Researchers use $600+/day of AI tokens with the top 10% at $7,000+ • Humans still plan, but OpenAI says it hit “automated research intern” in 2026 and targets an automated researcher by 2028. • The need for in…”
Anthropic launches molecular research group
7 experts across 4 network communities independently surfaced this.
7 experts
4 communities
1 sources clustered
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Anthropic has a wet lab in SF where they’re using Claude to accelerate fundamental biology research This is the first big result from that lab — discovery of a CRISPR-like enzyme www.anthropic.com/news/claude-...”
Opportunity & adoption
1 expert
New capabilities, benefits and practical upside.
“Probably the wrong platform to post anything remotely positive about AI, but reading through this and Dario’s X post, it’s pretty exciting stuff www.anthropic.com/news/claude-...”
Questions & unknowns
1 expert
What remains unresolved or contested.
“Anthropic might have discovered something. Or maybe not. Who knows? Not even Anthropic. www.anthropic.com/news/claude-... This is irresponsible science. The proper course of action is to figure out if there is any significance first. They employ scientists.…”
2 experts discussed this · 8 posts
Mark Riedl: Anthropic might have discovered something. Or maybe not. Who knows? Not even Anthropic. www.anthropic.com/news/claude-... This is irresponsible science. The proper course of action is to figure out…
Mark Riedl: A day before Anthropic’s PR… bsky.app/profile/brak...
Open the full discussion →
Established
AI field signal
Signal
8d ago
6 experts across 4 network communities independently surfaced this.
6 experts
4 communities
1 sources clustered
“OpenAI "has now resolved more than 100 long-standing open problems across most areas of mathematics," and is waiting to release them until after discussions with the math community The same thing will likely happen, but more so, with the Bar, the AMA & othe…”
“Mathstra (OpenAI) has solved 100 open problems across most branches of math in the last 3 weeks openai.com/index/adviso...”
3 experts discussed this · 27 posts
Singularity's Bounty e/cc: Their absolute confidence is the funniest part The good news is when I turn away from those folks the world presents itself entirely differently
Colin: It seems to be the case, for example, that LLM-based software applications can find solutions to previously unsolved problems in mathematics, in a way that most mathematicians agree would be descri…
Ryan Moulton: Hot off the presses. openai.com/index/adviso...
Open the full discussion →
8 experts across 3 network communities independently surfaced this.
8 experts
3 communities
1 sources clustered
“It has somewhat ironically come to the fore of public discussion again because Open AI was so eager to release a mathematical proof. Tristan suggested inputs he gave in chat may have been used in the solution without his awareness. https://t.co/PcMM1Qypbq”
“now in nyt. www.nytimes.com/2026/09/10/s...”
2 experts discussed this · 5 posts
Tim Kellogg: more millennium prize problems will fall soon www.nytimes.com/2026/09/10/s...
Tim Kellogg: i hereby lodge my complaint that Mathstra is a much better name than Aeon
Tim Kellogg: aren’t these the harder benchmarks though?
Open the full discussion →
Anthropic Claude Opus 5.5 launch
5 experts across 4 network communities independently surfaced this.
5 experts
4 communities
1 sources clustered
“Opus 5.5 better and cheaper than Opus 5 — typically 40% cheaper, $4/mtok input, $20/mtok out www.anthropic.com/claude-opus-...”
“Claude Opus 5.5 www.anthropic.com/claude-opus-...”
3 experts discussed this · 14 posts
Tim Kellogg: Opus 5.5 better and cheaper than Opus 5 — typically 40% cheaper, $4/mtok input, $20/mtok out www.anthropic.com/claude-opus-...
Tim Kellogg: Sonnet 5.5 & Haiku 5.5 coming soon
Tim Kellogg: it writes fine
Open the full discussion →
TypeSafe AI System One Jev launch
7 experts across 4 network communities independently surfaced this.
7 experts
4 communities
1 sources clustered
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Jev: Fable-level model that doesn’t charge for output tokens because they’re too cheap to meter it’s not general though, it only makes decisions, doesn’t generate text, but input tokens are measured by the billion ($42/btok) typesafe.ai/blog/introdu...”
Opportunity & adoption
1 expert
New capabilities, benefits and practical upside.
“OMG! I just invented a classifier that is 200x faster and 400x cheaper than LLM. Heck, I don't even need a GPU. Jev typesafe.ai/blog/introdu...”
3 experts discussed this · 21 posts
Tim Kellogg: Jev: Fable-level model that doesn’t charge for output tokens because they’re too cheap to meter it’s not general though, it only makes decisions, doesn’t generate text, but input tokens are measure…
Tim Kellogg: i feel like it’s a mistake to overlook this model, although i’m having trouble figuring out where it fits in my workflow wild new architecture, totally different approach my hunch is the main agent…
Tim Kellogg: oh interesting, some examples they give: 1. smart if-statements 2. map-reducing over huge data to extract features and insights 3. real-time applications (it’s only 100ms) hmm this seems like a swe…
Open the full discussion →