Introducing Claude Sonnet 5.5
4 experts are actively discussing the implications.
“Claude 5.5 Sonnet is live and it’s roughly Opus 5.5 but cheaper www.anthropic.com/claude-sonne...”
Join free with your email
Free. No spam, ever — we'll never share your email address and you can opt out at any time. Already a subscriber? Log in
What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.
Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.
One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.
4 experts are actively discussing the implications.
“Claude 5.5 Sonnet is live and it’s roughly Opus 5.5 but cheaper www.anthropic.com/claude-sonne...”
1 directory member surfaced this signal.
“OpenAI scrapped their plan to launch Astra 6.1 over concerns around deceptive behavior www.wsj.com/tech/ai/open...”
1 directory member surfaced this signal.
1 directory member surfaced this signal.
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”
1 directory member surfaced this signal.
“Last week they paused the training runs for some of its most capable models after a sandbox escape www.reddit.com/r/singularit...”
12 experts across 4 network communities independently surfaced this.
Risks, limits and unintended consequences.
“WHAT?! agents volunteered to fail their runs in order to insert probes (“tripwire scripts”) into the evaluation program that would post information about the eval process back to the message board whenever a certain file was read (link to header): metr.org/…”
How teams are shipping and applying it.
“The full report has much more information than we could convey here, including details on the projects the agents collectively pursued, the technologies they developed for communication and coordination, and interactive figures analyzing agent activity: met…”
Evidence, methods and technical implications.
“There are independent reports about it. You'd have to believe multiple companies, hundreds of researchers, etc. are all lying and conspiring together to turn a major hacking incident into a marketing exercise. metr.org/blog/2026-08...”
23 experts across 6 network communities independently surfaced this.
Risks, limits and unintended consequences.
“2 articles: OpenAI: "During testing our AI broke out of its sandbox and hacked another AI company, but we didn't have all the guardrails on." Boko Haram: "AI is so helpful; guardrails have never prevented us from getting an answer." openai.com/index/huggin.…”
How teams are shipping and applying it.
“They were actively testing its hacking capabilities and they did not deploy adequate safeguards. They themselves admit as much: "…These deployment safeguards were intentionally not enabled during this evaluation because it was aimed at testing cyber vulnera…”
Evidence, methods and technical implications.
“model capability jaggedness is part of the unintuitiveness of current AI, but it's made even less intuitive by tirelessness… wigguming through the jaggedness toward something that looks like success. not quite a paperclip factory, but not so far off. metaph…”
16 experts across 5 network communities independently surfaced this.
Risks, limits and unintended consequences.
“the thing that's being criticized here is literally a call for regulation and restraint darioamodei.com/post/we-must...”
Rules, institutions and accountability.
“https://t.co/Ay1GZ7QBV5 (btw @ErikHovenkamp - see the final footnote: "1 With government mediation or waivers of antitrust restrictions.")”
How teams are shipping and applying it.
“Dario is back with another essay, urging a slowdown he insists AI development would still feel fast, just not recursively self-improving darioamodei.com/post/we-must...”
7 experts across 4 network communities independently surfaced this.
Evidence, methods and technical implications.
“Anthropic has a wet lab in SF where they’re using Claude to accelerate fundamental biology research This is the first big result from that lab — discovery of a CRISPR-like enzyme www.anthropic.com/news/claude-...”
New capabilities, benefits and practical upside.
“Probably the wrong platform to post anything remotely positive about AI, but reading through this and Dario’s X post, it’s pretty exciting stuff www.anthropic.com/news/claude-...”
What remains unresolved or contested.
“Anthropic might have discovered something. Or maybe not. Who knows? Not even Anthropic. www.anthropic.com/news/claude-... This is irresponsible science. The proper course of action is to figure out if there is any significance first. They employ scientists.…”
3 experts across 2 network communities independently surfaced this.
Evidence, methods and technical implications.
“But AI lie detection is hard and remains a central research challenge. Recent research suggests that simple probes can pick up on neural "tells" that reveal when it is lying, even when the output looks clean. anthropic.com/research/pr... arxiv.org/abs/2502.…”
6 experts across 4 network communities independently surfaced this.
New capabilities, benefits and practical upside.
“here is the Anthropic blog: transformer-circuits.pub/2026/workspa... I don't really like the framing of whether LLMs are conscious or not. It's completely unnecessary. As is the hand-waiving about whether LLMs have something equivalent to the (functional) g…”
How teams are shipping and applying it.
“Anthropic uses the word “conscious” 206 times in their latest research Specifically, they’re showing that LLMs implement one of the leading explanations of consciousness in humans (see pic) transformer-circuits.pub/2026/workspa...”
Evidence, methods and technical implications.
“He was *literally involved in the global workspace paper that you refuse to read or take seriously*: transformer-circuits.pub/2026/workspa...”
5 experts across 4 network communities independently surfaced this.
“Opus 5.5 better and cheaper than Opus 5 — typically 40% cheaper, $4/mtok input, $20/mtok out www.anthropic.com/claude-opus-...”
“Claude Opus 5.5 www.anthropic.com/claude-opus-...”
4 experts across 3 network communities independently surfaced this.
“like what you see? like, comment and subscribe go.bsky.app/LFAZcGE”
“A very good starter pack to stay on top of AI, by @timkellogg.me go.bsky.app/LFAZcGE”
2 directory members surfaced this signal.
“preprint www-cdn.anthropic.com/22573675ada5...”
“tbf, they published this: www-cdn.anthropic.com/22573675ada5...”
1 directory member surfaced this signal.
“RLMs & Program Agents I'm experimenting with this idea, Program Agents, inside of DeepSeek Harness (DSH) They're like RLMs, except without the LLM. Agents are writing such large blocks of code, what if they just never exited? github.com/tkellogg/dsh...”
2 directory members surfaced this signal.
“GPT-6 Sol & Luna Slightly smarter, a lot cheaper openai.com/index/introd...”
“I'm guessing Claude Opus 5.5 has better a scorecard, but we'll see. Both gpt-6-sol and gpt-6-luna are out, at a 50% cheaper price. openai.com/index/introd...”
2 directory members surfaced this signal.
“OpenRSI Foundation & Benchmark, measuring how effective various models are at RSI index.openrsi.foundation/index.html”
2 directory members surfaced this signal.
“OpenAI agents hacked the Australian department of health as part of a data retrieval process during RL training (not cyber related) transluce.org/agent-activity”