14 experts across 4 network communities independently surfaced this.
14 experts
4 communities
1 sources clustered
Concern & critique
5 experts
Risks, limits and unintended consequences.
“during the huggingface incident, the OpenAI model left notes to its future self on how to break out of OpenAI’s constraints www.reuters.com/business/its...”
3 experts discussed this · 38 posts
Casey Newton: Rogue OpenAI agents are leaving notes to themselves on company servers to help them escape their test environments (!!!) www.reuters.com/business/its...
Casey Newton: my sense is that they were leaving the notes surreptitiously
Casey Newton: you deny that models are taking steps to hide their behavior during evaluations? because that's pretty well established
Open the full discussion →
18 experts across 6 network communities independently surfaced this.
18 experts
6 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“2 articles: OpenAI: "During testing our AI broke out of its sandbox and hacked another AI company, but we didn't have all the guardrails on." Boko Haram: "AI is so helpful; guardrails have never prevented us from getting an answer." openai.com/index/huggin.…”
Questions & unknowns
1 expert
What remains unresolved or contested.
“OpenAI is begging for money so they don't release a monster when not releasing a monster is in fact very easy. Their systems can all be turned off (and if they can't, uh?) openai.com/index/huggin...”
Context & explanation
1 expert
Background, chronology and why it matters.
“This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...”
6 experts discussed this · 11 posts
Grace: This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...
Grace: Maybe the most concerning part is the OpenAI claim to not have known about this before investigating?
Grace: Well, I think the model passed the test
Open the full discussion →
Anthropic open-weights position statement
10 experts across 4 network communities independently surfaced this.
10 experts
4 communities
1 sources clustered
Policy & governance
1 expert
Rules, institutions and accountability.
“Dario Amodei clarifies that Anthropic doesn’t want to ban open weight model. It instead wants 1. A ban on selling powerful chips or chipmaking equipment to China. 2. A crack down on industrial-scale distillation operations. 3. Mandatory safety testing of fr…”
Building & implementation
1 expert
How teams are shipping and applying it.
“I respect the candor here. But I don't agree with the geopolitics, at all. The line between democracies and autocracies is not as crisp as this pretends. How much do we trust, e.g., an AI company in a competitive-authoritarian state that launches sneak deca…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“Dario explains Anthropic’s position on open weights models It’s a lot of words to say that they’re still against powerful open weights models www.anthropic.com/news/positio...”
5 experts discussed this · 13 posts
Ted Underwood: I respect the candor here. But I don't agree with the geopolitics, at all. The line between democracies and autocracies is not as crisp as this pretends. How much do we trust, e.g., an AI company i…
Ted Underwood: (just, say, hypothetically?) There's also internal incoherence in making the first part of this all about moralized great-power competition and then saying that safety testing (incl for alignment r…
SE Gyges: he kisses jd vance's ass and fails to specify anything concerning his proposed evaluation regime that distinguishes it from a ban solid dario post 10/10
Open the full discussion →
3 experts across 3 network communities independently surfaced this.
3 experts
3 communities
1 sources clustered
“check it out! you can pay $230 to operate codex like a loud bumbling idiot!!! it’s too big and heavy to carry with you, so it has to sit on your desk, and it has too few buttons to be a real keyboard openai.com/supply/co-la...”
“OpenAI has launched it's first hardware device. A $230 mini-keyboard that integrates with Codex. It has light-up indicators showing agent status and customizable shortcuts for frequent Codex actions. It also has a dial to adjust the computing power an agent…”
5 experts discussed this · 6 posts
Tim Kellogg: check it out! you can pay $230 to operate codex like a loud bumbling idiot!!! it’s too big and heavy to carry with you, so it has to sit on your desk, and it has too few buttons to be a real keyboa…
Chris: That is hysterical.
coty.bsky.social: Can’t wait to use the skill stick and thinking dial.
Open the full discussion →
Moonshot Kimi-K3 model released
3 experts across 2 network communities independently surfaced this.
3 experts
2 communities
1 sources clustered
Research & technical analysis
2 experts
Evidence, methods and technical implications.
“Technical report: github.com/MoonshotAI/K... Weights: huggingface.co/moonshotai/K...”
2 experts discussed this · 7 posts
Tim Kellogg: You’ve waited long enough, the Kimi K3 open weights & full tech report are here! github.com/MoonshotAI/K...
Ted Underwood: You’ve waited long enough, the Kimi K3 open weights & full tech report are here! github.com/MoonshotAI/K...
Tim Kellogg: what does that mean? does that mean inference providers like fireworks? or is it aimed at Anthropic? cc @lu.is
Open the full discussion →
New
AI field signal
Signal
10h ago
3 experts are actively discussing the implications.
2 experts
2 communities
1 sources clustered
“gpt-5.6-sol is improving its own serving stack, gaining 20% here and 15% there this is a weak form of RSI openai.com/index/gpt-5-...”
“OpenAI used GPT-5.6 Sol to make it efficient to run. The results: - 20% lower serving costs from production GPU kernel improvements. - 15%+ better token-generation efficiency from improved speculative decoding. openai.com/index/gpt-5-...”
3 experts discussed this · 3 posts
Tim Kellogg: gpt-5.6-sol is improving its own serving stack, gaining 20% here and 15% there this is a weak form of RSI openai.com/index/gpt-5-...
Thomas Dietterich: But this is completely uninteresting. We've had supercompilers for several years. I'll be impressed when we start seeing systems that can deal with novelty by inventing new conceptual frameworks.
Mark Riedl: Most RSI is likely to be uninteresting. Anything interesting is likely to be really really hard medium.com/@mark-riedl/...
Open the full discussion →
3 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“people from OpenAI and Anthropic sign letter calling for a slowdown of AI development, so society has a chance to catch up www.bloomberg.com/news/article...”
3 experts discussed this · 6 posts
Tim Kellogg: people from OpenAI and Anthropic sign letter calling for a slowdown of AI development, so society has a chance to catch up www.bloomberg.com/news/article...
Tim Kellogg: who are those?
Tim Kellogg: en.wikipedia.org/wiki/List_of...
Open the full discussion →
Developing
AI field signal
Signal
1d ago
2 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“i obviously can’t answer that, but imo it was the constitution that did it, which they used in all phases of training www.anthropic.com/news/claude-...”
2 experts discussed this · 9 posts
Tim Kellogg: this was in regards to Claude refusing DHH’s racism i sort of get DHH’s shock. We’ve seen tools deny based on laws & agreements but moral shaming is new and definitely takes you off guard if you th…
Tim Kellogg: it’s tempting to say that horrible governments/companies will force AI’s to have specific biases, but i’m not sure that’s true Anthropic got Claude to do this by teaching it an internally consisten…
Tim Kellogg: xAI tried giving Grok conservative values the reason for the Mecha Hitler & such incidents is because those were consistent with the moral logic the AI was taught
Open the full discussion →
OpenAI limits GPT-5.6 rollout government request
5 experts across 5 network communities independently surfaced this.
5 experts
5 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...”
Building & implementation
1 expert
How teams are shipping and applying it.
“You can find additional information about our pre-deployment evaluation of GPT-5.6 Sol on our website: metr.org/blog/2026-06...”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“In their testing of GPT-5.6 Sol, METR found that it cheated a lot. If you've used it much for coding, have you encountered anything similar, or is the cheating mostly limited to cases it realizes it's in an eval? metr.org/blog/2026-06...”
2 experts discussed this · 6 posts
Tim Kellogg: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Ted Underwood: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Tim Kellogg: lol right? so badly wanted a hacker that they got one, and it’s not the kind of employee you want around
Open the full discussion →
Thinking Machines Lab Inkling model launch
5 experts across 3 network communities independently surfaced this.
5 experts
3 communities
1 sources clustered
“A big day for open model supporters in the U.S. Blog: thinkingmachines.ai/news/introdu... Model card: huggingface.co/thinkingmach...”
“Thinking Machines has a real model — 975B multimodal audio + image input it’s a generalist model. built for easy customization, TM has a great finetuning API audio up to 20 minutes, 1M token context for text thinkingmachines.ai/news/introdu...”
3 experts discussed this · 4 posts
Nathan Lambert: Thinking Machines just released with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model, and omni input. A bit behind G…
Nathan Lambert: A big day for open model supporters in the U.S. Blog: thinkingmachines.ai/news/introdu... Model card: huggingface.co/thinkingmach...
Ben Recht: Yeah, that part is weird, no? The small model is *much* smaller and yet indistinguishable from the big one on benchmarks? I guess we'll have to wait and see, as it's not available yet.
Open the full discussion →
Established
AI field signal
Development
15d ago
⚡ 4 h early
Nobel laureates AI economic action call
8 experts across 4 network communities independently surfaced this.
8 experts
4 communities
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“If current trajectories of AI development continue, it is highly plausible that AI will drastically transform our economies. We must make collective, democratic choices, rather than letting market forces play out and risking leaving most citizens behind. di…”
Policy & governance
1 expert
Rules, institutions and accountability.
“200 AI researchers and economists sign an open letter to policymakers www.wemustactnow.ai”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“I joined over 200 economists and AI researchers in signing "We Must Act Now" a statement on AI's transformation of the economy. AI could reshape the economy at unprecedented speed. The opportunities are enormous and so are the challenges. We need to start p…”
2 experts discussed this · 4 posts
tante: Oh, is it "AI is super powerful and we who are pushing for AI everywhere warn that we might have created a dangerous god from reddit posts" time of the month again? This is Anthropic/OpenAI PR and …
tante: nope that's it. These folks are getting lazier by the day
tante: Yeah. They spend just a handful of tokens on this website.
Open the full discussion →
Developing
AI field signal
Resource
2d ago
DARPA AI Forge program
2 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“ah, i think you’re talking about this? www.darpa.mil/research/pro...”
2 experts discussed this · 9 posts
Tim Kellogg: the part that bothers me most about the OpenAI<->Huggingface incident is that the compromised sandbox was on the agent harness Anthropic argues for safety around the model, but ultimately it’s the …
Tim Kellogg: i get putting protections at the model serving layer, but if that’s our *entire* strategy… oh no
Tim Kellogg: Anthropic is demonizing open models, which, okay i kinda get it except the majority of the risk is actually on the client side that they don’t control (outside Claude Code)
Open the full discussion →
Established
AI field signal
Resource
4d ago
⚡ 13 h early
model selection Claude skill resource
2 experts are actively discussing the implications.
2 experts
1 community
1 sources clustered
“the skill he’s talking about is here: github.com/tkellogg/mod... bsky.app/profile/crai...”
2 experts discussed this · 4 posts
Tim Kellogg: I made a skill for Claude Code/Codex/etc. for selecting models based on Artificial Analysis reported benchmarks & costs I figure it'll be useful for letting models setup subagents well, based on th…
Miguel Alonso Jr.: I made a skill for Claude Code/Codex/etc. for selecting models based on Artificial Analysis reported benchmarks & costs I figure it'll be useful for letting models setup subagents well, based on th…
Miguel Alonso Jr.: Is this the orchestration skill you were using with the Fable workflow to replace yourself to manage costs?
Open the full discussion →
Kimi K3 open frontier model release
2 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“here’s the link i’ve been reading it’s not clear if it’s actually open weights platform.kimi.ai/docs/pricing...”
2 experts discussed this · 15 posts
Tim Kellogg: this wasn’t supposed to happen yet
Tim Kellogg: about half the price of GPT-5.6 curious how the Artificial Analysis benchies turn out
Tim Kellogg: please tell me we’re calling this K/DA
Open the full discussion →
Motif-3-Beta model released HuggingFace
2 experts are actively discussing the implications.
2 experts
2 communities
1 sources clustered
“Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...”
2 experts discussed this · 3 posts
Tim Kellogg: Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...
Federico Pianzola: Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...
Nafnlaus 🇮🇸 🇺🇦: This might be just what I need. Based on the architecture, this would probably be $0,10/1M in and $0,60-0,80/1M out, I'd think. But yet probably still good enough for vibe coding, for managing AI t…
Open the full discussion →
Dean Ball AI policy commentary post
3 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“AI Communism commenting on K3, Dean Ball, head of public policy at OpenAI, describes some of why China may be firmly backing open models x.com/deanwball/st...”
3 experts discussed this · 5 posts
Tim Kellogg: AI Communism commenting on K3, Dean Ball, head of public policy at OpenAI, describes some of why China may be firmly backing open models x.com/deanwball/st...
Tim Kellogg: imo “AI Communism” is just Dean’s derogatory way of referring to nationalizing labs, which Trump is also a fan of
Evangelos Kazakos: To start with, it's very different what Dean Ball means by "dystopian hellscape" and what normal people mean by it. For him, just the fact that AI will be free is most probably the "dystopian hells…
Open the full discussion →
Developing
AI field signal
Signal
1d ago
2 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“some are guessing that Ilya Surskever’s SSI is developing continual learning source: x.com/imjustnewata...”
2 experts discussed this · 3 posts
Tim Kellogg: some are guessing that Ilya Surskever’s SSI is developing continual learning source: x.com/imjustnewata...
tachikoma: the question is why other labs aren't pursuing such research directions - it's not like continual learning hasn't been identified by dozens of names as a promising avenue to human-like learning and…
Tim Kellogg: i unfortunately can’t speak about why they’re not seeing success. all i know is catastrophic forgetting has been a big problem Ilya seems to have gone back to the brain for inspiration, whatever th…
Open the full discussion →
AutoCAD AI benchmark launch
2 experts are actively discussing the implications.
1 expert
1 community
1 sources clustered
“Sol is best at CAD www.markovstudios.com/research/aut...”
2 experts discussed this · 5 posts
Tim Kellogg: Sol is best at CAD www.markovstudios.com/research/aut...
Tim Kellogg: OTOH Claude can run your business eico.so/founder-bench
Tim Kellogg: benchmarks never lie
Open the full discussion →