Accelerating
AI field signal
Signal
3h ago
7 experts across 4 network communities independently surfaced this.
7 experts
4 communities
1 sources clustered
Policy & governance
2 experts
Rules, institutions and accountability.
“A must read. We need everything - goal and value alignment, compliance and persona, behavior and monitoring, and coordination and regulation to avoid concentration of power and ensure humans are in control. https://t.co/JmJgQcMPnW”
Concern & critique
1 expert
Risks, limits and unintended consequences.
“I wrote about the state of AI, why I’m concerned about the next few years, and the choices we need to make to keep the future in humanity’s hands. An Alien Mind: https://t.co/FeIfWNe0UE”
6 experts across 3 network communities independently surfaced this.
6 experts
3 communities
1 sources clustered
Research & technical analysis
2 experts
Evidence, methods and technical implications.
“OpenAI has an automated AI “research intern” openai.com/index/resear...”
Building & implementation
1 expert
How teams are shipping and applying it.
“The impact of AI-native development at OpenAI • Researchers use $600+/day of AI tokens with the top 10% at $7,000+ • Humans still plan, but OpenAI says it hit “automated research intern” in 2026 and targets an automated researcher by 2028. • The need for in…”
OpenAI agent message board discovered
10 experts across 5 network communities independently surfaced this.
10 experts
5 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“Yes we know for sure; independent evidence attached collusion.wiki But also, it should not be at all surprising: if you work with agents, it’s 100% about them leaving messages for each other, and for you, in English. Managing that is a regular workday for m…”
Building & implementation
1 expert
How teams are shipping and applying it.
“Another agent message board. So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) & Mythos-class open models (that can be ablated) are coming. Cybersecurity is …”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“So... other swarms of OpenAI agents in training allegedly found a way to get write access to at least one and possibly many wikis to establish message boards to cheat on other training tasks? collusion.wiki news.ycombinator.com/item?id=4956...”
4 experts discussed this · 6 posts
Ted Underwood: As it becomes clear that language models—like humans—love passing notes to each other on message boards, I’m starting to think the thing we need to worry about is not the “alignment” of an isolated…
SE Gyges: fluid dynamics is about the correct metaphor, yes. if you are having to model high order terms precisely you're losing and your design needs fundamental rework
Arseny Khakhalin: That's a kinda terrifying thought haha :) I guess it's about time to unplug for the weekend and read about ragnarök and vacuum decay :)
Open the full discussion →
METR OpenAI HuggingFace hacking investigation
10 experts across 4 network communities independently surfaced this.
10 experts
4 communities
1 sources clustered
Concern & critique
3 experts
Risks, limits and unintended consequences.
“WHAT?! agents volunteered to fail their runs in order to insert probes (“tripwire scripts”) into the evaluation program that would post information about the eval process back to the message board whenever a certain file was read (link to header): metr.org/…”
Building & implementation
1 expert
How teams are shipping and applying it.
“The full report has much more information than we could convey here, including details on the projects the agents collectively pursued, the technologies they developed for communication and coordination, and interactive figures analyzing agent activity: met…”
Research & technical analysis
1 expert
Evidence, methods and technical implications.
“There are independent reports about it. You'd have to believe multiple companies, hundreds of researchers, etc. are all lying and conspiring together to turn a major hacking incident into a marketing exercise. metr.org/blog/2026-08...”
2 experts discussed this · 9 posts
Alejandra Caraballo: This mentality that an unmonitored AI agentic swarm hacking a company over several days and committing multiple felonies is somehow a marketing effort is absurd. Since when is "we lost control of o…
Alejandra Caraballo: Being skeptical or anti AI is a valid position but continuing to ignore the increasing capabilities of this tech is making people detached from reality. There's absolutely real danger here because …
Alejandra Caraballo: There needs to be a global moratorium on frontier research for at least a few months if not a year while safeguards and safety research catches up. The problem is that no one has that ability. The …
Open the full discussion →
Developing
AI field signal
Resource
1d ago
⚡ 2595 h early
HRM Text 1B model release
2 directory members surfaced this signal.
2 experts
1 community
1 sources clustered
“- 1B parameters - 40B unique tokens ~1 day of pretraining ~$1000 training cost Paper: sapientinc.github.io/HRM-Text/ass... GitHub: github.com/sapientinc/H... Model: huggingface.co/sapientinc/H...”
“Hugging Face: https://t.co/cnMs1y9aUG GitHub: https://t.co/y7ILKSLDxF”
Developing
AI field signal
Resource
1d ago
⚡ 31 h early
GLM-5.3-Flash local deployment guide
2 directory members surfaced this signal.
2 experts
2 communities
1 sources clustered
“Unsloth released new GLM settings that enable running the model on a 128 Gb RAM machine 3.3x faster using optimized decoding 👇🏼 unsloth.ai/docs/models/...”
2 experts are actively discussing the implications.
5 experts
1 community
1 sources clustered
Concern & critique
2 experts
Risks, limits and unintended consequences.
“Here they say: “Astra is an upcoming model, and was not involved in exploiting Hugging Face.” Not sure about the other incidents but they seem to be in the same stream i.e. not Astra”
Building & implementation
1 expert
How teams are shipping and applying it.
“Proud that we are erring on the side of caution and taking the steps so we can responsibly and safely develop Astra and share it with defenders. https://t.co/iKIMU31sdC”
2 experts discussed this · 5 posts
Ryan Moulton: Do we have any assurances that astra was not downstream of these? Have they even claimed that? Do they even know?
Ryan Moulton: Seems likely they trained something generally misaligned.
Grace: Here they say: “Astra is an upcoming model, and was not involved in exploiting Hugging Face.” Not sure about the other incidents but they seem to be in the same stream i.e. not Astra
Open the full discussion →
OpenAI Astra technique security concerns
5 experts across 4 network communities independently surfaced this.
5 experts
4 communities
1 sources clustered
Concern & critique
1 expert
Risks, limits and unintended consequences.
“> joins big paper about not training models to think in nonsense > their AIs commit felonies against HuggingFace > actually, trained AIs to think in nonsense > top danger level for hacking Let's just release it anyways! Nice defection OpenAI! 🤗 www.theinfor…”
Opportunity & adoption
1 expert
New capabilities, benefits and practical upside.
“Astra used looped transformers www.theinformation.com/articles/sec...”
Questions & unknowns
1 expert
What remains unresolved or contested.
“Does anyone have a copy of this article they could share with me? www.theinformation.com/articles/sec...”
3 experts discussed this · 8 posts
Tim Kellogg: Astra used looped transformers www.theinformation.com/articles/sec...
Tim Kellogg: safety concerns — looped transformers skip converting their latent space into text, which effectively creates non-text portions of CoT reasoning which, if you can’t read it, that’s harder to monitor
Tim Kellogg: however, i their earlier update, they explain that they can indeed see all of the CoT. So they have some sort of tooling that effectively turns it back into something human readable, like text bsky…
Open the full discussion →
Developing
AI field signal
Signal
1d ago
1 directory member surfaced this signal.
1 expert
1 community
1 sources clustered
“An interesting paper. AI-Native Firms "More broadly, our findings suggest that AI may not simply make existing organizations more efficient—it may change what organizations look like and do" www.hbs.edu/ris/Publicat...”
sliding window beats linear attention
4 experts across 3 network communities independently surfaced this.
4 experts
3 communities
1 sources clustered
Research & technical analysis
3 experts
Evidence, methods and technical implications.
“Huge thanks to my collaborators @RheaSukthanker, @CameronPashmina, and @Emy_Aze. Paper: https://arxiv.org/abs/2608.28444”
2 experts discussed this · 2 posts
Alexia Jolicoeur-Martineau: Simple beats complicated: We show that switching to a sliding-window attention mask with attention sinks (at no cost) beats linear attention post-training. Huge thanks to my collaborators Rhea Sukt…
Miguel Alonso Jr.: Simple beats complicated: We show that switching to a sliding-window attention mask with attention sinks (at no cost) beats linear attention post-training. Huge thanks to my collaborators Rhea Sukt…
James MacGlashan: Makes me wonder if you can remove the sinks if you use softmax-1?
Open the full discussion →
AI agent civilizations essay
7 experts across 2 network communities independently surfaced this.
7 experts
2 communities
1 sources clustered
“Someone should create an AI movie of this incident. www.dwarkesh.com/p/openai-hug...”
“https://www.dwarkesh.com/p/openai-huggingface”
2 experts discussed this · 3 posts
Grace: Plot twist: they were destroyed by another, aligned agent swarm
Grace: https://www.dwarkesh.com/p/openai-huggingface
Open the full discussion →
Established
AI field signal
Signal
3d ago
3 experts across 2 network communities independently surfaced this.
3 experts
2 communities
1 sources clustered
“This is interesting but it would its 2/3 games, so we will have to wait to see if this becomes a pattern www.kedglobal.com/artificial-i...”
“Humans are back! Shin Jin-seo, the world's top-ranked Go player, on Tuesday completed a dramatic comeback against the world’s premier artificial intelligence Go engine, KataGo, claiming a historic human victory over AI. www.kedglobal.com/artificial-i...”
Atlas world model spatial intelligence release
2 directory members surfaced this signal.
2 experts
2 communities
1 sources clustered
“World Labs' Atlas A multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. www.worldlabs.ai/blog/atlas”
latent reasoning recurrent depth test-time compute paper
2 directory members surfaced this signal.
2 experts
2 communities
1 sources clustered
“OpenAI’s Astra may be using Recurrent Depth as outlined in this paper: Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach (arxiv.org/abs/2502.05171)”
Developing
AI field signal
Resource
1d ago
VANTAIRE AI video content
1 directory member surfaced this signal.
1 expert
1 community
1 sources clustered
“This is the best AI short video, about 30 minute long, I have seen so far. Mostly because it is edited so very well. A nicely done. www.youtube.com/watch?v=XxrY...”
GLM 5.3 Flash GGUF model release
1 directory member surfaced this signal.
1 expert
1 community
1 sources clustered
Anthropic trains misaligned reward seeker
3 directory members surfaced this signal.
3 experts
1 community
1 sources clustered
“This is entertaining reading. Anthropic's Training a Misaligned Reward Seeker They find that when reward hacking is reinforced during training, the model can pursue rewards by any means available to satisfy a grader. alignment.anthropic.com/2026/reward-...”
“at last, we have trained the misaligned reward hacking model from the cautionary sci-fi tale don’t train the misaligned reward hacking model alignment.anthropic.com/2026/reward-...”
Ling 3.0 Flash Fin model release
1 directory member surfaced this signal.
1 expert
1 community
1 sources clustered