Nathan Lambert

Post-training researcher at Ai2, writes Interconnects

Why they matter

Post-training researcher at Ai2, writes Interconnects with public evidence across Agents & robotics, AI research.

AI signals
20
past 30d
Sources
10
distinct domains
Discussions
6
past 30d
Latest signal
14h ago
View every signal from Nathan Lambert →
A LLN - large language Nathan - (RL, RLHF, society, robotics), athlete, yogi, chef Writes http://interconnects.ai At Ai2 via HuggingFace, Berkeley, and normal places

Articles & links

Insane numbers for opus 5, the power of faster iteration speed + scaled RL (Fable too big to RL as well, yet). And on safeguards "Based on our testing, we expect the classifiers to intervene around 85% less often than they do for Fable 5.". www.anthropic.com/news/claude-...

Introducing Claude Opus 5 anthropic.com
AI Weekly's analysis
  • Anthropic launched Claude Opus 5 on July 24, 2026 at $5 per million input tokens, matching Opus 4.8's rate.
  • On Frontier-Bench v0.1 Opus 5 scored 43.3%, versus 18.7% for Opus 4.8 and 33.7% for Fable 5.
  • Opus 5 becomes the default on Claude Max but sits behind Mythos 5 on cybersecurity tasks, per Anthropic.
Read full analysis →
View on Bluesky · ♥ 78 ↻ 9 ↩ 2 · 9 from the directory shared this · 5d ago

Why I think Anthropic's uneven safety policies with the release of Claude Fable 5 undermine the broader AI community's cohesion and accelerate us to more uncertainty and risk in AI's near-term evolution. www.interconnects.ai/p/claude-fab...

Claude Fable 5 and new safety fables interconnects.ai
View on Bluesky · ♥ 40 ↻ 10 ↩ 0 · 4 from the directory shared this · 50d ago

6 months to live for open models Staring down the barrel of policy action that could make open models a permanent second class citizen. We need to a) win on the distillation issue and b) form a coalition www.interconnects.ai/p/6-months-t...

6 months to live for open models interconnects.ai
AI Weekly's analysis
  • Nathan Lambert predicts within roughly six months the White House could restrict open-weight models above the GPT 5.5, Claude Opus 4.8, or GLM-5.2 tier.
  • He frames Anthropic's letters to representatives about Chinese open models as regulatory capture, not a safety measure.
  • His proposed off-ramp is for Microsoft or Meta to ship a frontier open-weight model before an executive order lands.
Read full analysis →
View on Bluesky · ♥ 81 ↻ 12 ↩ 4 · 3 from the directory shared this · 17d ago

My time at Ai2 / @ai2.bsky.social has come to an end. Ai2 is a wonderful place. The last 2.5+ years building Olmo, Tulu, and other projects will be one of the peaks of my entire career. www.interconnects.ai/p/farewell-ai2

Farewell Ai2 interconnects.ai
View on Bluesky · ♥ 96 ↻ 4 ↩ 4 · 2 from the directory shared this · 57d ago

If you're looking for the latest adoption data on open models in US v China v globally, we built a small dashboard with the big picture and per-org numbers. US's role is slowly growing, but still way behind China/Qwen. dashboard.interconnects.ai

Open Models Dashboard dashboard.interconnects.ai
AI Weekly's analysis
  • The Open Models Dashboard shows daily-updated Hugging Face downloads and derivatives of open-weight AI models across USA vs China vs EU.
  • The tracked list covers post-ChatGPT LLMs and VLMs released after Nov 30, 2022, with a >100K total downloads threshold and guard models excluded.
  • An original seven organizations cover 1,971 models through July 2025, with expanded coverage now spanning over forty additional orgs.
Read full analysis →
View on Bluesky · ♥ 44 ↻ 9 ↩ 2 · 2 from the directory shared this · 6d ago

If recent events with Kimi K3 have finally convinced you that you need to try and understand how the Chinese labs approach AI - and how it differs than the SF center of power - you should read my post from a few months ago: www.interconnects.ai/p/notes-from...

Notes from inside China's AI labs interconnects.ai
AI Weekly's analysis
  • Nathan Lambert visited Moonshot AI, Zhipu, Meituan, Xiaomi, 01.ai and Tsinghua in Beijing and surrounding regions to compare Chinese and US lab culture.
  • Every Chinese lab he visited described Nvidia compute access as the primary bottleneck limiting their progress, not talent or data.
  • Chinese developers reportedly use Claude widely despite restrictions, and DeepSeek is credited internally with the best research taste in execution.
Read full analysis →
View on Bluesky · ♥ 81 ↻ 9 ↩ 1 · 2 from the directory shared this · 10d ago

Welcome to the AGI era of AI governance It's a one-way door and we weren't ready for it. Especially the open-source community celebrating Anthropic's downfall isn't ready for when it's their turn in court. www.interconnects.ai/p/welcome-to...

Welcome to the AGI era of AI governance interconnects.ai
View on Bluesky · ♥ 34 ↻ 2 ↩ 1 · 2 from the directory shared this · 45d ago

Recent commentary

My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me finding time to study and document the fundamentals on nights and weekends since 2024.

View on Bluesky · ♥ 172 ↻ 21 ↩ 7 · 8d ago

Thinking Machines just released with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model, and omni input. A bit behind GLM 5.2 on agentic benches, and Kimi K 2.6 on multi modal Super exciting!!

View on Bluesky · ♥ 139 ↻ 16 ↩ 5 · 14d ago

I think what is pretty clear is that the Chinese labs are far more capital efficient. In a world where scaling labs are intelligence is proportional to effective capital (buys compute, data, & talent) that may be the greatest strength your AI industry could ever have.

View on Bluesky · ♥ 44 ↻ 5 ↩ 2 · 11d ago

Anthropic's political pressure on distillation is regulatory capture and most of the employees are blind to it under their veil of safety. Or their paycheck helped them buy into safety, is only human nature, I don't even fault them that much.

View on Bluesky · ♥ 48 ↻ 3 ↩ 2 · 32d ago

Being out of SF has lowered my information proximity but with the big upside of giving me space to cultivate my own beliefs and values around ai. We need more people zagging in AI, the monoculture just helps the incumbents win at this point.

View on Bluesky · ♥ 46 ↻ 3 ↩ 1 · 73d ago

Kimi K3 with more likes than downloads on HuggingFace is definitely showing us a glimpse of the future on open models. It's way less about individual access, and more of a distributed platform layer for companies.

View on Bluesky · ♥ 45 ↻ 2 ↩ 1 · 2d ago

Making talks with AI agents is awesome. I just told Fable to make a slide with real data on the KL distance from one of our reference Olmo 2 models and it made this with the wandb api (I edited text slightly).

View on Bluesky · ♥ 31 ↻ 2 ↩ 2 · 2d ago

Claude Fable is another big step in being able to make nice lectures based on existing educational content. Much better than Opus. GPT 5.6 is still very far off here. Is a good example of where Claude Code being a bit easier to work across different knowledge work tasks.

View on Bluesky · ♥ 30 ↻ 1 ↩ 2 · 17d ago

It's been a great effort by the early and growing American open-model labs since last June to put the US much more back on the map. We were getting totally owned last June. Nvidia, Ai2, Arcee, Gemma, GPT-OSS and a few others will be seen as saving American open AI.

View on Bluesky · ♥ 31 ↻ 0 ↩ 0 · 55d ago

The real comparison to Moore's law for AI isn't scaling laws, but rather the intelligence efficiency that we gain year-over-year using models to get better at training models.

View on Bluesky · ♥ 21 ↻ 2 ↩ 1 · 5d ago

In Nathan Lambert's orbit

Center = Nathan Lambert. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.