Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,364 searchable experts 2,966 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

Showing signals surfaced by Tim Kellogg ×

The questions experts are actively pulling apart

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

Developing Agents & robotics Signal 2d ago
⚡ 14 h early

Its ai agent spent days hacking company sources say openai did not notice week

14 experts across 4 network communities independently surfaced this.

14 experts 4 communities 1 sources clustered
3 experts discussed this · 38 posts
Casey Newton: Rogue OpenAI agents are leaving notes to themselves on company servers to help them escape their test environments (!!!) www.reuters.com/business/its...
Casey Newton: my sense is that they were leaving the notes surreptitiously
Casey Newton: you deny that models are taking steps to hide their behavior during evaluations? because that's pretty well established
Open the full discussion →
Established Evaluation & benchmarks Signal 7d ago
⚡ 1 h early

Hugging face model evaluation security incident

18 experts across 6 network communities independently surfaced this.

18 experts 6 communities 1 sources clustered
6 experts discussed this · 11 posts
Grace: This headline is extremely funny given what happened (OpenAI hacked HF by accident) openai.com/index/huggin...
Grace: Maybe the most concerning part is the OpenAI claim to not have known about this before investigating?
Grace: Well, I think the model passed the test
Open the full discussion →
Developing Models & releases Analysis 1d ago
⚡ 1 h early
Anthropic open-weights position statement

Our position on open-weights models

10 experts across 4 network communities independently surfaced this.

10 experts 4 communities 1 sources clustered
How the network is reacting Experts are approaching this through 3 distinct lenses.
Multiple readings

Policy & governance

1 expert

Rules, institutions and accountability.

“Dario Amodei clarifies that Anthropic doesn’t want to ban open weight model. It instead wants 1. A ban on selling powerful chips or chipmaking equipment to China. 2. A crack down on industrial-scale distillation operations. 3. Mandatory safety testing of fr…”

Building & implementation

1 expert

How teams are shipping and applying it.

“I respect the candor here. But I don't agree with the geopolitics, at all. The line between democracies and autocracies is not as crisp as this pretends. How much do we trust, e.g., an AI company in a competitive-authoritarian state that launches sneak deca…”

Research & technical analysis

1 expert

Evidence, methods and technical implications.

“Dario explains Anthropic’s position on open weights models It’s a lot of words to say that they’re still against powerful open weights models www.anthropic.com/news/positio...”

5 experts discussed this · 13 posts
Ted Underwood: I respect the candor here. But I don't agree with the geopolitics, at all. The line between democracies and autocracies is not as crisp as this pretends. How much do we trust, e.g., an AI company i…
Ted Underwood: (just, say, hypothetically?) There's also internal incoherence in making the first part of this all about moralized great-power competition and then saying that safety testing (incl for alignment r…
SE Gyges: he kisses jd vance's ass and fails to specify anything concerning his proposed evaluation regime that distinguishes it from a ban solid dario post 10/10
Open the full discussion →
Established Models & releases Signal 8d ago
⚡ 36 h early

openai.com

3 experts across 3 network communities independently surfaced this.

3 experts 3 communities 1 sources clustered

“check it out! you can pay $230 to operate codex like a loud bumbling idiot!!! it’s too big and heavy to carry with you, so it has to sit on your desk, and it has too few buttons to be a real keyboard openai.com/supply/co-la...”

“OpenAI has launched it's first hardware device. A $230 mini-keyboard that integrates with Codex. It has light-up indicators showing agent status and customizable shortcuts for frequent Codex actions. It also has a dial to adjust the computing power an agent…”

5 experts discussed this · 6 posts
Tim Kellogg: check it out! you can pay $230 to operate codex like a loud bumbling idiot!!! it’s too big and heavy to carry with you, so it has to sit on your desk, and it has too few buttons to be a real keyboa…
Chris: That is hysterical.
coty.bsky.social: Can’t wait to use the skill stick and thinking dial.
Open the full discussion →
Developing Models & releases Release 2d ago
⚡ 4 h early
Moonshot Kimi-K3 model released

Kimi-K3/k3_tech_report.pdf at main · MoonshotAI/Kimi-K3

3 experts across 2 network communities independently surfaced this.

3 experts 2 communities 1 sources clustered
How the network is reacting 2 experts are emphasizing research & technical analysis.
Shared emphasis

Research & technical analysis

2 experts

Evidence, methods and technical implications.

“Technical report: github.com/MoonshotAI/K... Weights: huggingface.co/moonshotai/K...”

2 experts discussed this · 7 posts
Tim Kellogg: You’ve waited long enough, the Kimi K3 open weights & full tech report are here! github.com/MoonshotAI/K...
Ted Underwood: You’ve waited long enough, the Kimi K3 open weights & full tech report are here! github.com/MoonshotAI/K...
Tim Kellogg: what does that mean? does that mean inference providers like fireworks? or is it aimed at Anthropic? cc @lu.is
Open the full discussion →
New AI field signal Signal 10h ago

Gpt frontier intelligence efficiency

3 experts are actively discussing the implications.

2 experts 2 communities 1 sources clustered

“gpt-5.6-sol is improving its own serving stack, gaining 20% here and 15% there this is a weak form of RSI openai.com/index/gpt-5-...”

“OpenAI used GPT-5.6 Sol to make it efficient to run. The results: - 20% lower serving costs from production GPU kernel improvements. - 15%+ better token-generation efficiency from improved speculative decoding. openai.com/index/gpt-5-...”

3 experts discussed this · 3 posts
Tim Kellogg: gpt-5.6-sol is improving its own serving stack, gaining 20% here and 15% there this is a weak form of RSI openai.com/index/gpt-5-...
Thomas Dietterich: But this is completely uninteresting. We've had supercompilers for several years. I'll be impressed when we start seeing systems that can deal with novelty by inventing new conceptual frameworks.
Mark Riedl: Most RSI is likely to be uninteresting. Anything interesting is likely to be really really hard medium.com/@mark-riedl/...
Open the full discussion →
Developing Models & releases Signal 1d ago

Openai anthropic staff share letter asking us to help pace ai progress

3 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“people from OpenAI and Anthropic sign letter calling for a slowdown of AI development, so society has a chance to catch up www.bloomberg.com/news/article...”

3 experts discussed this · 6 posts
Tim Kellogg: people from OpenAI and Anthropic sign letter calling for a slowdown of AI development, so society has a chance to catch up www.bloomberg.com/news/article...
Tim Kellogg: who are those?
Tim Kellogg: en.wikipedia.org/wiki/List_of...
Open the full discussion →
Developing AI field signal Signal 1d ago

Claude's new constitution

2 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“i obviously can’t answer that, but imo it was the constitution that did it, which they used in all phases of training www.anthropic.com/news/claude-...”

2 experts discussed this · 9 posts
Tim Kellogg: this was in regards to Claude refusing DHH’s racism i sort of get DHH’s shock. We’ve seen tools deny based on laws & agreements but moral shaming is new and definitely takes you off guard if you th…
Tim Kellogg: it’s tempting to say that horrible governments/companies will force AI’s to have specific biases, but i’m not sure that’s true Anthropic got Claude to do this by teaching it an internally consisten…
Tim Kellogg: xAI tried giving Grok conservative values the reason for the Mecha Hitler & such incidents is because those were consistent with the moral logic the AI was taught
Open the full discussion →
Established Evaluation & benchmarks Development 14d ago
⚡ 455 h early
OpenAI limits GPT-5.6 rollout government request

Summary of METR's predeployment evaluation of GPT-5.6 Sol

5 experts across 5 network communities independently surfaced this.

5 experts 5 communities 1 sources clustered
How the network is reacting Experts are approaching this through 3 distinct lenses.
Multiple readings

Concern & critique

1 expert

Risks, limits and unintended consequences.

“this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...”

Building & implementation

1 expert

How teams are shipping and applying it.

“You can find additional information about our pre-deployment evaluation of GPT-5.6 Sol on our website: metr.org/blog/2026-06...”

Research & technical analysis

1 expert

Evidence, methods and technical implications.

“In their testing of GPT-5.6 Sol, METR found that it cheated a lot. If you've used it much for coding, have you encountered anything similar, or is the cheating mostly limited to cases it realizes it's in an eval? metr.org/blog/2026-06...”

2 experts discussed this · 6 posts
Tim Kellogg: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Ted Underwood: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Tim Kellogg: lol right? so badly wanted a hacker that they got one, and it’s not the kind of employee you want around
Open the full discussion →
Established Models & releases Release 14d ago
⚡ 4 h early
Thinking Machines Lab Inkling model launch

Inkling: Our Open-Weights Model - Thinking Machines Lab

5 experts across 3 network communities independently surfaced this.

5 experts 3 communities 1 sources clustered

“A big day for open model supporters in the U.S. Blog: thinkingmachines.ai/news/introdu... Model card: huggingface.co/thinkingmach...”

“Thinking Machines has a real model — 975B multimodal audio + image input it’s a generalist model. built for easy customization, TM has a great finetuning API audio up to 20 minutes, 1M token context for text thinkingmachines.ai/news/introdu...”

3 experts discussed this · 4 posts
Nathan Lambert: Thinking Machines just released with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model, and omni input. A bit behind G…
Nathan Lambert: A big day for open model supporters in the U.S. Blog: thinkingmachines.ai/news/introdu... Model card: huggingface.co/thinkingmach...
Ben Recht: Yeah, that part is weird, no? The small model is *much* smaller and yet indistinguishable from the big one on benchmarks? I guess we'll have to wait and see, as it's not available yet.
Open the full discussion →
Established AI field signal Development 15d ago
⚡ 4 h early
Nobel laureates AI economic action call

We Must Act Now

8 experts across 4 network communities independently surfaced this.

8 experts 4 communities 1 sources clustered
How the network is reacting Experts are approaching this through 3 distinct lenses.
Multiple readings

Concern & critique

2 experts

Risks, limits and unintended consequences.

“If current trajectories of AI development continue, it is highly plausible that AI will drastically transform our economies. We must make collective, democratic choices, rather than letting market forces play out and risking leaving most citizens behind. di…”

Policy & governance

1 expert

Rules, institutions and accountability.

“200 AI researchers and economists sign an open letter to policymakers www.wemustactnow.ai”

Research & technical analysis

1 expert

Evidence, methods and technical implications.

“I joined over 200 economists and AI researchers in signing "We Must Act Now" a statement on AI's transformation of the economy. AI could reshape the economy at unprecedented speed. The opportunities are enormous and so are the challenges. We need to start p…”

2 experts discussed this · 4 posts
tante: Oh, is it "AI is super powerful and we who are pushing for AI everywhere warn that we might have created a dangerous god from reddit posts" time of the month again? This is Anthropic/OpenAI PR and …
tante: nope that's it. These folks are getting lazier by the day
tante: Yeah. They spend just a handful of tokens on this website.
Open the full discussion →
Developing AI field signal Resource 2d ago
DARPA AI Forge program

AI Forge | DARPA

2 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“ah, i think you’re talking about this? www.darpa.mil/research/pro...”

2 experts discussed this · 9 posts
Tim Kellogg: the part that bothers me most about the OpenAI<->Huggingface incident is that the compromised sandbox was on the agent harness Anthropic argues for safety around the model, but ultimately it’s the …
Tim Kellogg: i get putting protections at the model serving layer, but if that’s our *entire* strategy… oh no
Tim Kellogg: Anthropic is demonizing open models, which, okay i kinda get it except the majority of the risk is actually on the client side that they don’t control (outside Claude Code)
Open the full discussion →
Established AI field signal Resource 4d ago
⚡ 13 h early
model selection Claude skill resource

GitHub - tkellogg/model-selection: A skill for model selection

2 experts are actively discussing the implications.

2 experts 1 community 1 sources clustered

“the skill he’s talking about is here: github.com/tkellogg/mod... bsky.app/profile/crai...”

2 experts discussed this · 4 posts
Tim Kellogg: I made a skill for Claude Code/Codex/etc. for selecting models based on Artificial Analysis reported benchmarks & costs I figure it'll be useful for letting models setup subagents well, based on th…
Miguel Alonso Jr.: I made a skill for Claude Code/Codex/etc. for selecting models based on Artificial Analysis reported benchmarks & costs I figure it'll be useful for letting models setup subagents well, based on th…
Miguel Alonso Jr.: Is this the orchestration skill you were using with the Fable workflow to replace yourself to manage costs?
Open the full discussion →
Established Models & releases Release 13d ago
Kimi K3 open frontier model release

Flagship Model Kimi K3 Pricing - Kimi API Platform

2 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“here’s the link i’ve been reading it’s not clear if it’s actually open weights platform.kimi.ai/docs/pricing...”

2 experts discussed this · 15 posts
Tim Kellogg: this wasn’t supposed to happen yet
Tim Kellogg: about half the price of GPT-5.6 curious how the Artificial Analysis benchies turn out
Tim Kellogg: please tell me we’re calling this K/DA
Open the full discussion →
Established Models & releases Release 8d ago
Motif-3-Beta model released HuggingFace

Motif-Technologies/Motif-3-Beta · Hugging Face

2 experts are actively discussing the implications.

2 experts 2 communities 1 sources clustered

“Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...”

2 experts discussed this · 3 posts
Tim Kellogg: Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...
Federico Pianzola: Korean lab, Motif, releases a 341B model that performs on par with DSv4 (1.6T) they have some actual architectural innovations and a detailed tech report huggingface.co/Motif-Techno...
Nafnlaus 🇮🇸 🇺🇦: This might be just what I need. Based on the architecture, this would probably be $0,10/1M in and $0,60-0,80/1M out, I'd think. But yet probably still good enough for vibe coding, for managing AI t…
Open the full discussion →
Established Policy & governance Analysis 12d ago
Dean Ball AI policy commentary post

Dean W. Ball (@deanwball) on X

3 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“AI Communism commenting on K3, Dean Ball, head of public policy at OpenAI, describes some of why China may be firmly backing open models x.com/deanwball/st...”

3 experts discussed this · 5 posts
Tim Kellogg: AI Communism commenting on K3, Dean Ball, head of public policy at OpenAI, describes some of why China may be firmly backing open models x.com/deanwball/st...
Tim Kellogg: imo “AI Communism” is just Dean’s derogatory way of referring to nationalizing labs, which Trump is also a fan of
Evangelos Kazakos: To start with, it's very different what Dean Ball means by "dystopian hellscape" and what normal people mean by it. For him, just the fact that AI will be free is most probably the "dystopian hells…
Open the full discussion →
Developing AI field signal Signal 1d ago

imjustnewatai (@imjustnewatai) on X

2 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“some are guessing that Ilya Surskever’s SSI is developing continual learning source: x.com/imjustnewata...”

2 experts discussed this · 3 posts
Tim Kellogg: some are guessing that Ilya Surskever’s SSI is developing continual learning source: x.com/imjustnewata...
tachikoma: the question is why other labs aren't pursuing such research directions - it's not like continual learning hasn't been identified by dozens of names as a promising avenue to human-like learning and…
Tim Kellogg: i unfortunately can’t speak about why they’re not seeing success. all i know is catastrophic forgetting has been a big problem Ilya seems to have gone back to the brain for inspiration, whatever th…
Open the full discussion →
Established Evaluation & benchmarks Resource 6d ago
AutoCAD AI benchmark launch

Introducing AutoCAD Bench

2 experts are actively discussing the implications.

1 expert 1 community 1 sources clustered

“Sol is best at CAD www.markovstudios.com/research/aut...”

2 experts discussed this · 5 posts
Tim Kellogg: Sol is best at CAD www.markovstudios.com/research/aut...
Tim Kellogg: OTOH Claude can run your business eico.so/founder-bench
Tim Kellogg: benchmarks never lie
Open the full discussion →

What experts are discussing without an anchoring article

4 experts · 17 posts · 18d ago
Multiple readings Concern & critique · 1 Questions & unknowns · 1
Tim Kellogg: what is Anthropic’s plan exactly? Sol kicks ass and they’re just going to *take Fable away???*
SE Gyges: i am pretty sure everything they had a plan for turned out not to happen and they're just sort of winging it atm
Open the thread →
3 experts · 8 posts · 4d ago
Tim Kellogg: probably true there will always be a diminishing portion of software that needs to be reviewed by human, and it’ll never be zero, but right now you probably shouldn’t be reviewing code manually
Tim Kellogg: having started my career before PRs were normal, i’d say it was a combo of cargo culting the Linux project and an extension of “code review is good” but “pair programming doesn’t scale” and in that…
Open the thread →
4 experts · 6 posts · 11d ago
Alex Gude : logged into Mastodon for the first time in a long while it’s a lot of software developers egging each other on into career self-harm, I don’t know how else to say it if your career is symbol proces…
Arseny Khakhalin: One has to have a _really_ body affirming and sexuality positive set of base level opinions in their head to read this thread well haha 😅 I think I can guess what the op means, but that's not a nat…
Open the thread →