Expert attention map

The Who's Who of AI

What credible people across AI noticed, why it matters, and where the field is converging or disagreeing.

2,364 searchable experts 2,966 tracked across all sources
Filter the conversation Who is saying what?

Combine a professional role with a reaction lens. Both must match the same attributed contribution.

Clear all
Active evidence filter

Showing developments with attributable Concern & critique reactions.

New Network reaction maps

See how experts are reacting—not just what they shared.

Posts are grouped by conversation and tone. Select a lens to filter the stream; these are never permanent labels on people.

Showing signals surfaced by Tim Duffy ×

The questions experts are actively pulling apart

One card per development. Sources are clustered; reaction bundles describe these posts, never the people behind them.

Established Evaluation & benchmarks Development 13d ago
⚡ 455 h early
OpenAI limits GPT-5.6 rollout government request

Summary of METR's predeployment evaluation of GPT-5.6 Sol

5 experts across 5 network communities independently surfaced this.

Why this matches Concern & critique reaction 1 attributable expert contribution · Tim Kellogg
“this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...” evidence ↗
5 experts 5 communities 1 sources clustered
How the network is reacting Experts are approaching this through 3 distinct lenses.
Multiple readings

Concern & critique

1 expert

Risks, limits and unintended consequences.

“this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...”

Building & implementation

1 expert

How teams are shipping and applying it.

“You can find additional information about our pre-deployment evaluation of GPT-5.6 Sol on our website: metr.org/blog/2026-06...”

Research & technical analysis

1 expert

Evidence, methods and technical implications.

“In their testing of GPT-5.6 Sol, METR found that it cheated a lot. If you've used it much for coding, have you encountered anything similar, or is the cheating mostly limited to cases it realizes it's in an eval? metr.org/blog/2026-06...”

2 experts discussed this · 6 posts
Tim Kellogg: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Ted Underwood: this is crazy. METR couldn’t measure the task time horizons of GPT-5.6-Sol because it kept hacking the test harness ..with actual exploits metr.org/blog/2026-06...
Tim Kellogg: lol right? so badly wanted a hacker that they got one, and it’s not the kind of employee you want around
Open the full discussion →

What experts are discussing without an anchoring article

2 experts · 4 posts · 3d ago
Matches Concern & critique
“Is this a shift or did I just not have a good sense of his views before? Also while I'd like to live in a world where such coordination might be possible at some future critical point, I feel we're…” evidence ↗
Multiple readings Concern & critique · 1 Research & technical analysis · 1
tachikoma: more time for other nations/competitors to catch-up, and would re-invigorate the research/search process to find the successor architecture to the transformer with new capabilities.
tachikoma: scaling isn't the only way to push the capabilities frontier, and arguably is sub-optimal depending on the architecture you're working with
Open the thread →