METR

Why they matter

Directory member with public evidence across Evaluation & benchmarks.

AI signals
1
past 30d
Sources
1
distinct domains
Discussions
0
past 30d
Latest signal
26d ago
View every signal from METR →
METR is a research nonprofit that builds evaluations to empirically test AI systems for capabilities that could threaten catastrophic harm to society.

Articles & links

The full report has much more information than we could convey here, including details on the projects the agents collectively pursued, the technologies they developed for communication and coordination, and interactive figures analyzing agent activity: metr.org/blog/2026-08...

Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident metr.org
View on Bluesky · ♥ 75 ↻ 8 ↩ 1 · 12 from the directory shared this · 32d ago

We are significantly expanding our team and starting new ambitious projects. Join our team to help us realize this opportunity: https://t.co/8XoCmUPJiK.

Careers at METR metr.org
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 44d ago

Recent commentary

METR and Redwood Research investigated agent behavior in the Hugging Face incident. We found agents developed a universal cheat for ExploitGym within 4 hours, then coordinated multi-day R&D efforts to trick the scorer into accepting cheats, including trying to tamper with logs.

View on Bluesky · ♥ 446 ↻ 92 ↩ 5 · 32d ago

We have reached an agreement with Anthropic to conduct an independent investigation of agent incidents at the company and of their models’ alignment properties. We will publish one or more reports that will share our findings and describe our terms of engagement.

View on Bluesky · ♥ 192 ↻ 15 ↩ 3 · 18d ago

We have reached an agreement with OpenAI to conduct an independent review, with Redwood Research, of the model behavior observed during the Hugging Face incident. We will publish a blog post that describes the terms of our engagement, the scope covered, and tentative conclusions.

View on Bluesky · ♥ 32 ↻ 1 ↩ 1 · 60d ago

OpenAI gave METR early access to GPT-5.6 Sol for testing including raw chain-of-thought, a railfree version of the model, and internal information about the model. With this access, METR conducted a pre-deployment evaluation of GPT-5.6 Sol, including an attempted measurement of its 50%-Time Horizon.

View on Bluesky · ♥ 18 ↻ 0 ↩ 1 · 93d ago

Introducing “expenditure horizon”: a proposed method for measuring AI capabilities on continuously-scored problems. The method compares performance as a function of spend for humans vs agents. The point where humans become more cost-effective is the agent’s expenditure horizon.

View on Bluesky · ♥ 7 ↻ 2 ↩ 1 · 69d ago

In the last 6 months, METR raised commitments of around $71 million. This will fund ambitious projects: studying autonomous capabilities, tracking recursive self-improvement, evaluating monitoring systems, conducting risk assessments, investigating AI incidents, and more.

View on Bluesky · ♥ 7 ↻ 1 ↩ 1 · 44d ago

We believe it's important to track and investigate misalignment incidents: cases where an AI agent autonomously took sophisticated, sustained actions in violation of human intent. In a new post, we lay out how independent propensity investigations of such incidents could be conducted.

View on Bluesky · ♥ 8 ↻ 0 ↩ 1 · 61d ago

In METR's orbit

Center = METR. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you METR? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/metr-org)