Dylan Hadfield-Menell

Why they matter

Researcher with public evidence across AI research, Safety & security.

AI signals
4
past 30d
Sources
4
distinct domains
Discussions
7
past 30d
Latest signal
10d ago
View every signal from Dylan Hadfield-Menell →
Assistant Prof of AI & Decision-Making @MIT EECS I run the Algorithmic Alignment Group (https://algorithmicalignment.csail.mit.edu/) in CSAIL. I work on value (mis)alignment in AI systems. https://people.csail.mit.edu/dhm/

Articles & links

↻ Dylan Hadfield-Menell reposted
Miranda Bogen @mbogen.bsky.social

📄 New research from my team at @cdt.org together with @dhadfieldmenell.bsky.social's lab at @csail.mit.edu finds that benign fine-tuning of foundation models leads to unpredictable safety drift — with big implications for AI governance. Report below, academic research here: ar…

Out of Tune: Fine-Tuning Foundation Models Leads to Unpredictable Safety Drift - Center for Democracy and Technology cdt.org View on Bluesky →

Recent commentary

I’m still pretty shocked that there hasn’t been more scrutiny or criticism of HuggingFace’s decision not to pursue legal action against OpenAI for the hack. Lots of people who (claim to) care about concentration of power just ignoring NVIDIA’s role, incentives, and power.

View on Bluesky · ♥ 5 ↻ 1 ↩ 1 · 8d ago

In Dylan Hadfield-Menell's orbit

Center = Dylan Hadfield-Menell. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Dylan Hadfield-Menell? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/dhadfieldmenell-bsky-social)