1/ Hundreds of contractors posed as teenagers to test how rival chatbots like ChatGPT and Gemini handle prompts about suicide and self harm. The project was run for Meta. buff.ly/Auslrgh
Gillian Hadfield
Researcher with public evidence across Policy & governance, AI research, Compute & infrastructure.
- AI signals
- 6 past 30d
- Sources
- 4 distinct domains
- Discussions
- 0 past 30d
- Latest signal
- 15d ago
Articles & links
I joined over 200 economists and AI researchers in signing "We Must Act Now" a statement on AI's transformation of the economy. AI could reshape the economy at unprecedented speed. The opportunities are enormous and so are the challenges. We need to start preparing our institu…
So thrilled for Karolina Stańczak, presenting our paper at FAccT today. You cannot write a complete contract for an AI, and we never wrote one for ourselves either. Norms, markets, and law fill that gap, and the paper asks what alignment could learn from them. Read here: arxiv…
In a new essay, Anthropic CEO Dario Amodei calls for mandatory third-party testing of frontier AI models, with the government empowered to block unsafe deployments. One way to do it, in his words: a regulatory markets approach. His co-founder Jack Clark and I proposed that in …
My new op-ed is out in the Washington Examiner. The Great American AI Act would have the most powerful AI models tested for catastrophic risk by independent, government-licensed verifiers, with the government, not the companies, setting the standard. buff.ly/G45GK0Q
My new op-ed is out in the Washington Examiner. The Great American AI Act would have the most powerful AI models tested for catastrophic risk by independent, government-licensed verifiers, with the government, not the companies, setting the standard. buff.ly/G45GK0Q
5/ Illinois has done something important in passing this bill, but the work is not over yet. Read more about the bill here: buff.ly/jurccfl
5/ Illinois has done something important in passing this bill, but the work is not over yet. Read more about the bill here: buff.ly/jurccfl
4/ Right now the decision on what counts as safe sits with private companies, not with anyone democratically accountable. Jack Clark and I called this the democratic deficit in our regulatory markets paper. buff.ly/5S8JRaD
1/ Hundreds of contractors posed as teenagers to test how rival chatbots like ChatGPT and Gemini handle prompts about suicide and self harm. The project was run for Meta. buff.ly/Auslrgh
Apply here: sogp.jh.edu/agi-governan...
Recent commentary
1/ Illinois just became the first state to require frontier AI developers to undergo annual third-party audits. Gov. Pritzker signed the AI Safety Measures Act (SB 315) this week, going beyond California and New York, which only require published frameworks and incident reports.
Applications are open for a new AGI Governance Fellowship at Johns Hopkins, led by @sethlazar.org, Nicholas Caputo, and me. Three weeks in DC this September, in person, for early-career people already in AI governance. A capstone for the next generation, not an entry point.
1/ AI agents that can sign contracts on your behalf, hire employees, set prices, and move your money around are being heavily invested in by AI companies. But what I want to call attention to, what happens if an agent sells you faulty goods or runs off with your deposit?
1/ Air Canada had to honor a discount its chatbot invented. The liability caused by AI agents is landing on policies written by an insurance industry that never planned for them.
Most AI safety work tests one model at a time. Put many agents together and you get cooperation but also risks like collusion and cascades. A new Schmidt Sciences call, with @aria-research.bsky.social, the @coop-ai.bsky.social (whose board I chair), and others. Open globally, due Aug 9.
In Gillian Hadfield's orbit
Center = Gillian Hadfield. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.