Jonathan Stray

Why they matter

Researcher with public evidence across AI research, Culture, work & education.

AI signals
2
past 30d
Sources
2
distinct domains
Discussões
17
past 30d
Latest signal
8d ago
View every signal from Jonathan Stray →
Knowing things is a solved problem. Getting along is not. Working on AI, media, and inter-group conflict @CHAI_Berkeley. Got here from computational journalism.

Articles & links

Hi! Ready your paper, very interesting. Thought about it, and I'm not sure I find the computational complexity proof convincing because it bites only for learning all possible distributions D -- human-like D is plausibly easier. I found this, makes argument more rigorously. ar…

Barriers to Complexity-Theoretic Proofs that "AGI" Using Machine Learning is Impossible arxiv.org
AI Weekly's analysis
  • Guerzhoy argues van Rooij et al.'s 2024 proof that AGI via machine learning is intractable rests on an unjustified assumption about data distributions.
  • The same proof structure, applied consistently, would show ImageNet classification is intractable, yet that task demonstrably works.
  • Three barriers block any such proof: defining human-like behavior precisely, accounting for inductive bias, and specifying relevant data subsets.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 1 · 2 from the directory shared this · 54d ago

There's a long list of content-based signals we might want to rank on. We are limited by the classifiers we actually have. BUT we are working on LLM scoring with arbitrary prompts! The challenge is performance, but we have a cunning plan docs.google.com/document/d/1...

LLM-based scoring for GreenEarth docs.google.com
View on Bluesky · ♥ 3 ↻ 1 ↩ 2 · 32d ago

While there are many criticisms of LLMs I agree with, and correspondingly many lines of attack... this one I feel is a dead end. For two reasons. 1) there is deep theory showing that replicating the output of a system implies building a causal model of the system. E.g. arxiv.o…

Robust agents learn causal world models arxiv.org
View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 8d ago

Well, there are examples in the paper and the dataset is open source! github.com/humanCompati...

github.com
View on Bluesky · ♥ 0 ↻ 0 ↩ 1 · 52d ago

Recent commentary

Humanity's ability to know, reason, judge, and act well is the foundation of science, democracy, crisis response, & management of AI itself. AI poses serious risks to that foundation. New paper on epistemic risks by 30 experts calls for attention and proposes solutions. Link in thread.

View on Bluesky · ♥ 34 ↻ 19 ↩ 3 · 70d ago

Fusion power is gonna go a lot like AI. Empty promises for decades, then suddenly here faster than anyone can adjust to.

View on Bluesky · ♥ 11 ↻ 0 ↩ 1 · 32d ago

Me: I’m working on AI and human conflict BigLab employee: Cool, like multi-agent games? Me: No, like human conflict BigLab: Cool, so like simulating people’s responses? Me: no, like actual human field experiments 🤷‍♂️

View on Bluesky · ♥ 8 ↻ 0 ↩ 1 · 10d ago

What could it mean for an AI to be "politically neutral”? And can we measure it? New paper + dataset. We propose a definition that applies to any type of conflict on any topic: a neutral response should maximize approval on both sides of an issue, while keeping that approval balanced. 1/🧵

View on Bluesky · ♥ 4 ↻ 1 ↩ 2 · 73d ago

The AI models of today are the worst they will ever be. And yet, pretty much every "AI will never..." claim has now been shattered. I don't understand people who still bet against AI. How much more evidence do you need that these machines are going to be smarter than us in every way?

View on Bluesky · ♥ 4 ↻ 1 ↩ 1 · 26d ago

Seeing a flurry of evals and startups promising to test the mental health effects of AI. Literally all of them test what the model says in various conditions... none of them measure actual outcomes on actual people. A big gap, fixable with privacy-preserving experiments.

View on Bluesky · ♥ 5 ↻ 0 ↩ 1 · 84d ago

Is this a good logo for the GreenEarth feed? It's a healthier, user-controllable, open-source, LLM-powered, transparent feed we're building -- now in alpha testing. Try it? Tell us what you think! bsky.app/profile/did:...

View on Bluesky · ♥ 3 ↻ 1 ↩ 0 · 42d ago

I want to make sure AI doesn't incite human conflict. It's sometimes hard to explain what I do, but that's the core of it -- it won't happen automatically. And we're making progress! Both theoretically, and in field experiments that test how AI alters human relationships.

View on Bluesky · ♥ 2 ↻ 1 ↩ 0 · 47d ago

AI safety typically assumes one well-meaning user. I'm working on the case where two of them are at war.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 64d ago

Hello I am in Montreal for the week! Anyone interesting in AI safety I should meet here?

View on Bluesky · ♥ 1 ↻ 0 ↩ 0 · 49d ago

In Jonathan Stray's orbit

Center = Jonathan Stray. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.