Jonathan Stray

Why they matter

Researcher with public evidence across AI research, Culture, work & education.

AI signals
4
past 30d
Sources
3
distinct domains
Discussions
11
past 30d
Latest signal
9d ago
View every signal from Jonathan Stray →
Knowing things is a solved problem. Getting along is not. Working on AI, media, and inter-group conflict @CHAI_Berkeley. Got here from computational journalism.

Articles & links

This is a much better response to the same sort of AI threat to practicing scientists arxiv.org/abs/2602.10181

Why do we do astrophysics? arxiv.org
AI Weekly's analysis →
  • David W. Hogg says LLMs are beginning to design, execute, write up and referee scientific projects on the data-science side of astrophysics.
  • His proposed 'points of agreement' emphasise novelty, people-centrism, trust, and the lack of clinical value in astrophysics.
  • He argues strongly against both 'let-them-cook' and 'ban-and-punish' policies, warning good moderate rules will be hard to develop.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 3 from the directory shared this · 15d ago

Hi! Ready your paper, very interesting. Thought about it, and I'm not sure I find the computational complexity proof convincing because it bites only for learning all possible distributions D -- human-like D is plausibly easier. I found this, makes argument more rigorously. ar…

Barriers to Complexity-Theoretic Proofs that "AGI" Using Machine Learning is Impossible arxiv.org
AI Weekly's analysis →
  • Guerzhoy argues van Rooij et al.'s 2024 proof that AGI via machine learning is intractable rests on an unjustified assumption about data distributions.
  • The same proof structure, applied consistently, would show ImageNet classification is intractable, yet that task demonstrably works.
  • Three barriers block any such proof: defining human-like behavior precisely, accounting for inductive bias, and specifying relevant data subsets.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 1 · 2 from the directory shared this · 95d ago

There's a long list of content-based signals we might want to rank on. We are limited by the classifiers we actually have. BUT we are working on LLM scoring with arbitrary prompts! The challenge is performance, but we have a cunning plan docs.google.com/document/d/1...

LLM-based scoring for GreenEarth docs.google.com
View on Bluesky · ♥ 3 ↻ 1 ↩ 2 · 73d ago

While there are many criticisms of LLMs I agree with, and correspondingly many lines of attack... this one I feel is a dead end. For two reasons. 1) there is deep theory showing that replicating the output of a system implies building a causal model of the system. E.g. arxiv.o…

Robust agents learn causal world models arxiv.org
View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 49d ago

Well, there are examples in the paper and the dataset is open source! github.com/humanCompati...

github.com
View on Bluesky · ♥ 0 ↻ 0 ↩ 1 · 94d ago

There we go, another major LLM does some unintended hacking. This suggests that the problem is, in fact, hard to solve. www.reuters.com/business/gem...

reuters.com
View on Bluesky · ♥ 0 ↻ 1 ↩ 1 · 9d ago

There are concerns in the mathematician’s open letter I share, but on the whole I can’t endorse it. This satirical response kinda sums it up x.com/mbeisen/stat...

x.com
View on Bluesky · ♥ 2 ↻ 1 ↩ 1 · 15d ago

Recent commentary

Humanity's ability to know, reason, judge, and act well is the foundation of science, democracy, crisis response, & management of AI itself. AI poses serious risks to that foundation. New paper on epistemic risks by 30 experts calls for attention and proposes solutions. Link in thread.

View on Bluesky · ♥ 34 ↻ 19 ↩ 3 · 111d ago

Fusion power is gonna go a lot like AI. Empty promises for decades, then suddenly here faster than anyone can adjust to.

View on Bluesky · ♥ 11 ↻ 0 ↩ 1 · 73d ago

Me: I’m working on AI and human conflict BigLab employee: Cool, like multi-agent games? Me: No, like human conflict BigLab: Cool, so like simulating people’s responses? Me: no, like actual human field experiments 🤷‍♂️

View on Bluesky · ♥ 8 ↻ 0 ↩ 1 · 51d ago

What could it mean for an AI to be "politically neutral”? And can we measure it? New paper + dataset. We propose a definition that applies to any type of conflict on any topic: a neutral response should maximize approval on both sides of an issue, while keeping that approval balanced. 1/🧵

View on Bluesky · ♥ 4 ↻ 1 ↩ 2 · 114d ago

The AI models of today are the worst they will ever be. And yet, pretty much every "AI will never..." claim has now been shattered. I don't understand people who still bet against AI. How much more evidence do you need that these machines are going to be smarter than us in every way?

View on Bluesky · ♥ 4 ↻ 1 ↩ 1 · 68d ago

Seeing a flurry of evals and startups promising to test the mental health effects of AI. Literally all of them test what the model says in various conditions... none of them measure actual outcomes on actual people. A big gap, fixable with privacy-preserving experiments.

View on Bluesky · ♥ 5 ↻ 0 ↩ 1 · 125d ago

Is this a good logo for the GreenEarth feed? It's a healthier, user-controllable, open-source, LLM-powered, transparent feed we're building -- now in alpha testing. Try it? Tell us what you think! bsky.app/profile/did:...

View on Bluesky · ♥ 3 ↻ 1 ↩ 0 · 83d ago

I want to make sure AI doesn't incite human conflict. It's sometimes hard to explain what I do, but that's the core of it -- it won't happen automatically. And we're making progress! Both theoretically, and in field experiments that test how AI alters human relationships.

View on Bluesky · ♥ 2 ↻ 1 ↩ 0 · 89d ago

I have never expressed a p(doom), and I think @randomwalker.bsky.social is basically right that we have little basis for quantitative estimates of AI-induced catastrophic risk. OTH, no one has presented a knock-down argument that *any* of the various AI disaster scenarios are impossible.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 5d ago

Writing code and doing related research tasks with Fable has caused me to move up my estimated date for recursive self-improvement (RSI), which is a more precisely defined term than the nebulous "AGI." Anyway, I think we'll see intelligent machines self-improving by late next year.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 10d ago

In Jonathan Stray's orbit

Center = Jonathan Stray. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Jonathan Stray? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/jonathanstray-bsky-social)