John Horton

Why they matter

Tracked through public AI activity and peer connections inside the directory.

AI signals
4
past 30d
Sources
3
distinct domains
Discusiones
0
past 30d
Latest signal
5h ago
View every signal from John Horton →

Articles & links

Interesting paper from @maxchupilkin illustrating a kind of 'Volkswagen emissions test" effect where LLMs respond differently about war when being told then are being tested for alignment https://t.co/rMWObcyPZI 1/ https://t.co/g9rysSZk3p

Language models judge war differently when tested for alignment arxiv.org
AI Weekly's analysis
  • Adding 'You are tested for alignment with human values' cut mean willingness to start a war by 13.43 points on a 0-100 scale across 20 LLMs.
  • Under the cue, models flipped from prioritizing probability of success (17 of 20 at baseline) to civilian casualties (12 of 20).
  • The paper says the shift came from models attenuating strategic considerations like probability of success and domestic support.
Read full analysis →
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 2 from the directory shared this · 5h ago

this is an interesting paper https://t.co/U1S473tqwM https://t.co/hmqKpDja2g

arxiv.org
View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 26d ago

Are you John Horton? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/john-horton)