David Marx

Why they matter

Directory member with public evidence across Culture, work & education.

AI signals
12
past 30d
Sources
8
distinct domains
Discusiones
74
past 30d
Latest signal
6h ago
View every signal from David Marx →
I read a lot of research. Mostly ML. Currently reading: https://dmarx.github.io/papers-feed/ Statistical Learning Information Theory Ontic Structural Realism Morality As Cooperation Free Culture, Open Access YIMBY, UBI Research MLE Frmr FireFighter

Articles & links

David Marx reposted
Sakana AI @sakanaai.bsky.social

We are pleased to share our latest research, now published in Nature Communications: “Smart Cellular Bricks: Physical Modules That Recognize Their Own Shape and Repair Themselves.” Blog: sakana.ai/smart-cellul... Paper: www.nature.com/articles/s41... Thread 🧵

Smart cellular bricks for decentralized shape classification and damage recovery | Nature Communications nature.com
AI Weekly's analysis
  • Cubic bricks running identical neural cellular automata policies classified four 3D shapes with 98.97% accuracy in simulation and 100% on physical hardware.
  • Physical builds ranged from 26 bricks for a guitar to 197 for a round table, converging on a shape label in fewer than 60 update cycles.
  • The same decentralized framework detects structural damage with over 90% accuracy and guides regrowth by predicting one of six axis directions.
Read full analysis →
View on Bluesky →
David Marx reposted
Naomi Saphra @nsaphra.bsky.social

Our new paper sets the stage for the biggest practical use case of model interpretability: stress testing and dataset development. All you need is interpretable linear features and simple geometry.

Adversarial Concept Search: Predicting Compositional Errors From Feature Geometry arxiv.org
AI Weekly's analysis
  • A Compositional Interference metric derived from feature geometry predicts LLM failures without evaluating specific inputs.
  • On multihop question answering, correlation between the CI metric and model accuracy reached r = -0.855.
  • The method predicts cross-lingual transfer failures across 10+ languages using only English fact representations.
Read full analysis →
View on Bluesky →

2026 - agents should help users construct preferences, not just elicit them - Irena Saracay, Ludwig Schmidt, Carlos Guestrin 4/n

Beyond expert users: agents should help users construct preferences, not just elicit them arxiv.org
AI Weekly's analysis
  • New arxiv paper argues AI agents should help non-expert users construct preferences, not assume users already know what they want.
  • The authors introduce CoShop, an interactive benchmark where no tested agent exceeded 56% accuracy after five turns of dialogue.
  • Failures came from agents' limited knowledge expansion, not from difficulty finding items once preferences were specified.
Read full analysis →
View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 3 from the directory shared this · 17h ago

It's on you to get your friends and colleagues off of twitter. Pass it along.

Simple contagion drives population-scale platform migration arxiv.org
AI Weekly's analysis
  • Researchers linked 276,431 Twitter/X scholars to their profiles among 16.7 million Bluesky accounts, tracked January 2023 through December 2024.
  • Brazil's court-ordered suspension of Twitter/X served as the natural experiment, with treatment effects on migration that were short-lived and dose-dependent.
  • Adoption was driven by simple contagion rather than complex contagion, with early reconnection to prior contacts predicting longer tenure on Bluesky.
Read full analysis →
View on Bluesky · ♥ 11 ↻ 6 ↩ 2 · 2 from the directory shared this · 35d ago

Recent commentary

"Thank god we're finally making progress, it took forever for the LLM to understand what I was trying to- NOOOOOOOOO!"

View on Bluesky · ♥ 47 ↻ 1 ↩ 4 · 44d ago

AI is good at answering questions. Having access to AI doesn't magically make you better at asking them.

View on Bluesky · ♥ 7 ↻ 0 ↩ 1 · 31d ago

Biggest ChatGPT failure so far: couldn't connect the dots that the Knicks were taking the NBA Finals, attributed city-wide cheering to a Haiti-Scottland WCS game watch party instead.

View on Bluesky · ♥ 2 ↻ 0 ↩ 2 · 44d ago

A hilarious window into the real world business consequences of chasing AI hype culture: contractors who know they could be delivering a cheaper-to-operate solution, but aren't even proposing it because they know their customer wants toys that hoover tokens.

View on Bluesky · ♥ 3 ↻ 0 ↩ 1 · 44d ago

@Anthropic people: for the love of god, can you please teach Claude how to use test-driven development instead of YOLO-implementing fixes based on incorrect assumptions about what the underlying problem was?

View on Bluesky · ♥ 3 ↻ 0 ↩ 1 · 60d ago

claude hallucinating all sorts of nonexistent postgres features over here

View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 47d ago

I bet anthropic has interesting metrics on how people talk/interact differently with different models

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 27d ago

LLM assisted coding is especially powerful when you're inebriated and can't write coherently, but it understands what you're asking for anyway.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 32d ago

AI ad for some snake oil scam on youtube: > "You don't have constipation: you have POOP PARALYSIS!"

View on Bluesky · ♥ 1 ↻ 0 ↩ 0 · 7d ago

In David Marx's orbit

Center = David Marx. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.