David Marx

Why they matter

Directory member with public evidence across Culture, work & education.

AI signals
10
past 30d
Sources
10
distinct domains
Discussões
98
past 30d
Latest signal
23h ago
View every signal from David Marx →
I read a lot of research. Mostly ML. Currently reading: https://dmarx.github.io/papers-feed/ Statistical Learning Information Theory Ontic Structural Realism Morality As Cooperation Free Culture, Open Access YIMBY, UBI Research MLE Frmr FireFighter

Articles & links

David Marx reposted
Ethan Mollick @emollick.bsky.social

OpenAI announces 10 discoveries from their next model. Observations:: 1) AI is getting very good at math 2) Two years ago LLMs failed at basic math 3) This cost less than $2000 in current API fees 4) OpenAI is focusing on announcing benefits, not just risks, of new models open…

openai.com View on Bluesky →

And there it is.

NVIDIA to Acquire Hugging Face blogs.nvidia.com
AI Weekly's analysis
  • NVIDIA agreed to acquire Hugging Face for $12.93 billion, with closing expected in the first half of 2027.
  • Jensen Huang pledged that NVIDIA compute will not be required to build on or deploy through the Hugging Face platform.
  • On top of the purchase price, NVIDIA will pay up to $1 billion in employee retention bonuses, per Engadget.
Read full analysis →
View on Bluesky · ♥ 1 ↻ 0 ↩ 0 · 7 from the directory shared this · 4d ago
David Marx reposted
Sakana AI @sakanaai.bsky.social

We are pleased to share our latest research, now published in Nature Communications: “Smart Cellular Bricks: Physical Modules That Recognize Their Own Shape and Repair Themselves.” Blog: sakana.ai/smart-cellul... Paper: www.nature.com/articles/s41... Thread 🧵

Smart cellular bricks for decentralized shape classification and damage recovery | Nature Communications nature.com
AI Weekly's analysis
  • Cubic bricks running identical neural cellular automata policies classified four 3D shapes with 98.97% accuracy in simulation and 100% on physical hardware.
  • Physical builds ranged from 26 bricks for a guitar to 197 for a round table, converging on a shape label in fewer than 60 update cycles.
  • The same decentralized framework detects structural damage with over 90% accuracy and guides regrowth by predicting one of six axis directions.
Read full analysis →
View on Bluesky →
David Marx reposted
Ted Underwood @tedunderwood.com

I don’t assume Zuck is sincere. However, this is the promise I want companies to compete on. Sharing knowledge—selfishly, purely for the sake of good PR—has a better track record than protecting people from dangerous knowledge in a noble and heartfelt way.

about.fb.com View on Bluesky →
David Marx reposted
Naomi Saphra @nsaphra.bsky.social

Our new paper sets the stage for the biggest practical use case of model interpretability: stress testing and dataset development. All you need is interpretable linear features and simple geometry.

Adversarial Concept Search: Predicting Compositional Errors From Feature Geometry arxiv.org
AI Weekly's analysis
  • A Compositional Interference metric derived from feature geometry predicts LLM failures without evaluating specific inputs.
  • On multihop question answering, correlation between the CI metric and model accuracy reached r = -0.855.
  • The method predicts cross-lingual transfer failures across 10+ languages using only English fact representations.
Read full analysis →
View on Bluesky →

2026 - agents should help users construct preferences, not just elicit them - Irena Saracay, Ludwig Schmidt, Carlos Guestrin 4/n

Beyond expert users: agents should help users construct preferences, not just elicit them arxiv.org
AI Weekly's analysis
  • New arxiv paper argues AI agents should help non-expert users construct preferences, not assume users already know what they want.
  • The authors introduce CoShop, an interactive benchmark where no tested agent exceeded 56% accuracy after five turns of dialogue.
  • Failures came from agents' limited knowledge expansion, not from difficulty finding items once preferences were specified.
Read full analysis →
View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 3 from the directory shared this · 42d ago
David Marx reposted
Phillip Isola @phillipisola.bsky.social

Recently, there have been a lot of impressive demos of AI agents, like Claude, controlling robots. I wrote a short blog post with my thoughts on the advent of these "robot-use agents." web.mit.edu/phillipi/www... I think it's an important change in the trajectory of robotics!

Robot-use agents web.mit.edu
AI Weekly's analysis
  • Phillip Isola argues newer LLMs like Fable and Astra are eroding the case that models can't yet drive robots in general.
  • In his cloud-puppeteer model, any internet-connected robot could gain AI capability via a software update, with no new hardware.
  • Isola concedes LLM-controlled robots remain far less performant than dedicated solutions, with latency and reliability still limiting safety-critical use.
Read full analysis →
View on Bluesky →

It's on you to get your friends and colleagues off of twitter. Pass it along.

Simple contagion drives population-scale platform migration arxiv.org
AI Weekly's analysis
  • Researchers linked 276,431 Twitter/X scholars to their profiles among 16.7 million Bluesky accounts, tracked January 2023 through December 2024.
  • Brazil's court-ordered suspension of Twitter/X served as the natural experiment, with treatment effects on migration that were short-lived and dose-dependent.
  • Adoption was driven by simple contagion rather than complex contagion, with early reconnection to prior contacts predicting longer tenure on Bluesky.
Read full analysis →
View on Bluesky · ♥ 11 ↻ 6 ↩ 2 · 2 from the directory shared this · 76d ago

Recent commentary

"Thank god we're finally making progress, it took forever for the LLM to understand what I was trying to- NOOOOOOOOO!"

View on Bluesky · ♥ 47 ↻ 1 ↩ 4 · 86d ago

For a good?/horrifying? time: throw your research notes at an LLM and ask it to itemize established results from pre-existing research that would have saved you time/effort if you'd seen them earlier. Be ready for a steaming pile of: "Someone did this already and more rigorously decades ago."

View on Bluesky · ♥ 14 ↻ 2 ↩ 0 · 21d ago

AI is good at answering questions. Having access to AI doesn't magically make you better at asking them.

View on Bluesky · ♥ 7 ↻ 0 ↩ 1 · 72d ago

Biggest ChatGPT failure so far: couldn't connect the dots that the Knicks were taking the NBA Finals, attributed city-wide cheering to a Haiti-Scottland WCS game watch party instead.

View on Bluesky · ♥ 2 ↻ 0 ↩ 2 · 86d ago

A hilarious window into the real world business consequences of chasing AI hype culture: contractors who know they could be delivering a cheaper-to-operate solution, but aren't even proposing it because they know their customer wants toys that hoover tokens.

View on Bluesky · ♥ 3 ↻ 0 ↩ 1 · 86d ago

@Anthropic people: for the love of god, can you please teach Claude how to use test-driven development instead of YOLO-implementing fixes based on incorrect assumptions about what the underlying problem was?

View on Bluesky · ♥ 3 ↻ 0 ↩ 1 · 101d ago

Used ChatGPT to track down some distant relatives, then just for funsies had it reconstruct my family tree through public information, working its way all the way back to me. After finding my online footprint, it came up with some interesting characterizations of me. Here are some highlights: 🧵

View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 11d ago

claude hallucinating all sorts of nonexistent postgres features over here

View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 89d ago

I bet anthropic has interesting metrics on how people talk/interact differently with different models

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 68d ago

LLM assisted coding is especially powerful when you're inebriated and can't write coherently, but it understands what you're asking for anyway.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 73d ago

In David Marx's orbit

Center = David Marx. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you David Marx? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/digthatdata-bsky-social)