not even a month ago
David Marx
Directory member with public evidence across Culture, work & education.
- AI signals
- 12 past 30d
- Sources
- 8 distinct domains
- Discusiones
- 74 past 30d
- Latest signal
- 6h ago
Articles & links
2026 - agents should help users construct preferences, not just elicit them - Irena Saracay, Ludwig Schmidt, Carlos Guestrin 4/n
- New arxiv paper argues AI agents should help non-expert users construct preferences, not assume users already know what they want.
- The authors introduce CoShop, an interactive benchmark where no tested agent exceeded 56% accuracy after five turns of dialogue.
- Failures came from agents' limited knowledge expansion, not from difficulty finding items once preferences were specified.
apparently this nonsense is still going on
It's on you to get your friends and colleagues off of twitter. Pass it along.
- Researchers linked 276,431 Twitter/X scholars to their profiles among 16.7 million Bluesky accounts, tracked January 2023 through December 2024.
- Brazil's court-ordered suspension of Twitter/X served as the natural experiment, with treatment effects on migration that were short-lived and dose-dependent.
- Adoption was driven by simple contagion rather than complex contagion, with early reconnection to prior contacts predicting longer tenure on Bluesky.
If you're coming at this from the background of a traumatized logician, you'll probably enjoy this paper demonstrating how linear representations (i.e. "embeddings" of the kind learned by DNNs) are connected to boolean logic and formal concept analysis.
2025 - An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models - Kenneth Payne 10/n
2026 - The Geometry of Reasoning: Flowing Logics in Representation Space - Yufa Zhou, Yixiao Wang, Xunjian Yin, Shuyan Zhou, Anru R. Zhang 8/n
2025 - From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning - Chen Shani, Dan Jurafsky, Yann LeCun, Ravid Shwartz-Ziv 7/n
2025 - generative artificial intelligence does not have to undermine education - Myles Tan, Nicholle Maravilla 6/n
I haven't updated it in a minute, but I have a solid collection of important ML milestones which might include a few you'd want to add: github.com/dmarx/anthol...
Recent commentary
"Thank god we're finally making progress, it took forever for the LLM to understand what I was trying to- NOOOOOOOOO!"
AI is good at answering questions. Having access to AI doesn't magically make you better at asking them.
Biggest ChatGPT failure so far: couldn't connect the dots that the Knicks were taking the NBA Finals, attributed city-wide cheering to a Haiti-Scottland WCS game watch party instead.
A hilarious window into the real world business consequences of chasing AI hype culture: contractors who know they could be delivering a cheaper-to-operate solution, but aren't even proposing it because they know their customer wants toys that hoover tokens.
@Anthropic people: for the love of god, can you please teach Claude how to use test-driven development instead of YOLO-implementing fixes based on incorrect assumptions about what the underlying problem was?
claude hallucinating all sorts of nonexistent postgres features over here
I bet anthropic has interesting metrics on how people talk/interact differently with different models
LLM assisted coding is especially powerful when you're inebriated and can't write coherently, but it understands what you're asking for anyway.
AI ad for some snake oil scam on youtube: > "You don't have constipation: you have POOP PARALYSIS!"
In David Marx's orbit
Center = David Marx. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.