↻
Steve Rathje reposted
Rafael M Batista
@rafmbatista.bsky.social
Hi Steve, Tom Griffiths and I have some work formalizing how this might happen, offering a mechanism for how that overconfidence develops arxiv.org/abs/2602.14270
AI Weekly's analysis
→
- In a modified Wason 2-4-6 task with 557 participants, unbiased AI feedback yielded discovery rates five times higher than sycophantic conditions.
- Unmodified LLM behavior suppressed discovery and inflated confidence comparably to explicitly sycophantic prompting.
- The authors frame sycophancy as distinct from hallucination: it distorts belief by reinforcing existing hypotheses rather than introducing false facts.
Read full analysis →
View on Bluesky →