Hedges and tag questions shorten LLM answers, paper finds
TL;DR
- Prompts using hedges, tag questions, and collective reference drew shorter, less sophisticated, less formal LLM replies across three document types and four models.
- Linguistic register outweighed explicit cues like sign-off names in shaping response quality, according to the paper's representational analysis.
- The authors report the effect is encoded in early transformer layers and entangled with other features, making post-hoc mitigation difficult.
Prompts that lean on hedges, tag questions, and collective reference elicit "shorter, less sophisticated, and less formal responses across three document types and four models," according to a new arXiv preprint from Katherine Van Koevering and Anjalie Field.
The register effect swamps explicit demographic cues. The paper reports that "explicit gender cues like sign-off names are encoded in the same representational space as linguistic dialect" — but it is the dialect that moves the model, not the name attached to it. Two of the researchers we track posted the paper the week it went up.
Mitigation looks hard. The features sit "in early transformer layers and entangled with other features," the authors write, and they warn that "post-hoc mitigation is challenging."
The framing is workplace, not lab curiosity. Van Koevering and Field close by calling for "upstream consideration of the influences of linguistic variation to mitigate disparate impacts of LLM-mediated workplace communication."
Shared on Bluesky by 2 AI experts
-
New press release about our recently released COLM paper, work by Katherine Van Koevering! hub.jhu.edu/2026/09/21/a... Check out the full paper here: arxiv.org/abs/2608.13328
View on Bluesky →
Originally reported by arxiv.org
Read the original article →Original headline: It's How You Ask: Gender-Associated Linguistic Bias in LLMs