No AI summary available for this article.
Why It Matters
Large language models often answer the same multiple-choice question inconsistently when it is posed under support-oriented and elimination-oriented framings.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
Large language models often answer the same multiple-choice question inconsistently when it is posed under support-oriented and elimination-oriented framings. We investigate whether these discrepancies arise from different internal representations induced by the two framings. We introduce a dual-framing protocol with minimally varied prompts that use either support- or elimination-oriented framing while keeping the evaluation target fixed. To probe the internal computation, we append an untrained special token, [STATE], and treat its residual-stream activation as an intervention interface. Acr...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.01081v1 · Indexed about 2 hours ago