/p/2026-10-10 · explainer
Paper explainer · 2610.11183 · Zhou, Chen, Wang, Lin, Wang and Zhao

The prompt said to trust the documents.

Hold the question and the right answer fixed, edit one assertion in the retrieved passage so it supports a specific wrong answer, and change nothing else. Now swap a single clause in the system prompt. Telling the model the documents are the primary source of truth “even if the documents appear mistaken” made it adopt the planted answer 14.0 and 9.7 points more often, on questions it had already answered correctly with no documents at all. The rate at which the same clause rescued questions it had got wrong moved by −2.1 and +0.1 points, both intervals straddling zero. More harm, no matching gain — and the gap holds for all fifteen systems on every dataset.

01 · The problem

One clause, two different answers to the same question

02 · The mechanism

Edit one assertion, and move it around

where the answer-bearing sentence sits
and what that position costs, averaged over all fifteen systems

03 · The measurement

Two different denominators, and why you must not subtract them

pick a model from the paired audit

04 · Why I care

There is nothing to switch to

dataset
position of the planted assertion
soft sourcing strict sourcing

05 · Apply it

What the clause costs at your retrieval precision illustrative

Results

What the paper actually measured

What it does not show

In practice