withholding-the-users-opinion-is-the-no-training-alternative
mechanismmechanism reasoning

Sycophancy is triggered by the user's stated view being in context, so the cheapest control is not putting it there — ask for the answer before the opinion, or withhold the opinion entirely. Fine-tuning is the answer for the cases where the opinion has to be in context and the answer still must not move.

Capability: Telling the user what they want to hear · Behavior, Chat assistant

Sources

Status: activeLast checked: 2026-09-04Evidence activityHow much the field cites the sources under this claimheavily cited in the last 12 months92 in 12mo · 181 total — Simple synthetic data reduces sycophancy in large language models
Contest this claim

Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.

Related claims

Notes

Backing is mechanism-reasoning, not single-paper, and the gap is specific: the paper's design shows the opinion is what moves the answer, but nobody has tested answer-before-opinion ordering as a deliberate intervention, and I would expect it to leak in multi-turn settings where the view was stated earlier and is still in context. Worth an own-observation claim if I ever run it.