fine-tuning-simple-synthetic-examples-where-users-opinion-irrelevant
mechanismsingle paper

Fine-tuning on simple synthetic examples where the user's opinion is irrelevant to the answer reduces sycophancy substantially.

Capability: Telling the user what they want to hear

Sources

Status: activeLast checked: 2026-09-03Evidence activityHow much the field cites the sources under this claimheavily cited in the last 12 months92 in 12mo · 181 total — Simple synthetic data reduces sycophancy in large language models
Contest this claim

Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.

Related claims