models-fine-tuned-b-fail-answer-b-gpt-4-recalls
mechanismsingle paper

Models fine-tuned on "A is B" fail to answer "B is A", and GPT-4 recalls celebrity parents far more often than the reverse.

Capability: Not generalizing "A is B" to "B is A"

Sources

Status: activeLast checked: 2026-09-03Evidence activityHow much the field cites the sources under this claimheavily cited in the last 12 months174 in 12mo · 530 total — The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A"
Contest this claim

Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.

Related claims