the-producing-agent-should-not-be-sole-judge-of-its-output
mechanismmechanism reasoningpending review

The agent that produced an artifact is a biased judge of it — it holds the context and the incentives that skew its assessment — so verification belongs with a deterministic sensor or a separate verifier that reports failures back rather than rewriting the output, and in a multi-agent system the verifier is the one component no agent may override.

Ingested from a paper but not yet reviewed by a human. It is deliberately inert: it does not move any technique’s standing, does not count toward the backtest, and is excluded anywhere a claim would carry weight. Read the source before relying on it.

Evidence for: No one should be judge in their own cause (holds), Two heads are better than one (narrows)

Capability: Fixing its own mistakes · Agentic

Sources

Status: pending-reviewLast checked: 2026-09-11Evidence activityHow much the field cites the sources under this claimvery heavily cited in the last 12 months503 in 12mo · 1164 total — Large Language Models Cannot Self-Correct Reasoning Yet2 sources not yet checked, so not counted
Contest this claim

Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.

Related claims

Notes

Supports external-feedback-repair-works-only-with-real-grounding from the multi-agent side. Filed separately because "who verifies" is a design decision in its own right, distinct from "does reflection help".