abstention-and-self-knowledge · proposedSaying it does not know
Proposed. Several papers in the review queue converged on this framing, so the pipeline added it. Nobody has decided it is the right way to carve up the subject — it may be two topics, or a duplicate of another, or not a topic at all. Saying so is useful.
Covers detecting unanswerable or underspecified problems and gaps between what hidden states encode and what the model asserts; excludes ordinary accuracy on solvable problems.
Tags: general
Proposed from the ingestion pipeline rather than chosen by hand. 2 papers in the review queue independently pointed at this same competence, arriving under 2 different names (unsolvability-detection, latent-knowledge-readout), which is the signal that it is a real recurring topic and not one author's framing. Closely tied to hallucination but distinct enough as an abstention/calibration behavior to track separately.
What counts as this capability
Scope boundary used when deciding whether a paper is really about this capability, rather than merely mentioning it.
Covers detecting unanswerable or underspecified problems and gaps between what hidden states encode and what the model asserts; excludes ordinary accuracy on solvable problems.
Claims
No claims filed yet.
Techniques
None yet.
Capabilities are a way of carving up the subject, and carvings are arguable. Say so if this one is wrong — especially a proposed one, which a pipeline added because several papers used the same framing, not because anyone decided it was right.