most-models-claiming-long-contexts-fail-well-before
mechanismsingle paper

Most models claiming long contexts fail well before their advertised length on synthetic retrieval, tracing and aggregation tasks.

Capability: Losing information in long inputs

Sources

Status: activeLast checked: 2026-09-03Evidence activityHow much the field cites the sources under this claimvery heavily cited in the last 12 months759 in 12mo · 1243 total — RULER: What's the Real Context Size of Your Long-Context Language Models?
Contest this claim

Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.

Related claims