Equation 1 · Ten Failure Modes That Show Up Across Every Frontier Model, Not Just One
What does this equation mean?
Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.
Read it piece by piece
Symbol s_observed
bserved is part of the quantity the equation computes from the expression on the right.
Symbol s_capability
apability is one of the signed contributions combined to compute the quantity on the left.
Symbol Delta_contamination
inflation from the model having seen the benchmark, or something close to it, during training.
Symbol Delta_configuration
variance from everything about how the question was asked, formatted, and scored — prompt wording, decoding settings, which variant answered.
=
The expressions on both sides represent the same quantity under the stated assumptions.
See an illustrated explanation →subscript
The lower label selects a particular version, component, or indexed member of the quantity. For example, x₀ and xₜ can be values at different positions.
How to interpret it
Read it with the definitions, units, and assumptions supplied by the article.
What the article says around this equation
One assumption threads through everything that follows, and it is worth making explicit before the catalogue starts. A published benchmark score is not a direct read of capability; it is a sum of capability plus at least two other things that vary independently of it: . is inflation from the model having seen the benchmark, or something close to it, during training. is variance from everything about how the question was asked, formatted, and scored — prompt wording, decoding settings, which variant answered. Both terms are provider-specific and time-specific; neither is disclosed by the headline number. Three of…
Read the full surrounding passage
One assumption threads through everything that follows, and it is worth making explicit before the catalogue starts. A published benchmark score is not a direct read of capability; it is a sum of capability plus at least two other things that vary independently of it: . is inflation from the model having seen the benchmark, or something close to it, during training. is variance from everything about how the question was asked, formatted, and scored — prompt wording, decoding settings, which variant answered. Both terms are provider-specific and time-specific; neither is disclosed by the headline number. Three of the ten failure modes below are different ways that and get large without the label on the chart changing.
For background, read the article’s source list.
Return to Ten Failure Modes That Show Up Across Every Frontier Model, Not Just One