← Back to article

Equation 13 · Why Average Success Rate Hides the Failures That Matter Most

What does this equation mean?

BiB_i

Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.

the blast radius assigned to failure mode i : how much of a system, how much data, how many users or how much liability a given failure could plausibly reach, drawn from a small number of severity tiers rather than treated as a continuous unknown. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.

Read it piece by piece

BiB_i

Symbol B_i

the blast radius assigned to failure mode i : how much of a system, how much data, how many users or how much liability a given failure could plausibly reach, drawn from a small number of severity tiers rather than treated as a continuous unknown.

Understand this part →

subscript

subscript

The lower label selects a particular version, component, or indexed member of the quantity. For example, x₀ and xₜ can be values at different positions.

Understand this part →

How to interpret it

Read this expression with the definitions, units, and assumptions supplied by the article.

What the article says around this equation

where BiB_i is the blast radius assigned to failure mode i : how much of a system, how much data, how many users or how much liability a given failure could plausibly reach, drawn from a small number of severity tiers rather than treated as a continuous unknown. This is not a hypothetical scoring scheme; tiered, severity-weighted evaluation of exactly this shape is already how the field’s two most cited frontier-risk frameworks operate at the level of an entire model. Anthropic’s Responsible Scaling Policy defines a ladder of AI Safety Levels, modelled loosely on biosafety-level containment standards, under which a model’s required safety, security and deployment safeguards scale with the…
Read the full surrounding passage
where BiB_i is the blast radius assigned to failure mode i : how much of a system, how much data, how many users or how much liability a given failure could plausibly reach, drawn from a small number of severity tiers rather than treated as a continuous unknown. This is not a hypothetical scoring scheme; tiered, severity-weighted evaluation of exactly this shape is already how the field’s two most cited frontier-risk frameworks operate at the level of an entire model. Anthropic’s Responsible Scaling Policy defines a ladder of AI Safety Levels, modelled loosely on biosafety-level containment standards, under which a model’s required safety, security and deployment safeguards scale with the severity of catastrophic risk its own capability evaluations place it at [ 11 ] . OpenAI’s Preparedness Framework runs a parallel structure at the level of specific risk domains — cybersecurity, chemical/biological/radiological/nuclear capability, persuasion, and model autonomy — scoring each as Low, Medium, High or Critical, and committing not to deploy a model that scores High in a given category until mitigations bring the score back down to Medium [ 12 ] . Neither framework was built for scoring a single agent’s task-level failures; both were built for gating whether an entire model ships. The analogy worth drawing is not that the two problems are identical, but that the field’s own most consequential risk-management structures already reject an averaged, ungraded pass/fail count in favour of exactly this kind of tiered, severity-weighted scoring — and an evaluation harness built for consequential agent tasks has good reason to borrow the same shape at a smaller scale, alongside the general risk-taxonomy work national standards bodies have separately published for generative systems more broadly [ 9 ] .

Read the equation in its article →

Sources cited in the surrounding passage

These citations give research context. Read each source to check which claims it supports.

Return to Why Average Success Rate Hides the Failures That Matter Most

Browse the mathematical compendium →