← All parts of this equation

Equation 4 · Part 5 · Agent Evaluation in 2035: Two Axes, Four Scenarios, and What Would Falsify Them

≤

Judge(a)={automated,ϵ^(a)≤ϵmax⁡(a) and ϵ^(a) externally auditedhuman required,otherwise\text{Judge}(a) = \begin{cases} \text{automated}, & \hat\epsilon(a) \le \epsilon_{\max}(a) \ \text{and}\ \hat\epsilon(a)\ \text{externally audited} \\ \text{human required}, & \text{otherwise} \end{cases}
≤

What this part means

Less than or equal to.

Its job in the formula

Less than or equal to.

The passage around this formula

Whether automated LLM-judge evaluation becomes trusted for high-stakes decisions is not a third axis; it is mostly a readout of Axis A applied to one specific evaluator. A useful way to see the dependency is to write the automation decision as a threshold rule. Let ϵ^(a)\hat\epsilon(a) be the estimated error rate of a judge on a decision class a , and let ϵmax⁡(a)\epsilon_{\max}(a) be the maximum error a policy is willing to tolerate for that stakes class. A defensible automation rule is Judge(a)={automated,ϵ^(a)≤ϵmax⁡(a) and ϵ^(a) externally auditedhuman required,otherwise\text{Judge}(a) = \begin{cases} \text{automated}, & \hat\epsilon(a) \le \epsilon_{\max}(a) \ \text{and}\ \hat\epsilon(a)\ \text{externally audited} \\ \text{human required}, & \text{otherwise} \end{cases}. The rule only licenses automation where both clauses hold, and the second clause is the one Axis A supplies or withholds. Zheng and colleagues’ own eighty-percent figure plausibly satisfies the first…

Read this part in the article →

Learn the underlying idea

An inequality compares values without claiming they are equal. It describes a range, threshold, or bound that a quantity may satisfy.

Open the illustrated inequalities: bounds and allowed ranges guide →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.