← All parts of this equation

Equation 3 · Part 6 · The Hardest Unsolved Problems in Frontier AI Model Comparisons

=

s^i,b=gb(θi)+δi,b+εi,b,\hat{s}_{i,b} = g_b(\theta_i) + \delta_{i,b} + \varepsilon_{i,b},
=

What this part means

The expressions on both sides represent the same quantity under the stated assumptions.

Its job in the formula

The equals sign connects the complete expression on the left with the complete expression on the right. Both sides must have compatible units.

The passage around this formula

The underlying reason none of these approaches resolves the problem is structural. A published score for model i on benchmark b can be decomposed, at least conceptually, as s^i,b=gb(θi)+δi,b+εi,b\hat{s}_{i,b} = g_b(\theta_i) + \delta_{i,b} + \varepsilon_{i,b}. where θi\theta_i is some unobservable underlying ability vector for model i , gbg_b is the mapping from ability to expected score that is specific to benchmark b and not shared across benchmarks, δi,b\delta_{i,b} is a contamination or leakage term specific to that model-benchmark pair, and εi,b\varepsilon_{i,b} is sampling and decoding noise. Because gbg_b differs across benchmarks by construction — a multiple-choice knowledge test and a pairwise human-preference vote are not measuring the same projection of…

Read this part in the article →

Learn the underlying idea

An equals sign says that the expression on its left and the expression on its right have the same value under the stated definitions and assumptions.

Open the illustrated equality: what the equals sign claims guide →

Sources cited in the article section

These citations provide research context; check each source for the exact claim it supports.