Equation 1 · The Hardest Unsolved Problems in Evaluating xAI's Claims
What does this equation mean?
Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.
Read it piece by piece
Symbol u_k
is part of the quantity the equation computes from the expression on the right.
Symbol k
k is part of the quantity the equation computes from the expression on the right.
=
The expressions on both sides represent the same quantity under the stated assumptions.
See an illustrated explanation →subscript
The lower label selects a particular version, component, or indexed member of the quantity. For example, x₀ and xₜ can be values at different positions.
superscript
A raised number can be a power. When it is a label or bound, it selects a case or the upper limit of a sum; the formula’s structure distinguishes these uses.
See an illustrated explanation →How to interpret it
Read it with the definitions, units, and assumptions supplied by the article.
What the article says around this equation
Epoch AI, a research organisation that tracks compute at the frontier, publishes exactly this kind of reconstruction for Colossus and states its own confidence in bands rather than single numbers: an estimate graded “Confident” carries roughly ±3x uncertainty, “Likely” roughly ±10x, and “Speculative” roughly ±31x [ 4 ] . Those three multipliers are not arbitrary; they sit at half-order-of-magnitude steps, . giving 3.16 , = 10 , and 31.6 — which is a formal way of saying that even a well-resourced independent estimate of a frontier training run can be off by an order of magnitude or more, and everyone doing the estimating knows it going in.…
Read the full surrounding passage
Epoch AI, a research organisation that tracks compute at the frontier, publishes exactly this kind of reconstruction for Colossus and states its own confidence in bands rather than single numbers: an estimate graded “Confident” carries roughly ±3x uncertainty, “Likely” roughly ±10x, and “Speculative” roughly ±31x [ 4 ] . Those three multipliers are not arbitrary; they sit at half-order-of-magnitude steps, . giving 3.16 , = 10 , and 31.6 — which is a formal way of saying that even a well-resourced independent estimate of a frontier training run can be off by an order of magnitude or more, and everyone doing the estimating knows it going in. Epoch’s own account of its Grok 4 figures is direct about why: “the FLOP and GPU-hour estimates for Grok-4 are largely based on public statements from xAI, which are often vague,” producing significant uncertainty in the resulting compute total [ 4 ] . Their published point estimate — roughly 5.0×10²⁶ FLOP, built from about 246 million H100-hours — is a best reconstruction, not a confirmed figure, and Epoch says so on the page.
Sources cited in the surrounding passage
These citations give research context. Read each source to check which claims it supports.
Return to The Hardest Unsolved Problems in Evaluating xAI's Claims