Equation 28 · Why Coding Agents Fail: Long-Horizon Reliability in OpenAI Codex
What does this equation mean?
Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.
Read it piece by piece
=
The expressions on both sides represent the same quantity under the stated assumptions.
See an illustrated explanation →How to interpret it
Read it with the definitions, units, and assumptions supplied by the article.
What the article says around this equation
Define C as the agent’s completion claim and V as external task success. Four states follow: . False completion is operationally expensive because it transfers discovery to the reviewer. It can arise from missing tests, misread output, stale state, premature stopping, or incentives in the prompt to present a polished result. A reliable harness should make “blocked,” “partially complete,” and “evidence inconclusive” legitimate terminal states.
For background, read the article’s source list.
Return to Why Coding Agents Fail: Long-Horizon Reliability in OpenAI Codex