Equation 1 · AI Feeds on the Distance Between an Intention and an Outcome
What does this equation mean?
Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.
This mathematical expression combines the displayed quantities; its precise role follows from the surrounding article text. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.
Read it piece by piece
Symbol C_tr
r is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.
subscript
The lower label selects a particular version, component, or indexed member of the quantity. For example, x₀ and xₜ can be values at different positions.
How to interpret it
Read this expression with the definitions, units, and assumptions supplied by the article.
What the article says around this equation
This makes false positives expensive in the right way. Suppose a redacted request causes every annotator to invent ten fanciful implementation routes, but each route is plainly inconsistent with the repository’s public interfaces. They do not enter . Suppose another request has only two plausible repair stories, but choosing between them changes a promise on which distant callers depend. They remain separate classes even if their patches share most lines. The annotation manual therefore needs positive examples, counterexamples, and a rule that every class name records the observable behavior and scope commitment that distinguish it. It must be frozen before an agent or a…
Read the full surrounding passage
This makes false positives expensive in the right way. Suppose a redacted request causes every annotator to invent ten fanciful implementation routes, but each route is plainly inconsistent with the repository’s public interfaces. They do not enter . Suppose another request has only two plausible repair stories, but choosing between them changes a promise on which distant callers depend. They remain separate classes even if their patches share most lines. The annotation manual therefore needs positive examples, counterexamples, and a rule that every class name records the observable behavior and scope commitment that distinguish it. It must be frozen before an agent or a human is scored. Otherwise the panel will learn which ambiguity was useful to a particular system and retroactively call that the intention gap.
Sources cited in the article section
- [1] SWE-bench: Can Language Models Resolve Real-World GitHub Issues? ↗
- [14] Defects4J: A Database of Existing Faults to Enable Controlled Testing Studies for Java Programs ↗
- [13] RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems ↗
- [7] AgentBench: Evaluating LLMs as Agents ↗
- [10] AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents ↗
These citations give research context. Read each source to check which claims it supports.
Return to AI Feeds on the Distance Between an Intention and an Outcome