Symbol L_eff
a measured variable — specific to a task, a threshold, a benchmark design, and a date — and every result surveyed above shows it running well below W well before W is reached.
Read this term in its guide →Published equation contexts
Put the documentation and the independent benchmarks side by side and a pattern emerges that is more informative than any single score. Formalise the distinction the whole comparison rests on: let W be the context window a vendor documents for a given model, a fixed engineering figure set by attention implementation, position encoding, and what the serving stack supports. Let S(n) be some benchmark’s measured accuracy at input length n W , and the same benchmark’s accuracy at a short reference length. For a chosen retention threshold , define the effective context length as . W is a documented constant, published on day one of a model’s release, identical no…
a measured variable — specific to a task, a threshold, a benchmark design, and a date — and every result surveyed above shows it running well below W well before W is reached.
Read this term in its guide →τ is an argument of the function-like quantity on the left; its role is set by that function’s stated inputs.
Read this term in its guide →n appears in the objective or constraint used by the optimization on the right.
Read this term in its guide →W appears in the objective or constraint used by the optimization on the right.
Read this term in its guide →S appears in the objective or constraint used by the optimization on the right.
Read this term in its guide →Read it with the definitions, units, and assumptions supplied by the article.
A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.
Equation 6 · Model Evaluation
This equation states a bound: one expression must stay on the indicated side of the other under the article’s assumptions.
Put the documentation and the independent benchmarks side by side and a pattern emerges that is more informative than any single score. Formalise the distinction the whole comparison rests on: let W be the context window a vendor documents for a given model, a fixed engineering figure set by attention implementation, position encoding, and what the serving stack supports. Let S(n) be some benchmark’s measured accuracy at input length n W , and the same benchmark’s accuracy at a short reference length. For a chosen retention threshold , define the effective context length as . W is a documented constant, published on day one of a model’s release, identical no…