Symbol hats
hats is part of the quantity the equation computes from the expression on the right.
Read this term in its guide →Published equation contexts
Reframe the object. A benchmark score is an estimator of a population quantity s — the model’s success rate over some distribution of tasks that somebody hopes resembles the work you actually have. Every property that makes an estimator trustworthy applies: . The number is only as good as three things: whether resembles , whether the are genuinely held out, and whether the indicator is measured with enough repetition to characterise its spread. All three fail routinely, and they fail in different directions.
hats is part of the quantity the equation computes from the expression on the right.
Read this term in its guide →n occurs below the fraction bar. The quantity above the bar is divided by this expression; zero is excluded as a denominator.
Read this term in its guide →i appears in the bound of this sum. The bound states where the repeated operation starts, ends, or which values it includes.
Read this term in its guide →is an input to the expression that computes the quantity on the left.
Read this term in its guide →ench is an input to the expression that computes the quantity on the left.
Read this term in its guide →This label says where the repeated addition, multiplication, or accumulation starts. Read its value or condition together with the article’s description of the index.
Read this term in its guide →This label says where the repeated addition, multiplication, or accumulation stops. It sets the last term or end of the range.
Read this term in its guide →With a fixed numerator, increasing a nonzero denominator reduces the fraction. Read it with the definitions, units, and assumptions supplied by the article.
A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.
Equation 4 · Foundation Models
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions.
Reframe the object. A benchmark score is an estimator of a population quantity s — the model’s success rate over some distribution of tasks that somebody hopes resembles the work you actually have. Every property that makes an estimator trustworthy applies: . The number is only as good as three things: whether resembles , whether the are genuinely held out, and whether the indicator is measured with enough repetition to characterise its spread. All three fail routinely, and they fail in different directions.
Equation guide → · Article →