← Mathematical compendium

Published equation contexts

weff=w+skw_{\mathrm{eff}} = w + \frac{s}{k}

Why this formula appears here

Fact. The move from FP8 to sub-8-bit formats has already happened once, and the mechanism by which it happened exposes the trade-off that will decide whether it happens again. A microscaling format with element width w bits, block size k , and a shared scale of s bits carries an effective per-element cost of weff=w+skw_{\mathrm{eff}} = w + \frac{s}{k}. with the second term the amortised metadata tax. The OCP MX alliance’s MXFP4 uses w = 4 , k = 32 , and an 8-bit shared scale, for weffw_{\mathrm{eff}} = 4.25 bits per element — a roughly 6% tax [ 4 ] . NVIDIA’s NVFP4 instead uses a smaller block, k = 16 , with the same 8-bit block scale plus a near-negligible per-tensor term, for weffw_{\mathrm{eff}} ≈\approx 4.5 bits per…

Read the full article-specific guide →

Read the representative guide

weffw_{\mathrm{eff}}

Symbol w_eff

wew_eff is part of the quantity the equation computes from the expression on the right.

Read this term in its guide →

How to interpret it

With a fixed numerator, increasing a nonzero denominator reduces the fraction. Read it with the definitions, units, and assumptions supplied by the article.

Research cited beside this formula

Published contexts (1)

A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.

weff=w+skw_{\mathrm{eff}} = w + \frac{s}{k}

Equation 9 · Semiconductors

AI Accelerator Architecture in 2035: Scenarios, Signals, and Falsifiable Predictions

This equation states an equality: the expressions on both sides have the same value under the article’s assumptions.

Fact. The move from FP8 to sub-8-bit formats has already happened once, and the mechanism by which it happened exposes the trade-off that will decide whether it happens again. A microscaling format with element width w bits, block size k , and a shared scale of s bits carries an effective per-element cost of weff=w+skw_{\mathrm{eff}} = w + \frac{s}{k}. with the second term the amortised metadata tax. The OCP MX alliance’s MXFP4 uses w = 4 , k = 32 , and an 8-bit shared scale, for weffw_{\mathrm{eff}} = 4.25 bits per element — a roughly 6% tax [ 4 ] . NVIDIA’s NVFP4 instead uses a smaller block, k = 16 , with the same 8-bit block scale plus a near-negligible per-tensor term, for weffw_{\mathrm{eff}} ≈\approx 4.5 bits per…

Meanings in this article

  • kk: the block size.
Equation guide → · Article →