Equation 24 · How a Model Actually Gets Small Enough to Run on a Phone
What does this equation mean?
Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.
the dequantised approximation actually used in arithmetic. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.
Read it piece by piece
How to interpret it
Read this expression with the definitions, units, and assumptions supplied by the article.
What the article says around this equation
where is round-to-nearest and is the dequantised approximation actually used in arithmetic. Jacob and colleagues’ integer-arithmetic scheme, still the reference formulation for this mapping, showed that with weights and activations quantized this way, “inference can be carried out using integer-only arithmetic” end to end on ordinary integer hardware, with no floating-point unit required at all, which is what makes the technique valuable on the cheapest edge silicon rather than merely on the largest [ 7 ] .
Sources cited in the surrounding passage
These citations give research context. Read each source to check which claims it supports.
Return to How a Model Actually Gets Small Enough to Run on a Phone