← All parts of this equation

Equation 1 · Part 7 · Comparing the Main Approaches to AI Inference Economics

subscript

CdenseCMoE≈NdenseNactive\frac{C_{\mathrm{dense}}}{C_{\mathrm{MoE}}} \approx \frac{N_{\mathrm{dense}}}{N_{\mathrm{active}}}
subscript

What this part means

The lower label selects a particular version, component, or indexed member of the quantity. For example, x₀ and xₜ can be values at different positions.

Its job in the formula

A subscript distinguishes a version, component, step, or member of a quantity. It does not automatically mean multiplication.

The passage around this formula

For inference cost specifically, what matters is that per-token compute tracks activated parameters, not total parameters. Approximating compute per token as proportional to the parameter count actually touched [ 13 , 11 ] , the ratio of dense to MoE compute at equal activated size is approximately CdenseCMoE≈NdenseNactive\frac{C_{\mathrm{dense}}}{C_{\mathrm{MoE}}} \approx \frac{N_{\mathrm{dense}}}{N_{\mathrm{active}}}. for a dense model with NdenseN_{\mathrm{dense}} parameters compared against an MoE model activating NactiveN_{\mathrm{active}} of its NtotalN_{\mathrm{total}} parameters per token. What this ratio hides is exactly what a compute-only comparison always hides: memory. Serving an MoE model requires holding all NtotalN_{\mathrm{total}} parameters resident — on one device or, more often, sharded across…

Read this part in the article →

Learn the underlying idea

A subscript is a label attached below a symbol. It often selects a time step, component, category, or member of a sequence.

Open the illustrated subscripts: which member of a family? guide →

Sources cited in the article section

These citations provide research context; check each source for the exact claim it supports.