← Back to article

Equation 34 · What an AI Accelerator Actually Is: Silicon, Packaging, and the Memory It Can Reach

What does this equation mean?

pp

Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.

the latency term grows linearly in. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.

Read it piece by piece

pp

Symbol p

the latency term grows linearly in.

Understand this part →

How to interpret it

Read this expression with the definitions, units, and assumptions supplied by the article.

What the article says around this equation

with α\alpha the per-step latency and βlink\beta_{\mathrm{link}} the per-link bandwidth. The bandwidth term approaches 2N/βlink\beta_{\mathrm{link}} and stops growing with p ; the latency term grows linearly in p . Small, frequent collectives are therefore latency bound and scale badly, while large ones are bandwidth bound and scale well — which is precisely why gradient bucketing and overlapping communication with backward computation are standard practice rather than optimisations.

Read the equation in its article →

Sources cited in the article section

These citations give research context. Read each source to check which claims it supports.

Return to What an AI Accelerator Actually Is: Silicon, Packaging, and the Memory It Can Reach

Browse the mathematical compendium →