← Back to article

Equation 47 · The Economics and Physical Limits of Running AI Agents at Scale

What does this equation mean?

ss

Read the formula alongside the article passage below. Each part has a deeper page with its role in the equation, the supporting passage and nearby citations.

the number of tokens. Read the equation part by part below; each part has a contextual explanation and a link to its mathematical background.

Read it piece by piece

ss

Symbol s

the number of tokens.

Understand this part →

How to interpret it

Read this expression with the definitions, units, and assumptions supplied by the article.

What the article says around this equation

Compaction and editing are not free, and the model above shows why they earn their keep anyway. Reset the accumulated history to a small fixed-size summary of s tokens every m steps, and the trajectory becomes a sequence of n/m short quadratic segments instead of one long one: each segment’s own retransmission cost scales with m2m^2 , but the number of segments scales with n/m , so the total across the whole trajectory scales with m ⋅\cdot n — linear in n for a fixed segment length m , not quadratic. The summarization step itself is not free — reading through a full segment once to compress it costs roughly what one ordinary step near the end of that segment costs — so an interval m chosen too…
Read the full surrounding passage
Compaction and editing are not free, and the model above shows why they earn their keep anyway. Reset the accumulated history to a small fixed-size summary of s tokens every m steps, and the trajectory becomes a sequence of n/m short quadratic segments instead of one long one: each segment’s own retransmission cost scales with m2m^2 , but the number of segments scales with n/m , so the total across the whole trajectory scales with m ⋅\cdot n — linear in n for a fixed segment length m , not quadratic. The summarization step itself is not free — reading through a full segment once to compress it costs roughly what one ordinary step near the end of that segment costs — so an interval m chosen too short pays that overhead too often, and one chosen too long lets the quadratic term inside each segment grow large again before it is cut. Anthropic’s own reasoning behind sub-agent summaries, returning one to two thousand tokens rather than a full working transcript [ 6 ] , is exactly a statement about keeping s small relative to m , which is what makes the reset worth performing at all rather than merely deferring the same bill.

Read the equation in its article →

Sources cited in the surrounding passage

These citations give research context. Read each source to check which claims it supports.

Return to The Economics and Physical Limits of Running AI Agents at Scale

Browse the mathematical compendium →