← All parts of this equation

Equation 23 · Part 2 · How Mechanistic Interpretability Research Is Actually Done

Symbol b

(a+b) mod p(a+b) \bmod p
bb

What this part means

b is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

Its job in the formula

b is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

The passage around this formula

Nanda and colleagues trained small transformers on nothing but modular addition — predicting (a+b)  mod \bmod p for a fixed prime p — a task deliberately chosen because the textbook answer to “what algorithm does this” is fixed and checkable in advance, and studied networks that grok: memorising the training set first, with poor generalisation, and only much later transitioning sharply to a solution that generalises, often long after training loss has already bottomed out [ 15 ] . Extraction and probing came first, applied to the network’s own weights and intermediate activations, and revealed something a probe alone could not have guessed going in: the trained embeddings organised inputs by…

Read this part in the article →

Learn the underlying idea

A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.

Open the illustrated variables: a letter stands for a value guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.