← All parts of this equation

Equation 7 · Part 3 · Actually Deploying an Open-Weight Model in Production

Symbol k

r(d+k)r(d + k)
kk

What this part means

k is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

Its job in the formula

k is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

The passage around this formula

Only B and A are trained; W0W_0 never moves during fine-tuning [ 5 ] . The effect on trainable parameter count for that one matrix is to fall from dk to r(d + k) , which is small whenever the chosen rank r is small relative to d and k — and because BA can be merged back into W0W_0 after training, LoRA adds no extra inference latency once deployed. Hu and colleagues report, for their comparison against full fine-tuning of GPT-3 175B, up to a 10,000-fold…

Read this part in the article →

Learn the underlying idea

A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.

Open the illustrated variables: a letter stands for a value guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.