← All parts of this equation

Equation 6 · Part 10 · Retrieval Is an Evidence System, Not a Memory

Symbol y_1:i-1

p(y∣x)≈∑z∈Zk(x)pη(z∣x)∏i=1∣y∣pθ(yi∣x,z,y1:i−1)p(y \mid x) \approx \sum_{z \in \mathcal{Z}_k(x)} p_\eta(z \mid x) \prod_{i=1}^{|y|} p_\theta\left(y_i \mid x, z, y_{1:i-1}\right)
y1:i−1y_{1:i-1}

What this part means

y1y_1:i-1 is one of the signed contributions combined to compute the quantity on the left.

Its job in the formula

y1y_1:i-1 is one of the signed contributions combined to compute the quantity on the left.

The passage around this formula

The retrieval literature said the opposite. Lewis and colleagues introduced RAG explicitly as a combination of parametric and non-parametric memory, and the architectural point is that the second is a different kind of object, not an extension of the first [ 1 ] . Their formulation treats the retrieved passage as a latent variable to be marginalised over. Writing x for the query, y for the output, and Zk(x)\mathcal{Z}_k(x) for the top k passages returned by a retriever with parameters η\eta : p(y∣x)≈∑z∈Zk(x)pη(z∣x)∏i=1∣y∣pθ(yi∣x,z,y1:i−1)p(y \mid x) \approx \sum_{z \in \mathcal{Z}_k(x)} p_\eta(z \mid x) \prod_{i=1}^{|y|} p_\theta\left(y_i \mid x, z, y_{1:i-1}\right). Read the structure rather than the arithmetic. The generator is never conditioned on the corpus. It is conditioned on z — one span, or a handful — and its output distribution is a…

Read this part in the article →

Learn the underlying idea

A subscript is a label attached below a symbol. It often selects a time step, component, category, or member of a sequence.

Open the illustrated subscripts: which member of a family? guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.