← Mathematical compendium

Published equation contexts

lim⁡m→∞var⁡ ⁣(dm(Pm,Qm)pE[dm(Pm,Qm)p])=0    ⟹    lim⁡m→∞Pr⁡ ⁣[Dmax⁡(m)≤(1+ε) Dmin⁡(m)]=1\lim_{m \to \infty} \operatorname{var}\!\left( \frac{d_m(P_m, Q_m)^p}{\mathbb{E}\left[ d_m(P_m, Q_m)^p \right]} \right) = 0 \;\;\Longrightarrow\;\; \lim_{m \to \infty} \Pr\!\left[ D_{\max}^{(m)} \le (1 + \varepsilon)\, D_{\min}^{(m)} \right] = 1

Why this formula appears here

Beyer and colleagues proved the canonical result. Under broad conditions on the data and query distributions — much broader than independence and identical distribution across dimensions — as dimensionality rises the distance to the nearest data point approaches the distance to the farthest. Formally, writing dmd_m for the distance function in m dimensions, PmP_m for a data point and QmQ_m for a query point: lim⁡m→∞var⁡ ⁣(dm(Pm,Qm)pE[dm(Pm,Qm)p])=0    ⟹    lim⁡m→∞Pr⁡ ⁣[Dmax⁡(m)≤(1+ε) Dmin⁡(m)]=1\lim_{m \to \infty} \operatorname{var}\!\left( \frac{d_m(P_m, Q_m)^p}{\mathbb{E}\left[ d_m(P_m, Q_m)^p \right]} \right) = 0 \;\;\Longrightarrow\;\; \lim_{m \to \infty} \Pr\!\left[ D_{\max}^{(m)} \le (1 + \varepsilon)\, D_{\min}^{(m)} \right] = 1. for every ε\varepsilon > 0 [ 11 ] . The condition is on the relative variance of the distance distribution: when distances stop varying much relative to their own mean, the nearest neighbour stops being distinguishable from everything else. The authors call such a query…

Read the full article-specific guide →

Read the representative guide

PmP_m

Symbol P_m

PmP_m is an argument of the function-like quantity on the left; its role is set by that function’s stated inputs.

Read this term in its guide →
QmQ_m

Symbol Q_m

QmQ_m is an argument of the function-like quantity on the left; its role is set by that function’s stated inputs.

Read this term in its guide →
Dmax⁡(m)D_{\max}^{(m)}

Symbol D_max^(m)

DmD_max^(m) appears in the objective or constraint used by the optimization on the right.

Read this term in its guide →
Dmin⁡(m)D_{\min}^{(m)}

Symbol D_min^(m)

DmD_min^(m) appears in the objective or constraint used by the optimization on the right.

Read this term in its guide →
Pr⁡\Pr

Probability operator

The probability operator gives the chance of the event named inside its brackets or parentheses.

Read this term in its guide →
E[dm(Pm,Qm)p]\mathbb{E}\left[ d_m(P_m, Q_m)^p \right]

Denominator: E[ d_m(P_m, Q_m)^p ]

The complete quantity below the fraction bar; it must be nonzero for this division.

Read this term in its guide →

How to interpret it

With a fixed numerator, increasing a nonzero denominator reduces the fraction. Read it with the definitions, units, and assumptions supplied by the article.

Research cited beside this formula

Published contexts (1)

A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.

lim⁡m→∞var⁡ ⁣(dm(Pm,Qm)pE[dm(Pm,Qm)p])=0    ⟹    lim⁡m→∞Pr⁡ ⁣[Dmax⁡(m)≤(1+ε) Dmin⁡(m)]=1\lim_{m \to \infty} \operatorname{var}\!\left( \frac{d_m(P_m, Q_m)^p}{\mathbb{E}\left[ d_m(P_m, Q_m)^p \right]} \right) = 0 \;\;\Longrightarrow\;\; \lim_{m \to \infty} \Pr\!\left[ D_{\max}^{(m)} \le (1 + \varepsilon)\, D_{\min}^{(m)} \right] = 1

Equation 6 · AI Agents & Systems

Embeddings and the Geometry of Similarity

This equation states a bound: one expression must stay on the indicated side of the other under the article’s assumptions.

Beyer and colleagues proved the canonical result. Under broad conditions on the data and query distributions — much broader than independence and identical distribution across dimensions — as dimensionality rises the distance to the nearest data point approaches the distance to the farthest. Formally, writing dmd_m for the distance function in m dimensions, PmP_m for a data point and QmQ_m for a query point: lim⁡m→∞var⁡ ⁣(dm(Pm,Qm)pE[dm(Pm,Qm)p])=0    ⟹    lim⁡m→∞Pr⁡ ⁣[Dmax⁡(m)≤(1+ε) Dmin⁡(m)]=1\lim_{m \to \infty} \operatorname{var}\!\left( \frac{d_m(P_m, Q_m)^p}{\mathbb{E}\left[ d_m(P_m, Q_m)^p \right]} \right) = 0 \;\;\Longrightarrow\;\; \lim_{m \to \infty} \Pr\!\left[ D_{\max}^{(m)} \le (1 + \varepsilon)\, D_{\min}^{(m)} \right] = 1. for every ε\varepsilon > 0 [ 11 ] . The condition is on the relative variance of the distance distribution: when distances stop varying much relative to their own mean, the nearest neighbour stops being distinguishable from everything else. The authors call such a query…

Meanings in this article

  • dmd_m: the writing.
  • E\mathbb{E}: The expected value operator: the probability-weighted average of the quantity inside its brackets.
Equation guide → · Article →