← All parts of this equation

Equation 9 · Part 5 · How Training Data and Synthetic Data Actually Work

Symbol delta

Pr⁡[M(D)∈S]≤eε⋅Pr⁡[M(D′)∈S]+δ\Pr[\mathcal{M}(D) \in S] \le e^{\varepsilon} \cdot \Pr[\mathcal{M}(D') \in S] + \delta
δ\delta

What this part means

delta is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

Its job in the formula

delta is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

The passage around this formula

The final mechanism worth separating out is synthetic data generation aimed specifically at privacy rather than capability — producing a dataset that preserves the statistical properties of a sensitive real dataset (medical records, private communications) without preserving any individual record well enough to be re-identified. The standard mechanical tool here is differential privacy, formalized by Abadi and colleagues’ DP-SGD algorithm, which modifies ordinary stochastic gradient descent by clipping each individual training example’s gradient contribution to a bounded norm and then adding calibrated random noise before the aggregated update is applied [ 9 ] . The guarantee this produces…

Read this part in the article →

Learn the underlying idea

A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.

Open the illustrated variables: a letter stands for a value guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.