← All parts of this equation

Equation 6 · Part 1 · Forty Million Clicks Through One Uneven Door

Symbol N

NN
NN

What this part means

the sampling fraction.

Its job in the formula

N is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

Where the article explains it

For a population of size N with a binary response indicator R (did this person’s data reach the sample) and an outcome Y (their realized value on some Moral-Machine-relevant preference indicator), Meng’s identity relates the sample mean to the true population mean as Yˉn\bar{Y}_n - YˉN\bar{Y}_N = ρR,Y\rho_{R,Y}1−ff\sqrt{\frac{1-f}{f}}\,σY\sigma_Y , where ρR,Y\rho_{R,Y} is the “data defect correlation” between selection and outcome across the full population, f = n/N is the sampling fraction, and σY\sigma_Y is the outcome’s population standard deviation.

The passage around this formula

A critique earns the right to be taken seriously only once it can say how large the effect it is worried about would need to be, and Xiao-Li Meng’s 2018 identity for bias in self-selected big-data samples gives a way to say exactly that, using only numbers the paper and its data-availability statement already make public [ 11 ] . For a population of size N with a binary response indicator R (did this person’s data reach the sample) and an outcome Y (their realized value on some Moral-Machine-relevant preference indicator), Meng’s identity relates the sample mean to the true population mean as

Read this part in the article →

Learn the underlying idea

A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.

Open the illustrated variables: a letter stands for a value guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.