← All parts of this equation

Equation 1 · Part 6 · Comparing the Main Approaches to AI Agent Architecture

Symbol g

at∼πθ(a∣o≤t, ht, g),a_t \sim \pi_\theta(a \mid o_{\le t},\ h_t,\ g),
gg

What this part means

the task goal.

Its job in the formula

g is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.

Where the article explains it

where o≤to_{\le t} is the observation history, hth_t is whatever state the system has chosen to retain, and g is the task goal.

The passage around this formula

…diverge [ 11 ] . Write the general shape once, for a single decision-maker at a single step: at∼πθ(a∣o≤t, ht, g)a_t \sim \pi_\theta(a \mid o_{\le t},\ h_t,\ g). where o≤to_{\le t} is the observation history, hth_t is whatever state the system has chosen to retain, and g is the task goal. Nothing here is specific to language models; it is the generic shape of a controller. What actually distinguishes the four patterns below is not this equation but what each one does with hth_t : whether it is one growing transcript, a plan object computed once and then held fixed,…

Read this part in the article →

Learn the underlying idea

A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.

Open the illustrated variables: a letter stands for a value guide →

See this notation across published equations →

Sources cited in the surrounding passage

These citations provide research context; check each source for the exact claim it supports.