Published equation contexts
Why this formula appears here
Formally, this replaces the single-loop policy with two functions and one artifact: . The plan P is computed against the belief state available at time zero and then held fixed; the executor applies it against a stream of later observations without necessarily invoking the planning policy again. That is precisely the assumption a classical open-loop plan makes, and it is a specific instance of the belief-state fragility decision theory already names: a policy computed once from a belief state is only as good as that belief state stays accurate, and nothing in the equation above notices when the two have quietly come apart [ 11 ] .
Read the representative guide
Symbol pi_plan
plan is an input to the expression that computes the quantity on the left.
Read this term in its guide →Symbol o_0
is an input to the expression that computes the quantity on the left.
Read this term in its guide →Symbol g
g is an input to the expression that computes the quantity on the left.
Read this term in its guide →Symbol a_t
is an input to the expression that computes the quantity on the left.
Read this term in its guide →Symbol pi_exec
pxec is an input to the expression that computes the quantity on the left.
Read this term in its guide →Symbol o_t
is an input to the expression that computes the quantity on the left.
Read this term in its guide →How to interpret it
Read it with the definitions, units, and assumptions supplied by the article.
Research cited beside this formula
Published contexts (1)
A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.
Equation 6 · AI Agents & Systems
Comparing the Main Approaches to AI Agent Architecture
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions.
Formally, this replaces the single-loop policy with two functions and one artifact: . The plan P is computed against the belief state available at time zero and then held fixed; the executor applies it against a stream of later observations without necessarily invoking the planning policy again. That is precisely the assumption a classical open-loop plan makes, and it is a specific instance of the belief-state fragility decision theory already names: a policy computed once from a belief state is only as good as that belief state stays accurate, and nothing in the equation above notices when the two have quietly come apart [ 11 ] .