Equation 3 · Part 1 · Comparing the Main Approaches to Robotics and Embodied AI
Symbol s
What this part means
s is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.
Its job in the formula
s is a part of this expression. Its role is fixed by the surrounding article and by the operations shown in the formula.
Full expression→Symbol s→Article meaning
The passage around this formula
Nothing in this update requires knowing or estimating p(s' s, a) ; it only requires having experienced (s, a, r, s') . A model-based approach instead fits an explicit dynamics model — a “world model” — and plans against it directly:
Learn the underlying idea
A variable is a named place for a value. Its letter is a local label: x can mean position in one formula and a data point in another.
Open the illustrated variables: a letter stands for a value guide →
See this notation across published equations →
Sources cited in the article section
- [6] Model-Based Reinforcement Learning: A Survey ↗
- [10] Dynamic Locomotion in the MIT Cheetah 3 Through Convex Model-Predictive Control ↗
- [9] Mastering Diverse Domains through World Models ↗
- [3] Learning to Walk in Minutes Using Massively Parallel Deep Reinforcement Learning ↗
These citations provide research context; check each source for the exact claim it supports.