Equation 5 · Part 13 · Comparing the Main Approaches to Robotics and Embodied AI
Starting index or lower bound: t=1
What this part means
This label says where the repeated addition, multiplication, or accumulation starts. Read its value or condition together with the article’s description of the index.
Its job in the formula
t=1 appears in the bound of this sum. The bound states where the repeated operation starts, ends, or which values it includes.
Full expression→Starting index or lower bound: t=1→Article meaning
The passage around this formula
Nothing in this update requires knowing or estimating p(s' s, a) ; it only requires having experienced (s, a, r, s') . A model-based approach instead fits an explicit dynamics model — a “world model” — and plans against it directly: . A comprehensive survey of the model-based literature frames the trade this way: fitting and planning against typically buys sample efficiency, because every transition teaches the model something reusable across many hypothetical future plans rather than updating one value estimate, and it buys interpretability, because the model can be queried and its predictions checked against reality…
Learn the underlying idea
Σ adds a collection of terms. Π multiplies them. The lower and upper labels tell you which terms belong to the collection.
Open the illustrated sums and products: repeat an operation over an index guide →
Sources cited in the surrounding passage
- [6] Model-Based Reinforcement Learning: A Survey ↗
- [10] Dynamic Locomotion in the MIT Cheetah 3 Through Convex Model-Predictive Control ↗
These citations provide research context; check each source for the exact claim it supports.