Equation 5 · Part 14 · Comparing the Main Approaches to Robotics and Embodied AI
Ending index or upper bound: H
What this part means
This label says where the repeated addition, multiplication, or accumulation stops. It sets the last term or end of the range.
Its job in the formula
H appears in the bound of this sum. The bound states where the repeated operation starts, ends, or which values it includes.
Full expression→Ending index or upper bound: H→Article meaning
The passage around this formula
…queried and its predictions checked against reality independent of the policy it supports; the cost is that policy quality is now bounded by model accuracy, and errors in compound over the planning horizon H in ways that are hard to detect from the outside [ 6 ] . Classical robotics has practiced a version of this for decades without calling it “model-based reinforcement learning”: model-predictive control on the MIT Cheetah 3 solves a convex optimization over ground reaction forces against a…
Learn the underlying idea
Σ adds a collection of terms. Π multiplies them. The lower and upper labels tell you which terms belong to the collection.
Open the illustrated sums and products: repeat an operation over an index guide →
See this notation across published equations →
Sources cited in the surrounding passage
- [6] Model-Based Reinforcement Learning: A Survey ↗
- [10] Dynamic Locomotion in the MIT Cheetah 3 Through Convex Model-Predictive Control ↗
These citations provide research context; check each source for the exact claim it supports.