Symbol L_joint
oint is computed from the expected values combined on the right.
Read this term in its guide →Published equation contexts
The documented limitation of native pretraining in this literature is not architectural instability — that problem belongs to the next section — but capability gaps that persist despite the joint training. Kosmos-1’s authors introduce a nonverbal Raven-style IQ test specifically to probe reasoning that is not mediated by language, and report that the model reaches only 26 percent accuracy against a random baseline of 17 percent, describing “a large performance gap between the current model and the average level of adults” even as they credit the model with demonstrating the basic capability [ 5 ] . Joint training from scratch buys deeper cross-modal integration; it does not, on this…
oint is computed from the expected values combined on the right.
Read this term in its guide →θ is an argument of the function-like quantity on the left; its role is set by that function’s stated inputs.
Read this term in its guide →E_(ext, mg, ud) sim D appears inside an expected value, so its contribution is averaged under the distribution or condition shown by that operator.
Read this term in its guide →ext appears inside an expected value, so its contribution is averaged under the distribution or condition shown by that operator.
Read this term in its guide →mg appears inside an expected value, so its contribution is averaged under the distribution or condition shown by that operator.
Read this term in its guide →ud appears inside an expected value, so its contribution is averaged under the distribution or condition shown by that operator.
Read this term in its guide →Read it with the definitions, units, and assumptions supplied by the article.
A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.
Equation 6 · Foundation Models
This equation states an equality: the expressions on both sides have the same value under the article’s assumptions.
The documented limitation of native pretraining in this literature is not architectural instability — that problem belongs to the next section — but capability gaps that persist despite the joint training. Kosmos-1’s authors introduce a nonverbal Raven-style IQ test specifically to probe reasoning that is not mediated by language, and report that the model reaches only 26 percent accuracy against a random baseline of 17 percent, describing “a large performance gap between the current model and the average level of adults” even as they credit the model with demonstrating the basic capability [ 5 ] . Joint training from scratch buys deeper cross-modal integration; it does not, on this…
Equation guide → · Article →