← Mathematical compendium

Published equation contexts

Δm=Accfull−Accablate(m)\Delta_m = \mathrm{Acc}_{\mathrm{full}} - \mathrm{Acc}_{\mathrm{ablate}(m)}

Why this formula appears here

The formal version is a simple, reportable metric. For a task with a full-input accuracy Accfull\mathrm{Acc}_{\mathrm{full}} and an accuracy Accablate(m)\mathrm{Acc}_{\mathrm{ablate}(m)} measured with modality m removed, blanked, or replaced with noise, define Δm=Accfull−Accablate(m)\Delta_m = \mathrm{Acc}_{\mathrm{full}} - \mathrm{Acc}_{\mathrm{ablate}(m)}. A small Δm\Delta_m for a modality the task specification says should matter is the signature of a unimodal shortcut: the model is scoring well on the full-input evaluation without actually depending on that channel. This is not a hypothetical failure mode. Agrawal and colleagues showed that VQA models trained and evaluated under the field’s original data splits could reach strong scores while relying heavily on the language prior in…

Read the full article-specific guide →

Read the representative guide

How to interpret it

Read it with the definitions, units, and assumptions supplied by the article.

Research cited beside this formula

Published contexts (1)

A symbol can carry a different meaning in another article. Each occurrence keeps its own guide and term definitions.

Δm=Accfull−Accablate(m).\Delta_m = \mathrm{Acc}_{\mathrm{full}} - \mathrm{Acc}_{\mathrm{ablate}(m)}.

Equation 7 · Foundation Models

Building a Multimodal AI Application That Actually Uses Its Inputs

This equation states an equality: the expressions on both sides have the same value under the article’s assumptions.

The formal version is a simple, reportable metric. For a task with a full-input accuracy Accfull\mathrm{Acc}_{\mathrm{full}} and an accuracy Accablate(m)\mathrm{Acc}_{\mathrm{ablate}(m)} measured with modality m removed, blanked, or replaced with noise, define Δm=Accfull−Accablate(m)\Delta_m = \mathrm{Acc}_{\mathrm{full}} - \mathrm{Acc}_{\mathrm{ablate}(m)}. A small Δm\Delta_m for a modality the task specification says should matter is the signature of a unimodal shortcut: the model is scoring well on the full-input evaluation without actually depending on that channel. This is not a hypothetical failure mode. Agrawal and colleagues showed that VQA models trained and evaluated under the field’s original data splits could reach strong scores while relying heavily on the language prior in…

Meanings in this article

Equation guide → · Article →