← All parts of this equation

Equation 3 · Part 5 · Anthropic in 2035: Four Scenarios, Their Signals, and What Would Falsify Them

=

Safeguard(m)={ASL-2,R(m)<τ3ASL-3,τ3≤R(m)<τ4ASL-4,R(m)≥τ4\text{Safeguard}(m) = \begin{cases} \text{ASL-2}, & R(m) < \tau_3 \\ \text{ASL-3}, & \tau_3 \le R(m) < \tau_4 \\ \text{ASL-4}, & R(m) \ge \tau_4 \end{cases}
=

What this part means

The expressions on both sides represent the same quantity under the stated assumptions.

Its job in the formula

The equals sign connects the complete expression on the left with the complete expression on the right. Both sides must have compatible units.

The passage around this formula

The threshold-triggered structure underneath all three versions is stable even as what triggers what has moved. Write R(m) for a model’s assessed risk on a given threat class and τ3\tau_3, τ4\tau_4 for the thresholds separating safety tiers: Safeguard(m)={ASL-2,R(m)<τ3ASL-3,τ3≤R(m)<τ4ASL-4,R(m)≥τ4\text{Safeguard}(m) = \begin{cases} \text{ASL-2}, & R(m) < \tau_3 \\ \text{ASL-3}, & \tau_3 \le R(m) < \tau_4 \\ \text{ASL-4}, & R(m) \ge \tau_4 \end{cases}. What changed between versions is not this shape but the codomain attached to crossing a threshold . Under the original policy, crossing the top implied a commitment to halt. From version 3.0 onward, the same crossing implies an obligation to publish a Risk Report and justify the decision publicly — a disclosure requirement, not a hard constraint. That distinction is the entire disagreement about whether the architecture is strengthening…

Read this part in the article →

Learn the underlying idea

An equals sign says that the expression on its left and the expression on its right have the same value under the stated definitions and assumptions.

Open the illustrated equality: what the equals sign claims guide →

Sources cited in the article section

These citations provide research context; check each source for the exact claim it supports.