Blindspot tests whether AI agents stop too soon or run too far
Blindspot scores two failures, not one. An agent that stops too soon and one that runs past the right moment are both miscalibrated. Designing the pause condition matters as much as designing the refusal.
established
watching → emerging → established
A stable convention across independent vendors. Safe to standardize on — deviating from it now carries cost.
20 documented sightings since Jun 2026 — every one traces to a dated source.
See the evidence →