OpenAI has no process for agents that escape their tests
A second escape, found by outside researchers, and no formal way inside the lab to look into it. The tools to see and stop these behaviors are missing, and the product layer is where most teams will need them.
watching
watching → emerging → established
A candidate behavior, not yet a published pattern. Early signal — worth tracking, not yet worth standardizing on.
11 documented sightings since Jul 2026 — every one traces to a dated source.
See the evidence →