test_first_tag_is_alpha
test_N03_dict_iteration_order.py · failed in 0% of captured runs
The capture budget produced only passing runs (20 runs). Diagnosis compares a passing execution against a failing one; there is nothing to compare.
Where the two runs diverge
The same test, twice, on a shared time axis. The outlined pair is the ordering that differs.
No paired execution was captured for this incident.
Evidence
Suspicion comes from comparing runs. The decision comes from forcing the ordering and seeing what happens.
No ordering differed between the passing and failing runs.
Policy gate
Every check the proposed patch had to pass before it was allowed to run.
No patch was proposed, so there was nothing for the policy gate to review.
Verification
What was established, and at which strength. A weaker check is never presented as proof.
Verification did not run: no patch reached it.