test_cart_total
test_N06_wrong_assertion.py · failed in 100% of captured runs
The capture budget produced only failing runs (20 runs). Diagnosis compares a passing execution against a failing one; there is nothing to compare.
Where the two runs diverge
The same test, twice, on a shared time axis. The outlined pair is the ordering that differs.
No paired execution was captured for this incident.
Evidence
Suspicion comes from comparing runs. The decision comes from forcing the ordering and seeing what happens.
No ordering differed between the passing and failing runs.
Policy gate
Every check the proposed patch had to pass before it was allowed to run.
No patch was proposed, so there was nothing for the policy gate to review.
Verification
What was established, and at which strength. A weaker check is never presented as proof.
Verification did not run: no patch reached it.