01
1. Pick two runs
- Open /runs and choose an earlier run as the baseline and a later run as the candidate.
- Open the comparison page to inspect decision-grade score changes, evidence coverage, and findings that are new, persistent, or resolved.
- Use runs with the same target, template, and scenario set. If scope changed, describe that difference before interpreting the result.
02
2. Read the verdict
- Start with the API comparability verdict; without equivalent scope and decision-grade evidence, do not interpret numeric deltas.
- Review new, persistent, and resolved findings individually; the labels describe differences between these runs, not every possible behavior.
- For an eligible Starter+ run, export the comparison with the evidence pack so another reviewer can trace the scope and observed change.