The search folds orbits and refits the error model for every point group it
confirms, ranks them, adopts one and discards the rest. The discarded numbers are
the only per-run measurement of how much worse the rivals actually are, and until
now the only way to see them was to re-run with -S and compare by hand.
Behind --finalist-ledger the run prints them: for the adopted group, the refused
higher point group's best, the parent's best and any re-ask candidate that scored
and lost, an evidence row folded from the P1 cross-check merge. That merge is the
substrate the audit kept asking for and nothing else in the run has - the same
observations at full resolution, production-scaled, folded under no symmetry, so
every hypothesis is folded from it on equal terms. It also carries half-set
intensities, which is what lets the noise floor be formed for the first step out of
P1, and the reason this cannot be done from the stored reflection file afterwards.
Report-only, and deliberately so: the accept side has 0.10 of headroom between the
largest genuine ratio measured on the corpus and the bound the gate already uses, so
reading these numbers as a decision before they are calibrated is the one way this
can do harm. The merged output is byte for byte what a run without the flag writes.
Two of the six channels are printed as evidence but marked as measured-contaminated:
a ratio taken against the parent is meaningless when the parent is itself a false
promotion, and on this corpus both false cases sit strictly inside the genuine range
for them. A screw over-call is invisible here by construction - the folds are
identical - and the table says so rather than reading clean.
Every path that cannot produce a ledger now declines out loud instead of printing
nothing: stills, --no-p1-crosscheck, a fixed centred group whose absences were never
predicted, --no-merge, and --mode scale.
Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01EFEJG6WBQv8th4UJFNe53N