mirror of
https://github.com/vladmandic/automatic
synced 2026-08-26 06:30:44 +02:00
b159daabc9
Repeat-pair runs measured 2-4% between-run drift on one machine, enough to flip threshold verdicts near the margin on every rerun. - bench keeps its iteration samples: rows carry a median sigma, and per-shape sentinel re-measurements sample run-level clock drift - on/off verdicts are three-zone at the run's own noise level; too close to the margin keeps the current setting and says so - pv candidates (now including fp16) are tested independently against the margin at a sidak-adjusted z instead of min-then-threshold - unmeasured toggle stacks are estimated additively in the composition check - per-shape qk verdicts print alongside the reference-shape verdict - cross-gpu error sanity bands flag corrupted measurements