jsbai-aaron's picture
Correct generalization row: audit found 27 of 121 gate instances had entered the training pool pre-measurement; honest split (94 never-trained: 68.1%/60.8%; 27 overlapping: 81.5%/73.6%) replaces the inflated 82.9%. Correction note added with methodology disclosure.
20eaa79 verified