- Add reproducible v6 boundary, blessing, and consensus-adjudication
corpora, plus the tiny-transformer trainer and v6 release-gate
evaluator that gate every candidate on the deployed baselines.
- Wire consensus-label merging, product-policy anchor evaluation, and
sealed blessing benchmark review with their pytest coverage.
- Refresh open-training corpus generation, iterative retraining runner,
and random-holdout evaluation so v6 candidates can be benchmarked
end-to-end.