โ€บNavigation

Iteration 41 โ€” H-021 EILLS Passes the Entrance Exam EXAM PASS ยท M3 CENSUS NEXT

In plain words: the next candidate answers the invariance question with a completely different engine: instead of testing feature-sets one by one, it folds "predict well" and "stay steady across regimes" into a single price, and whoever minimizes the total price wins. On the ground-truth exam it was flawless โ€” handed a target that IS one of the predictors byte-for-byte, it named exactly that predictor as the answer (total price exactly zero), threw out every wrong answer with a price that GROWS as the steadiness weight is turned up (the penalty amplifies wrongness rather than hiding it), and โ€” the encouraging part โ€” its answer didn't budge as the steadiness weight swept three orders of magnitude, so the dial that doomed other candidates may be harmless here. Also structurally hopeful: unlike the last candidate (killed as a shadow of an existing instrument), this one does NOT ride our certified engine โ€” so the "do you add anything new?" fight is genuinely open. That fight is next.

Preflight (A0): 2026-07-10 00:45 โ€” load1 2.92/32c (โ‰ค24) ยท no heavy nasimubd jobs ยท 33 GiB avail, si/so=0 ยท CH + sidecar + kintsugi active โ†’ ALL PASS. Sync note: main moved (#600 โ€” operator amendments to the rotation-probe campaign, not this one; contract unchanged); branch rebased cleanly onto main, force-with-lease push (PR pickup verified). One capped run, 0.43 min, readonly=2 loader, deterministic algebra over real rows only โ€” no permutations even needed.

Exam results (pre-registered form, eills_exam.py)

LegReadingRuling
E1 identification (every ฮณ โˆˆ {1, 10, 100, 1000})All Q-minimizers contain duration_us; smallest minimizer = exactly {duration_us} at Q = 0.0 (byte-identical twin โ†’ exact zero; supersets tie at 0, resolved by the pre-registered parsimony rule โ€” the v5 degeneracy lesson applied cleanly)PASS
E2 rejection (every ฮณ)Every duration-free support strictly separated: Q โ‰ˆ 0.89โ€“0.96 at ฮณ=0 (pure MSE) โ†’ 137โ€“234 at ฮณ=1000 โ€” the penalty amplifies mis-specification. The perfect-fit breakdown that killed causalicp/F014 cannot occur here (no residual-variance division anywhere)PASS
E3 mechanism-alive (descriptive, not gated)Invariance penalty at the pooled-OLS solution on a real second target (burstiness) = 0.601 > 0 โ€” the objective has something to trade on real dataALIVE
ฮณ dial (early read)Selection invariant across three orders of magnitude on the exam anchor โ€” census-scale behavior still owed at M3 (the row-51 ฮบ-sweep pattern)PROMISING ยท VERIFY AT CENSUS
Verdict (row 52)EXAM PASS โ€” checkpoint; ladder walk unlocked, no certificate yet

Why this candidate is structurally different from the one just killed: H-017's stability screen literally ran the admitted v5 e-value oracle, so its readout was that instrument's shadow (ฯ=0.977, row 51). EILLS computes stability as a penalty on per-environment score discrepancy โ€” independent machinery end to end. The M3 redundancy question is therefore genuinely open, not structurally pre-answered.

โ–ถ Next iteration

Iteration 41 ยท 2026-07-10 ยท CHECKPOINT (row 52) โ€” exam PASS ยท capped (0.43 min) ยท readonly=2 ยท zero generated values ยท append-only ยท evidence: eills_exam.py ยท eills_exam_results.json