eills_exam.py)| Leg | Reading | Ruling |
|---|---|---|
| E1 identification (every ฮณ โ {1, 10, 100, 1000}) | All Q-minimizers contain duration_us; smallest minimizer = exactly {duration_us} at Q = 0.0 (byte-identical twin โ exact zero; supersets tie at 0, resolved by the pre-registered parsimony rule โ the v5 degeneracy lesson applied cleanly) | PASS |
| E2 rejection (every ฮณ) | Every duration-free support strictly separated: Q โ 0.89โ0.96 at ฮณ=0 (pure MSE) โ 137โ234 at ฮณ=1000 โ the penalty amplifies mis-specification. The perfect-fit breakdown that killed causalicp/F014 cannot occur here (no residual-variance division anywhere) | PASS |
| E3 mechanism-alive (descriptive, not gated) | Invariance penalty at the pooled-OLS solution on a real second target (burstiness) = 0.601 > 0 โ the objective has something to trade on real data | ALIVE |
| ฮณ dial (early read) | Selection invariant across three orders of magnitude on the exam anchor โ census-scale behavior still owed at M3 (the row-51 ฮบ-sweep pattern) | PROMISING ยท VERIFY AT CENSUS |
| Verdict (row 52) | EXAM PASS โ checkpoint; ladder walk unlocked, no certificate yet | |
Why this candidate is structurally different from the one just killed: H-017's stability screen literally ran the admitted v5 e-value oracle, so its readout was that instrument's shadow (ฯ=0.977, row 51). EILLS computes stability as a penalty on per-environment score discrepancy โ independent machinery end to end. The M3 redundancy question is therefore genuinely open, not structurally pre-answered.