Iteration 7 โ k-of-N: Parametric Variant Killed by Its Own Null ยท "Nothing Is Invariant" Reinterpreted 1 KILL1 MAJOR FINDING
What happened in this iteration, in plain words: we built the instrument the statuses have been waiting for โ one that asks, per feature, "in HOW MANY of the 10 market regimes does your relationship hold?" (instead of the old all-or-nothing question). Then the lab's own safety net caught it cheating: fed real data with the regime structure deliberately destroyed, the instrument still shouted "regime break!" almost everywhere โ a truth-teller must go quiet there. A precision probe showed exactly why: the textbook test it used is fine on homogeneous data (calibrated to within half a percent) but breaks when one side of the comparison is a blend of nine different regimes. So the first version of the instrument is dead โ killed at the assumptions gate by its own control, mechanism identified, replacement designed (calibrate each feature against its own shuffled-regime distribution). Kills are wins. And on the way we found something bigger: the campaign's headline "nothing is regime-invariant (0 of 92)" turns out to be the wrong reading of the instrument. Cracking open the invariance engine shows that for all 100 features it never actually rejected invariance โ it just couldn't identify which anchor combination is the invariant one (many candidates pass, they don't overlap). "Nothing is invariant" becomes "nothing is identified" โ a very different scientific claim that reshapes the STABLE-ceiling question.
Preflight (A0)
2026-07-03 10:19 UTC โ load1 4.81 / 32 cores (โค24) ยท no competing nasimubd jobs ยท 21 GiB available, si/so=0 ยท clickhouse-server active ยท sidecar active + /health healthy โ ALL PASS. Runs: two capped read-only spikes (0.26 min + probe), readonly=2, 5c/5G scope, pre-authorized (03b), zero synthetic values (03c). SSoT maintenance: rule 3e (no-synthetic) folded into LOOP-PROMPT.md's PROMPT block (was in the running loop, missing from the file).
Part 1 โ the kill: parametric Chow k-of-N dies at M2, caught by its own null
THE INSTRUMENT: per feature f, per regime e โ coefficient-equality F-test
(anchors โ f) between env e and the pooled REST of the envs.
k(f) = #regimes where the relationship holds. ฮฑ=0.05 (G3),
BH q=0.10 (G11). All conventions inherited, no new knobs.
MEASURED (real envs): mean k = 1.88 of 10 (k_BH = 0.77)
MATCHED NULL (env structure destroyed by permuting REAL rows):
mean k = 1.91 of 10 โ MUST have been โ10. FAIL.
โ
VERIFY-BEFORE-REPORT PROBE (single env, random halves โ pure exchangeability):
rejection rate 4.57% at nominal 5% โ TEST IS CALIBRATED
โ
MECHANISM (by elimination): not heavy tails, not the harness โ
the e-vs-REST design itself: "rest" = a 9-REGIME MIXTURE; mixture
heteroskedasticity violates the F-test's iid-error assumption โ
over-rejection BY CONSTRUCTION on any multi-regime panel.
โผ
VERDICT: EXCLUDE(TEXTBOOK-ONLY) at M2 โ LEDGER row 9, BONEYARD entry.
Per ยง0: its counts may be reported, never gated.
SUCCESSOR (next Frontier #2 slice): permutation-calibrated per-feature k-of-N โ
each feature's Chow statistic judged against its OWN env-shuffle permutation
distribution: matched null by construction, no distributional assumptions.
Part 2 โ the finding: "ICP 0/92" is an identification failure, not an invariance rejection
The admitted wiring reads len(estimate) โ the INTERSECTION of all accepted
anchor-subsets. Zero can mean two OPPOSITE things:
(a) NO subset accepted โ invariance genuinely rejected
(b) MANY subsets accepted, โ invariance NOT rejected;
intersection empty the invariant set is UNIDENTIFIED
This iteration extracted the accepted-sets family itself, all 100 features:
rejected_everywhere: 0
unidentified_nonempty_family: 100 โ EVERY feature is case (b)
identified: 0
CORRECT READING of the committed campaign evidence:
"0/92 regime-invariant" โ "0/92 IDENTIFIED; invariance rejected for NONE"
Practical consequence unchanged (declarations stay regime-conditional).
Scientific consequence large: the STABLE ceiling may be an IDENTTIFIABILITY
limit, not an absence of invariance โ Frontier #3 (F014) now carries both
questions, converging with the NP-hardness/ฮต-frontier reframe already on file.
Technical record
Item
Value
Frontier item
#2 per-environment k-of-N (crypto) โ bounded slice 1 of the successor design
Harnesses
k_of_n_invariance_crypto.py (measurement + disambiguation + null) ยท k_of_n_size_probe.py (calibration probe) โ both capped, readonly=2, real rows only
Ledger rows
9 (ADMISSION โ EXCLUDE(TEXTBOOK-ONLY) at M2, parametric variant โ kill recorded proudly) ยท 10 (NOTE โ evidence reinterpretation of the committed 0/92)
BONEYARD
parametric Chow e-vs-rest k-of-N, with revisit path = permutation-calibrated successor
Conventions
ฮฑ=0.05 (G3) ยท BH q=0.10 (G11) ยท anchors/act[:4]/target standardization mirrored from the admitted v3 wiring โ zero new magic numbers
Numbers
real mean k 1.88 ยท null mean k 1.91 (fail) ยท single-env calibration 4.57% @ 5% (pass) ยท accepted-family disambiguation 0/100/0 (rejected/unidentified/identified) ยท wall 0.26 min + probe
Forex note
the reinterpretation presumably applies to iteration 6's forex zeros too โ verification rides the full-grid slice
Compute / prod impact
two sub-minute capped spikes ยท ~11 readonly=2 SELECTs ยท zero writes outside audit folder + dashboard
โถ Next iteration
Iteration 8 = Frontier #2, successor slice: the permutation-calibrated k-of-N. In plain words: same question ("in how many regimes does each feature's relationship hold?"), but each feature's regime-break score is now judged against that feature's own shuffled-regime distribution โ a built-in lie detector for every single measurement, immune to the mixture problem that killed version 1. Technically: per feature, Chow statistic per env vs its env-shuffle permutation distribution (real rows only, ~200 draws), permutation p-values โ k(f) at ฮฑ=0.05, BH q=0.10 panel summary; the env-destroyed global null must now read kโN (the successor's acceptance test); runtime estimate ~40โ90s capped. If it passes its own null, the FRAGILEโCONDITIONAL split becomes measurable and the instrument enters the ladder for admission. Steering alternatives: full forex grid extension (fix-plan step 6), or jump to Frontier #3 (F014 โ now sharpened by the identifiability finding).
Iteration 7 ยท 2026-07-03 ยท Frontier #2 slice 1: parametric variant killed at M2 by its own matched null + committed-evidence reinterpretation ยท capped read-only ยท zero synthetic data ยท append-only