E4 — live (Day 0: Aug 21, 2026 · read Sep 20)
Charter stamped Aug 21. Standing caveat, pre-committed: E3 matured in a single bullish regime, so long-side results were flattered and shorts penalized by drift. E4 verdicts carry whatever regime E4 delivers; none may be excused or upgraded by regime narrative after the fact. Kill criteria: pooled 7d n < 300 at the read downgrades every verdict to observational; the Best Available suspension lifts only if H1 or H2 passes with a deployable positive signal.
E4-H1
Range-state conditioning (lead): COILED picks beat TRENDING-WEAK picks by ≥ 1.0 pt at 7d, COILED positive with win rate > 50%, n ≥ 100 per cell.
Promoted from the E3 observation (COILED +0.83 vs TRENDING-WEAK −0.77). If it passes: candidate second gate condition for E5. If it fails: stamped as in-sample noise.
SEALED
E4-H2
Ranking-inversion persistence: long book minus short book at 7d, n ≥ 150 per side.
≥ +1.0 predictive · ≤ −1.0 persistently anti-predictive (triggers a scanner-weight review charter for E5) · between: no cross-sectional signal.
SEALED
E4-H3
Grade monotonicity: A ≥ B+ ≥ B at 7d (longs) with A−B separation ≥ 0.5 pt, n ≥ 30 per cell.
Any inversion demotes grades to internal-only pending a recalibration charter. As of Sep 3 the B cell will not reach n by the read — that cell will read UNTESTABLE, not pass.
SEALED
E4-X1
Day-1 kill rule: a pick whose first full trading day closes against entry is marked KILLED-D1. Pass if that cohort's eventual-stop rate exceeds its eventual-target rate by ≥ 15 pts, n ≥ 100.
Designed from the Aug 20 stop/target diagnostic; the mechanical basis (an adverse day 1 sits nearer the stop) is acknowledged in the charter. Measured, never executed.
SEALED
E4-X2
Day-5 time stop: picks unresolved at the day-5 close are marked EXPIRED-D5. Pass if their forward expectancy from day 5 is ≤ +0.10R, n ≥ 80.
Day 5 was chosen over day 3 before any E4 data existed and cannot be adjusted after. If it passes, a ~5-day horizon becomes the stamped plan horizon for E5 only — never retrofitted to E3/E4 rows.
SEALED
B1 and A1 — live (Day 0: Aug 24 / Aug 25, 2026)
Two of the three concurrent slots. Both scoreboards are sealed; the platform shows evidence accumulation only.
B1
Benchmark cohorts: do simple ghost rankings (equal-weight, trend-momentum, relative-strength) match the composite scanner? ±0.5 pt bars at 7d, n ≥ 150 per lane.
Read on or about Sep 23. A composite that fails to beat its own simple benchmarks is a finding, and it will be published as one.
SEALED
A1-H1
Disagreement as a filter: scanner picks the second-opinion model endorses beat the ones it rejects by ≥ 1.0 pt at 7d, n ≥ 100 endorsed / 60 rejected.
Read on or about Sep 24. Known before the read: the scorer stored a direction default for short-side rows; the read derives beat-SPY as-called from the relative return and the scanner direction and never reads the stored result field.
SEALED
A1-H2
Calibration: Brier score < 0.25 and below the base rate; overconfidence gap ≤ 15 pts against the prior 53.5.
Model recorded per row (two models across the window). Verdicts are sliced by model.
SEALED
Mirror Backtest — run Aug 22, 2026 (diagnostic)
The one backtest this platform runs, and only as a check on itself: the frozen E3 configuration replayed over the E3 window, single run, published regardless of outcome. The platform does not backtest to find strategies.
MIRROR
Replay of the frozen scanner over the E3 window lands within ±1.5 pts of the forward result.
Replay +0.726 vs forward −0.13 — inside the pre-registered band, so the forward record is not a logging artifact. The long-short spread flipped sign versus live E3; that finding was handed to E4-H2 rather than interpreted here.
CONSISTENT
E3 — read complete (Day 0: Jul 20 · 7-day checkpoint Aug 3 · 30-day read Aug 20, 2026)
Pre-registered Jul 19. Cohorts: core scanner, universe-x1 (a zero-overlap defensive/dividend list), TSP allocation, Daily Brief. The 30-day read executed a pre-committed suspension: Best Available moved to Parked. That consequence was written before E3 began.
E3-H3
The scanner's 7-day edge persists out-of-sample (confidence interval above 50, n ≥ 30).
By pick-week, 7d avg excess vs SPY: Jul 20 +2.28 (n=122) → Jul 27 −1.93 (n=150) → Aug 3 +0.03 (n=150) → Aug 10 −0.66 (n=90). Pooled −0.13 (n=512, 47.9% win rate). The launch-week edge seen at the Aug 3 checkpoint (+1.49, n=81) did not persist. SUSPEND-to-Parked fired as registered.
FAILS
E3-H1
The regime gate discriminates out-of-sample: the approved side beats the blocked side by ≥ 2.0 pts.
Inverted. The flag read "bull" every night of the matured window; approved longs returned −1.28 vs SPY (7d, n=256) while the blocked short-ranked names returned +1.01 raw (n=256). Single-regime caveat stamped: the gate never faced a non-bull flag in the window.
FAILS
E3-H2
A+ conviction picks outperform A picks by ≥ 1.5 pts (pooled n ≥ 30).
Zero A+ grades fired in 30 days (n=0). Logged calibration finding, not a verdict: the grade ladder was non-monotonic on longs at 7d — A −0.83 (n=52), B −1.28 (n=59), B+ −1.59 (n=141). Became an E4-H3 design input.
UNTESTABLE
E3-H4
The universe-x1 expansion list performs within 1 pt of the core list.
Core −0.12 (n=341) vs x1 −0.16 (n=171); delta 0.04. The machine generalizes across a zero-overlap universe — and the signal it generalized was not working in this window. Both halves stated.
CONFIRMS
Q-R1/R2
Range-state questions: does the edge differ across TRENDING / RANGING / COILED tape?
COILED +0.83 (7d, n=186, 55.9% win) vs TRENDING-WEAK −0.77 (7d, n=420, 43.8%). The sharpest split in the dataset, found post-hoc — so it cannot be claimed as validated. Pre-registered as E4-H1 instead.
OBSERVATIONAL
E3-BRIEF
Daily Brief v1 (AI free-picks from the watchlist) adds value vs SPY.
Negative at the Aug 3 checkpoint; −1.26 (7d, n=94) at the 30-day read. Product retired Aug 12 — see Retirements.
NEGATIVE
E2 — closed (Jun 1 – Jul 17, 2026; addendum Jul 19)
Ran across three sub-regimes: melt-up tail, geopolitical whipsaw (strikes, FOMC), and the relief rally. Closed with an addendum after a scoring-corruption repair; the addendum had power to confirm or mark low-confidence only, never to reverse. Zero verdict reversals.
E2-GATE
Regime-gated long picks beat SPY at 7 days.
+2.93 avg excess, n=158 at close. The addendum's rescored master table landed on the same number.
CONFIRMS
E2-H4
The edge is a 7-day edge; it decays by 30 days.
Completed 30-day cohort −1.41 — the picks that won at 7d faded by 30d, reinforcing the pre-registered horizon.
REINFORCES
E2-RELIEF
Relief-rally cohort at 30 days (journal cell, not a hypothesis).
25% beat rate / −8.97 avg, n=40 — matured directly into the July semiconductor rout. Stamped as hostile-tape context per the pre-committed journal rule, and displayed, not excused.
HOSTILE MARKET
Retirements
Products killed by their own scoreboard. Records preserved in full on the Track Record page; v1 and v2 numbers are never blended.
BRIEF-V1
Daily Brief v1 — AI-selected picks from the user watchlist.
Failed the Aug 3 read (negative excess vs SPY), confirmed unhealthy at the Aug 20 read. Retired Aug 12, 2026. Replaced by v2, whose picks are restricted to the scanner's own validated candidate pool and logged as a fresh cohort under its own tag, starting at n=0.
RETIRED
SCAN-V2
Legacy scanner-v2 demo logging.
Disabled Jul 2, 2026. Its era's corrupted outcome rows (677) were deleted in the all-time cleanup with the deletion documented — the only rows ever removed, removed for being unscoreable, not for being losses.
RETIRED
Data integrity — caught and repaired
A ledger that only listed wins would be marketing. These are the bookkeeping failures found, fixed and disclosed. Since Sep 3 a nightly integrity check grades every lane at 11:45 PM ET and writes the grade to a log the public page reads.
SEP-3
Production stuck on an Aug 30 build for four days: a truncated integrity-check file failed every deploy.
Rebuilt Sep 3; the check now writes integrity_log nightly. A holiday weekday graded FAIL until Sep 5 — it now grades SKIPPED from the same trading calendar every logging lane uses.
CONFIRMS
AUG-4
Benchmark-snapshot gap: the TSP logger wrote no SPY snapshot after Jul 21; the Daily Brief snapshot was context-dependent; a zero-value bug ate exactly-flat returns.
Found because the Aug 3 read's cohort and gate averages disagreed by 0.74 pts. Root-caused to three bugs within 24 hours, repaired with official EOD closes (63 predictions stamped, 30 outcomes re-based, zero scoring windows contaminated). A snapshot-health watchdog now runs in the read tooling.
CONFIRMS
JUL-20
Regime-gate auth bug: server-to-server calls carried no session, silently neutralizing the macro tilt.
Caught hours before E3 Day 0; fixed by threading authentication through every caller. Disclosed because a silently-neutral gate would have invalidated E3-H1.
CONFIRMS
JUL-19
E2 scoring corruption: a scorer zero-fill wrote fake outcomes (a −84.59% "loss" on a real +9.66% win).
Rescored against real window-end closes under a pre-committed addendum with confirm/low-confidence power only. True numbers for the Jun 15 cohort were worse than the corrupted ones and were recorded as-is.
CONFIRMS
Next gates
Nothing about the live methodology changes between reads. A deploy freeze holds Sep 19–24 around the three reads.
E4-READ
E4 30-day read, Sep 20, 2026 (7-day cutoff Sep 13).
At stake: whether Best Available leaves Parked. Only H1 or H2 passing with a deployable positive signal can lift the suspension.
PENDING
B1-READ
B1 benchmark read, on or about Sep 23, 2026.
Unseals the three ghost lanes against the composite.
PENDING
A1-READ
A1 second-opinion read, on or about Sep 24, 2026.
Sliced by model. The scorer direction fix and the regime backfill are bundled after this read, never before it.
PENDING