Decision-grade verdict from a rigorous refit study (5 purged/embargoed folds, 5,671 OOS windows, real BTC CLOB labels): the signal is HEALTHY and the edge PERSISTS — it is NOT being priced out. Honest recalibration of the frozen map is a NO-GO for EV (recovers ~zero, just refuses near-breakeven expensive-ask fires). This note ALSO visibly retracts an earlier same-session read that the edge was eroding.
Provenance & scope
Single-session, read-only, un-peer-reviewed (main agent, 2026-07-11). No live change, no code merged,
calibration.jsonuntouched. Rigorous statistical folds are BTC-only (binance_ticksis BTC-only); the 5 alt assets were cross-checked (+EV late week) but not independently refit. Treat the alt read as directional, not gate-cleared.
RETRACTION — the “edge is eroding” read was WRONG
RETRACTED (2026-07-11): "the signal edge is being priced out"
An earlier read in this same session claimed the signal edge was decaying — that real win-rate had fallen 80%→73% and the at-ask ceiling had collapsed +7¢→~0. This is WRONG. Do not carry it forward.
Why it was wrong: it was a coarse CROSS-ASSET by-day aggregate, skewed by SOL. On 07-10, sol printed −11.2¢ (n=62, a PARTIAL day), which dragged the pooled number down. Measured PER-ASSET the same day, BTC was +9.2¢ and ETH was +3.8¢ — the opposite of erosion.
Lesson (binding): measure erosion PER-ASSET, never as a cross-asset aggregate. A single weak/partial-day lane in the pool manufactures a fake “decay” trend. This is the same asset-blind-aggregate family as the window_id cross-asset join trap — pooling across lanes that differ by construction fabricates artifacts. The rigorous refit study below (per-asset, purged folds) is the correct read and it overturns the erosion claim.
VERDICT — MARGINAL / UNDERPOWERED; the edge PERSISTS, is NOT priced out
The rigorous study (5 purged/embargoed folds, 5,671 OOS windows, real BTC CLOB labels) refits the calibration honestly and compares it to the frozen map through the live edge gate.
1. The signal is HEALTHY
- The late window (≥07-05) is the STRONGEST, not the weakest — directly contradicting the retracted “eroding” read.
- Honest signal win-rate is stable ~0.77–0.79, +3.46¢/fire.
- All 6 assets are +EV in the late week.
2. Honest recalibration is a NO-GO for EV
Frozen vs honest-refit, through the gate:
| Fires | Win-rate | Net PnL | own_z | |
|---|---|---|---|---|
| Frozen (current live map) | 2,633 | 76.3% | +$42.36 | 2.00 |
| Honest-refit | 1,679 (~40% fewer) | 77.0% | +$37.13 | 2.30 |
- Δz = −0.30 (a statistical tie). Best of 16 refit variants = +0.16, far below the deflated multiple-testing bar of 3.99.
- The frozen map IS overconfident ~5–11pt/bin (it claims 0.912 where the real rate is 0.808) — but the gate absorbs the overconfidence at the ~0.72 ask. So recalibration recovers ZERO EV; it merely refuses ~40% of fires that were near-breakeven expensive-ask fires. Fewer trades, same edge.
3. The loss is FILL-side, not the signal
This reconfirms the standing finding: the bleed is A0 first-touch adverse fill, not the model. See the FILL adverse-selection decomposition and the A0-lever confirmation. The lever is fill-capture. DO NOT shelve the strategy.
Actionable
What to actually do
- A0 fill fix = the #1 lever. The toxic first-touch fill is where the money goes; that is the work, not the model.
- The durable edge lives in cheap asks (< 0.75). The 0.78–0.85 tail is thin and near-breakeven → lower
EMIT_MAX_ASKfrom 0.85 → ~0.78–0.80 AND re-gate when re-arming. (This tightens, does not remove, the validatedEMIT_MAX_ASKbound — see CLAUDE.md rule: keepEMIT_MAX_ASK + buffer ≤ EMIT_MAX_LIMIT, re-run the gate on any change.)- Recalibrate ONLY for sizing-honesty, never for EV. The frozen map over-sizes by ~10pt (claims 0.912 where real 0.808) — a honest map would size correctly. But since sizing is currently gate-absorbed, this is a hygiene item, not an EV win.
calibration.jsonstays frozen unless/until the sizing path consumes the honest probability.
Genuine-erosion re-check trigger
Do NOT declare erosion off a cross-asset aggregate again. The real trigger for genuine signal decay is: frozen-gated FILL win-rate < 73% sustained over a full ≥5-day window (per-asset, not a partial day, not a pool). Below that sustained bar → re-open the refit question.
Caveats
BTC-only rigor
The purged/embargoed folds are BTC-only because
binance_tickscarries only BTC. The 5 alts were cross-checked (+EV late week) but not independently refit. Resolve this before any live GO — an alt lane could have drifted without the BTC-only study seeing it. (Note the standing [[crypto-shortterm-db-index-audit-2026-07-09|binance_tickspkey(trade_ts, price)has no asset column]] gotcha — alt ticks would need their own populated tape before an alt refit is even possible.)
Current live-state
Live exposure = NONE
- All lanes DRY — no live exposure.
- The overnight re-arm was PULLED after night-1 lost −$189 (directional). This reverses the provisional 07-10 overnight re-arm — consistent with that note’s own framing (a provisional risk tilt, NOT a validated clock edge; night-1 landing negative is within the underpowered-noise band it flagged).
Artifacts
Where the work lives
- Branch:
feat/calibration-refit-degraded-rerun(commitae31a0c) — NOT pushed, NOT merged.- Report:
docs/superpowers/refit-honest-degraded-verdict-2026-07-11.md- Scripts:
scripts/gate.py,scripts/gate2.pycalibration.jsonuntouched (frozen, as required).
Related
- crypto-shortterm-disagreement-ceiling-nogo-2026-07-03 — the prior refit-study NO-GO (07-03): a selection ceiling run through the same refit/exit walk-forward rig, rejected directionally. Same rig family, same “overlays don’t beat the frozen edge” conclusion.
- crypto-shortterm-loss-decomposition-fill-adverse-selection-2026-07-10 — the FILL adverse-selection / A0 decomposition this verdict reconfirms.
- crypto-shortterm-overnight-gate-a0-lever-rearm-2026-07-10 — the A0-lever confirmation + the overnight re-arm that was just pulled.
- crypto-shortterm-depth-aware-sizing-nogo-2026-07-10 — depth-aware sizing NO-GO (wrong lever; not size, not depth).
- crypto-shortterm-loss-investigation-real-vs-onchain-edge-2026-07-10 — the
resolved_upoverstates traded win-rate ~10pt read (why frozen over-sizes). - crypto-shortterm-risk-management — portfolio risk, caps, kill switch.
- crypto-shortterm-observations-window-id-not-asset-unique — the asset-blind-aggregate trap family the retraction belongs to.
- signal-healthy-fill-is-the-problem — memory pointer.
- loss-is-adverse-selection-at-fill — memory pointer.
- onchain-label-overstates-traded-winrate — memory pointer.