Backtest v7 — corrected forward-drawdown evidence

Evaluation frame 2000-08-15 → 2026-04-07 · evidence base v7_backtest_results.md (pre-computed, rendered verbatim) · page generated 2026-10-01T15:09:07+00:00 · live implementation: monitor v7 (frozen 2026-10-01)

RESEARCH RECORD — v7 corrects the metric the v1–v6 tracks were evaluated on. The tables and the chart on this page are the pre-computed evidence assets, rendered without recomputation; the v7 triggers they selected are live on the monitor. Methodology + track archive: handover.
SPY (log scale) with v7 defensive stages: yellow Level 1, red Level 2, orange Confirmation, grey known drawdown episodes

SPY (log scale) with the v7 defensive stages replayed on history: yellow = Level 1 (Elevated, ≥2 of 4 signals), red = Level 2 (Maximum, B ≥ 450), orange = Confirmation (B+65 / CCC+200), grey bands = known drawdown episodes. Chart generated once by backtest_v7_chart.py and embedded unchanged.

Methodology note. The v1–v6 framework measured “arming” as P(day is within 252 sessions before a known peak | gate hot). This metric is flawed for defensive portfolios because it rewards signals that fire during late-cycle rallies with +10–15% remaining upside — precisely when a defensive portfolio underperforms. v7 corrects this by measuring P(≥10% drawdown from signal-day close within 6 months | fire-run event), which answers the economically relevant question: if I reduce risk today, will the market be lower 6 months from now? All statistics are computed at fire-run level (each contiguous raw-fire event = one test) to avoid distortion from the 45-session latch persistence. The symmetric ±3-month window is provided as supplementary reference to distinguish leading signals (forward hits > backward hits) from lagging confirmation.

The v7 live framework (deployed on the monitor, frozen 2026-10-01)

Base rates (random expectation)

HorizonP(≥10% drawdown)Mean ReturnMedian Return
3 months7.8%3.1%3.9%
6 months13.8%6.3%7.0%

Any signal must beat the 6-month base rate of 13.8% to earn a defensive role.

Table 1: all signals ranked by forward 6M drawdown rate

Fire-run level: each contiguous raw-fire event = one test. Looks forward from fire start.

RankSignalFiresDD 3mDD 6mMean 3mMed 3mMean 6mMed 6mP10 3mP10 6mType
1Baa-G1 +65 OR ≥4251100.0%100.0%17.9%17.9%24.9%24.9%17.9%24.9%reference
2B ≥ 4504017.5%35.0%3.9%3.1%6.9%9.4%-5.4%-6.4%credit_lvl
3B +507420.3%33.8%3.8%4.8%6.7%6.3%-5.7%-7.9%credit_vel
4T10Y2Y ≥ 12m min+0.255820.7%32.8%1.8%2.8%5.1%5.6%-7.0%-5.1%tier2
5ICSA 4wk ≥ 1y min+25k293.4%31.0%4.3%4.3%7.5%9.8%-5.0%-11.2%tier2
6B +408219.5%30.5%3.9%4.2%7.0%6.4%-4.8%-6.7%credit_vel
7CCC +1502326.1%30.4%6.9%6.5%12.0%12.6%1.3%1.2%credit_vel
8B +606725.4%29.9%3.8%4.9%7.7%8.4%-5.0%-4.5%credit_vel
9VIX > 2013020.0%29.2%3.5%4.9%5.1%6.4%-6.1%-8.1%macro
10B +656225.8%27.4%4.5%5.0%9.0%9.7%-6.0%-3.0%credit_vel
11VIX > 1.5x SMA202213.6%27.3%5.5%5.9%10.9%10.2%-4.8%-0.9%macro
12G1′ B+65 OR ≥6006725.4%26.9%3.6%3.8%8.5%7.1%-6.3%-2.8%composite
13CCC +2001216.7%25.0%8.3%8.1%13.5%14.3%-3.6%0.7%credit_vel
14CCC′ +200 OR ≥18001216.7%25.0%8.3%8.1%13.5%14.3%-3.6%0.7%composite
15DGS10 > 4.5%1216.7%25.0%6.2%8.1%12.4%13.6%-3.2%4.0%macro
16BNO-Brent roll ≤ -2%22114.0%24.9%1.9%2.7%5.3%6.2%-7.1%-5.5%tier2
17VIX > 304114.6%24.4%5.9%6.9%9.7%10.8%-4.2%-7.7%macro
18B ≥ 6002623.1%23.1%2.1%1.3%8.5%6.0%-5.5%1.2%credit_lvl
19G3 DGS10≥5.1 OR z≥+32222.7%22.7%1.4%6.0%5.1%10.0%-13.7%-16.0%reference
20B +804520.0%22.2%5.0%5.9%9.6%10.1%-4.6%-1.1%credit_vel
21B ≥ 5003215.6%21.9%5.0%6.0%7.3%7.8%-1.4%-3.0%credit_lvl
22VIX > 257212.5%20.8%6.1%7.0%10.1%13.4%-4.3%-2.6%macro
23SPHB/SPLV < SMA5011712.0%20.5%3.0%4.5%6.6%6.8%-6.9%-3.3%tier2
24CCC +250520.0%20.0%6.4%7.4%12.0%16.1%2.7%2.0%credit_vel
25T10Y2Y < 0520.0%20.0%2.3%9.1%-0.5%3.3%-10.0%-10.9%macro
26B +100238.7%13.0%8.9%9.6%14.3%15.2%2.1%4.2%credit_vel
27B ≥ 700812.5%12.5%11.3%9.5%16.7%16.0%7.5%6.5%credit_lvl
28CCC ≥ 150040.0%0.0%16.8%16.1%24.2%21.3%8.8%14.1%credit_lvl
29CCC ≥ 180020.0%0.0%33.5%33.5%40.6%40.6%28.9%35.7%credit_lvl
30CCC ≥ 20000————————credit_lvl
31CCC ≥ 22000————————credit_lvl
32GLD/CPER > 1.15xSMA50110.0%0.0%11.0%8.6%15.6%10.5%-4.0%4.2%tier2
33DGS10 > 5.0%0————————macro
34T10Y2Y < -0.5160.0%0.0%7.9%8.8%9.3%9.3%4.1%4.2%macro

Table 2: top 15 signals — forward return + drawdown magnitude detail

RankSignalFiresDD 3mDD 6mAvgMag 3mMedMag 3mAvgMag 6mMedMag 6m
1Baa-G1 +65 OR ≥4251100.0%100.0%-18.7%-18.7%-18.7%-18.7%
2B ≥ 4504017.5%35.0%-13.7%-12.0%-13.8%-13.5%
3B +507420.3%33.8%-16.0%-13.2%-16.3%-14.5%
4T10Y2Y ≥ 12m min+0.255820.7%32.8%-17.4%-14.0%-20.2%-18.9%
5ICSA 4wk ≥ 1y min+25k293.4%31.0%-16.5%-16.5%-17.5%-16.7%
6B +408219.5%30.5%-16.1%-13.0%-16.1%-14.2%
7CCC +1502326.1%30.4%-16.4%-13.3%-15.6%-12.4%
8B +606725.4%29.9%-14.9%-12.1%-15.6%-13.2%
9VIX > 2013020.0%29.2%-14.7%-14.2%-16.5%-15.2%
10B +656225.8%27.4%-13.7%-11.7%-14.7%-13.5%
11VIX > 1.5x SMA202213.6%27.3%-24.3%-26.3%-19.5%-15.6%
12G1′ B+65 OR ≥6006725.4%26.9%-13.7%-11.7%-14.7%-13.3%
13CCC +2001216.7%25.0%-19.5%-19.5%-16.6%-12.8%
14CCC′ +200 OR ≥18001216.7%25.0%-19.5%-19.5%-16.6%-12.8%
15DGS10 > 4.5%1216.7%25.0%-18.4%-18.4%-17.3%-17.9%

Table 3: symmetric ±3-month window (reference)

For each fire run: does a ≥10% drawdown occur within 3 months BEFORE or AFTER the signal? Useful for understanding whether signals are leading, lagging, or coincident.

RankSignalFiresSymHitsSymRateFwdOnlyBwdOnlyBoth
1Baa-G1 +65 OR ≥42511100.0%———
2B ≥ 7008450.0%———
3CCC ≥ 18002150.0%———
4B ≥ 600261246.2%———
5DGS10 > 4.5%12541.7%———
6T10Y2Y ≥ 12m min+0.25582339.7%———
7VIX > 30411536.6%———
8G3 DGS10≥5.1 OR z≥+322836.4%———
9G1′ B+65 OR ≥600672334.3%———
10VIX > 201304333.1%———
11SPHB/SPLV < SMA501173429.1%———
12B +60671928.4%———
13B +65621727.4%———
14GLD/CPER > 1.15xSMA5011327.3%———
15CCC +15023626.1%———
16BNO-Brent roll ≤ -2%2215725.8%———
17B +50741925.7%———
18B ≥ 50032825.0%———
19CCC ≥ 15004125.0%———
20B +40822024.4%———

B+40 vs B+50 — head-to-head (credit velocity trigger selection)

MetricB +40B +50B +65
Fire events (25y)827462
Fires per year3.33.02.5
DD 3m rate19.5%20.3%25.8%
DD 6m rate30.5%33.8%27.4%
Mean 3m return3.9%3.8%4.5%
Mean 6m return7.0%6.7%9.0%
Median 3m return4.2%4.8%5.0%
Median 6m return6.4%6.3%9.7%
P10 3m (worst 10%)-4.8%-5.7%-6.0%
P10 6m (worst 10%)-6.7%-7.9%-3.0%
Avg drawdown (3m hits)-16.1%-16.0%-13.7%
Avg drawdown (6m hits)-16.1%-16.3%-14.7%
Symmetric ±3m rate24.4%25.7%27.4%

Evidence-base verdict: Use B+40 — B+50 has a marginally higher DD 6m rate (+0.5pp) but fires 7 fewer times over 25 years, and in defensive positioning sensitivity matters more than specificity. Superseded by the owner decision noted above: the live Level-1 leg 1 and the chart use B+50.