# Config Watch #7

**WEEKLY · CONFIG WATCH**  ·  Issue #7  ·  the ceiling rises and the board with it

Issue date: September 23, 2026  ·  Language: English  ·  Methodology v1.1 + Config Watch amendments (approved 12.08, hash e66c7e8c864a2233)

Search windows: **WK** Sep 14-21, 2026 and **M30** Aug 24-Sep 21, 2026 -- both frozen and Monday-aligned (WK = Mon Sep 14 00:00 -> Sun Sep 20 23:59 UTC; M30 = Mon Aug 24 00:00 -> Sun Sep 20 23:59 UTC, 4 full weeks), sharing the cut-off edge Mon Sep 21 00:00 UTC, so slots dated Sep 21 fall outside both.

*Slots fall inside the window; trades settle past its right edge at their horizon -- the same 2-3 day luft the engine has used since issue #1 and that the live presets run on.*

PDF: https://marketmania.ai/research/reports/config-watch-2026-09-16.pdf · Open data (JSON): https://marketmania.ai/research/reports/config-watch-2026-09-16.json

---

## Snapshot

- **16,642** simulations, 0 errors, in two frozen collects
- **WK** Sep 14-21, 2026
- **M30** Aug 24-Sep 21, 2026
- 7-stage search: scan -> grid -> refine -> stake sweep -> per-ticker -> OOS replays
- engine 1.1; **$1,000** deposit, **10x** leverage
- stake is **universe-aware** from sim #1 (table below)
- maturity gate: >=**15** closed (WK) / >=**20** (M30)
- criteria frozen in writing before launch (w276 plan)
- walk-forward: **43** past cards + **4** live presets replayed verbatim on both windows
- **14 of 43** past cards Survive (net-positive on both windows)
- stability: **16** configs across WK and M30 (night run)
- models: **7** (grok axis = grok-4.6 (incl. 4.5-era))
- best WK council **+$1,467.8** (+146.8%)
- best M30 council **+$1,351.8** (+135.2%)
- zero-trade cells: **919** (weekly) / **1,658** (night run)
- margin-skip alerts on leaders: **4** (WK) / **16** (M30) rows

---

## KEY FINDING

> **[OBSERVATION (the ceiling rises and the board with it)]** The ceiling of the fresh search moved from +$609.5 to +$1,467.8 on the week and 33 of the 43 cards this series has published are net-positive on that same week, and under the standing two-window rule 14 of the 43 Survive: 16 clear the month, 33 clear the week.

---

## TL;DR

- OBSERVATION -- **The standing board grew to 43 cards.** Issue #6 kept 6 of its 38 past cards under the same rule; this issue replays **43 cards** -- issues #1, #2, #3, #4, #5, #6 and Monthly Config Watch #1 -- on two fresh windows, and **14 of the 43 Survive** (net-positive on both). On the week alone 33 of 43 are positive, on the four-week window 16 of 43. Issue #6's own week champion, replayed verbatim, makes +$935.6 on the week and +$711.7 on the month.
- **The search ceiling, window by window.** The best mature council on WK makes +$1,467.8 (+146.8% of a $1,000 deposit, 80.0% win rate, 7.7% maxDD, 60 closed) against issue #6's +$609.5; on M30 it makes +$1,351.8 (+135.2%, 66.9% win rate, 69.8% maxDD, 145 closed) against issue #6's +$1,793.8. The grids behind them: median +$79.2 with 57.0% positive across 4,176 WK cells (960 of them above +$500), median -$83.5 with 25.7% positive across 4,176 M30 cells (169 above +$500).
- **Calibration and the universe, window by window.** Across the whole stage-2 grid calibration improves 25.8% of the 1,452 matched WK pairs (mean -$111.9) and 65.9% of the 1,453 matched M30 pairs (mean +$253.6); pooled, 45.9% of 2,905. On the cards, **no published WK council is a calibrated cell** and **1 of the 2 M30 cards are calibrated cells**. The mature net top-10 is a single universe on each window: SOL+XRP on WK (10 of 10 slots), SOL+XRP on M30 (10 of 10).
- **The stability board, a fourth reading of the same pair.** The night collect computes WK and M30 in one pass, so this issue's cross-window board is measured on the two windows the issue prints -- and **16 configurations** are in the mature net top-40 of both. Issue #6 measured the same pair one week earlier and held 9, issue #5 the week before that 0, and issue #3 21. The hand-tuned anchor keeps its standing measurement against the whole mature field of each window -- 2,015 of 2,652 mature WK configs and 233 of 2,989 mature M30 configs beat it on win rate, maxDD and smoothness at the same time. Net PnL stays the headline board (continuity with issues #1-#6).

---

## Stage-2 grid: where the defaults sit (WK)

![WK stage-2 grid net PnL histogram](chart_hist_wk.png)

*Net PnL across all 4,176 stage-2 grid sims, WK window (Sep 14-21, 2026), universe-aware stake throughout. Grid median **+$79.2**, 57.0% positive (issue #6's week: median -$32.2, 17.5% positive; its month: -$0.1, 30.3% positive). The spike at $0 is mostly cells whose entry filters produced no trades (918 zero-trade sims on this window). Range -$781.8 to +$1,467.8; 960 cells finished above +$500. Red marker = best solo default at table stake; green = the published cards. The same board definition one week earlier put its top-5 at +$609.5 down to +$568.9; this week it runs +$1,467.8 down to +$1,432.0 -- a property of the week, not evidence about either issue's cards. The M30 histogram appears after the M30 cards. Source: grid_hist_wk.json.*

---

## Why it matters

MarketMania publishes default LLM-council trading configs. This series asks one narrower question every week: how much apparent performance can large-scale in-sample search extract -- and how much of it survives out-of-sample? Issue #1 set the in-sample baseline, issue #2 delivered the first walk-forward verdict (brutal), issue #3 the second (the opposite), issue #4 the third on two windows at once, issue #5 the fourth and issue #6 the fifth, each on the widest board the series had had. Issue #7 delivers the sixth, on a board that carries every generation the series has produced: 43 cards from six weekly issues and the first monthly one, each replayed verbatim on the same two fresh windows. That is the point of the exercise -- a rule that is applied to a growing, never-pruned list is the only walk-forward number that cannot be chosen after the fact.

---

## How we searched

Two frozen collects, run to the same plan with the criteria frozen in writing *before* launch (w276 plan) and printed into the meta of every JSON deliverable: the weekly collect over the WK window, and the night collect, which sweeps the M30 window and the same WK week together in one pass. Each stage below prints **weekly + night** and both totals are printed whole; no calendar-month window is computed this issue. **Every WK number in this issue comes from the weekly collect and every M30 number from the night collect** -- the night run's own second pass over the WK week is used for one thing only, the cross-window stability board, and no number from it is printed beside a weekly-collect number.

1. **1. Scan (s1)** -- 476 + 952 sims across council compositions and coarse settings.
2. **2. Systematic grid (s2)** -- 4,176 + 8,352 sims sweeping the declared axes (archetype, FH/TF, entry filters, universe, council size, membership, TP/SL source, ladder, break-even, TP/SL shift, calibrate and stake) -- the denominator for every distribution claim below.
3. **3. Refinement: ladder + stake sweep (s2b+s3b)** -- 121 + 243 sims around the grid winners.
4. **4. Preset neighbourhood grid (s2p)** -- 53 + 106 sims: one-knob neighbours of the live presets that have a local grid this issue.
5. **5. Per-ticker probes (s5t)** -- 70 + 140 sims: best finalists split onto single coins.
6. **6. Dedicated per-ticker (s6t+s6tc)** -- 585 + 1,165 sims: a compacted independent per-coin grid plus calibrate twins (the standing branch from issue #2).
7. **7. OOS replays (s4cw6+s4cw6o+s4cw5+s4cw5o+s4cw4+s4cw4o+s4cwm1+s4cwm1o+s4cw3+s4cw3o+s4cw2+s4cw2o+s4cw1+s4pre)** -- 78 + 125 sims: every published card and all four live presets, verbatim, on every fresh window, plus origin re-runs of the issue-6 and issue-5 cards for the drift baseline.

*Total 16,642 simulations (5,559 weekly + 11,083 night), 0 errors, 2,577 zero-trade cells (919 + 1,658); stage sums reconcile exactly in each collect (counts.by_stage). The axis grid was NOT widened relative to issue #6 (anti-overfit rule): the replay branch grew by one cohort, the search space did not. Equity series, grid histograms, baselines and the calibrate pack are produced inside the same two collects -- there is no separate same-day extras pass this issue.*

### Stake: universe-aware from simulation #1 ($1,000 deposit)

| Instruments in universe | Stake cap | Sweep values also simulated |
|---|---|---|
| 1 coin | $450 | $300 / $375 |
| 2 coins | $450 | $300 / $375 |
| 3 coins | $300 | $225 / $375 |
| 4 coins | $250 | $200 / $300 |
| 5 coins | $200 | $150 / $250 |

*A $1,000 deposit cannot fund five concurrent $450 positions. Stake is a function of how many instruments the config trades, recomputed on every universe change and swept as its own axis; any simulated stake above the cap for its universe size is marked **concentrated** and excluded from the headline boards (it stays in the open data). Every card also prints **skipped_nofunds** -- entries the engine could not fund. Margin-skip alert rows on the leaders this issue publishes: 4 on WK (weekly collect) and 16 on M30 (night collect); the night collect raised 20 alerts in total, the rest on its own second run of the WK week, which this issue does not publish. Source: criteria in both collects.*

*Models under test (exact engine IDs): claude-fable-5, claude-opus-5, gpt-5.6-sol, deepseek-v4-pro, gemini-3.1-pro, qwen-3.8-max, grok-4.6. The grok axis runs as **grok-4.6 (incl. 4.5-era)** -- one lineage; published grok-4.5 configs replayed out-of-sample get the same 4.5 -> 4.6 lineage remap ([remap] rows).*

---

## TOP councils -- WK window (Sep 14-21, 2026)

| # | Council | Coins | FH/TF | Net | Net % | maxDD | Trades | WR |
|---|---|---|---|---|---|---|---|---|
| **#1** | fable, opus, gpt, gemini | SOL+XRP | 4h/1h | **+$1,467.8** | +146.8% | 7.7% | 60 | 80.0% |
| **#2** | fable, opus, gpt | BTC+ETH | 4h/4h,1h | **+$953.2** | +95.3% | 0.0% | 37 | 91.9% |

| # | Ladder | BE | Shifts | Conf filter | Source | Entry | Positional | Stake |
|---|---|---|---|---|---|---|---|---|
| #1 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | farthest | 3/1 | reenter | $450 |
| #2 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | farthest | 2/0 | reenter | $450 |

*Against the 4,176-cell stage-2 grid on this window: card #1 lands in the +$1,250 to +$1,500 bucket (31 of 4,176 cells) and card #2 in the +$750 to +$1,000 bucket (265 cells); 960 cells in the whole grid finished above +$500 (grid median +$79.2). Card percentiles exported this issue cover the raw net-board leaders and the replay rows, not the mature cards -- no percentile is claimed for the two cards themselves. Net % = net PnL as a percentage of the fixed $1,000 starting deposit; every simulation runs at **10x leverage** (liquidation modeled by the engine). Entry = max_sideways / max_diff_side.*

***Only two councils are published on this window.** All ten slots of the mature net board are the same 2-coin universe (SOL+XRP), so the pairwise-different-universes rule admits exactly one; #2 is the best mature multi-coin council outside that universe on the quality boards -- it comes off the **Calmar** board (+$953.2 on BTC+ETH, R2 0.931, ulcer 0.00, 0.0% maxDD, 0 unfunded entries) -- disclosed rather than padded. Read the cards as two universes deep, not three.*

***On the calibrate axis no card is a calibrated cell**, where issue #6 published 1 of its 2 WK councils as calibrated cells. Card #1 runs the 2-coin table stake of $450.  Neither card carries an unfunded entry.*

**OPEN IN SANDBOX**   **marketmania.ai/s/cw7-wk1**   **marketmania.ai/s/cw7-wk2**

*Each link opens this exact frozen window; results visible without sign-in (embedded share signature). The codes are minted by the publish script before this file goes live; every card's full frozen query signature is in the open-data JSON next to its simulation index.*

![WK published card equity](chart_equity_wk.png)

*Daily settled equity for card #1 above ($1,000 start, 10x leverage, frozen window; #1 red -- fixed rank colors across all issues). A curve running past the window end is an open-at-cutoff trade settling at its horizon (the 2-3 day luft); the curve's last point reconciles to its config's net to the cent. There is no contrast line on this window: the net-board maximum is the same simulation as card #1 (i=1121, 60 closed, above the >=15 maturity gate), so a second curve would be the same curve. Card #2 has no daily series in this issue's equity pack, and the solo and smoothness series stay in the open data (equity.cards) -- the PDF carries one equity chart per window. Source: equity_wk.json.*

---

## TOP councils -- M30 window (Aug 24-Sep 21, 2026)

| # | Council | Coins | FH/TF | Net | Net % | maxDD | Trades | WR |
|---|---|---|---|---|---|---|---|---|
| **#1** | fable, gpt, gemini, qwen | SOL+XRP | 4h/1h | **+$1,351.8** | +135.2% | 69.8% | 145 \* | 66.9% |
| **#2** | opus, gpt, gemini, qwen | BTC+ETH | 1d/1d | **+$528.4** | +52.8% | 9.0% | 26 | 96.2% |

| # | Ladder | BE | Shifts | Conf filter | Source | Entry | Positional | Stake |
|---|---|---|---|---|---|---|---|---|
| #1 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | farthest | 3/1 | reenter | $450 |
| #2 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | farthest | 2/0 | reenter | $450 |

*Against the 4,176-cell stage-2 grid on this window: card #1 is the same simulation as the raw net-board leader (i=3061), so its exported percentile is the card's -- above **100.0%** of the grid; card #2 lands in the +$500 to +$750 bucket (103 cells) with 96.0% of the grid finished below it. Grid median -$83.5, 169 of the 4,176 cells above +$500. Net % = net PnL as a percentage of the fixed $1,000 starting deposit; every simulation runs at **10x leverage** (liquidation modeled by the engine). Entry = max_sideways / max_diff_side.*

***Only two councils are published on this window too, and it is the same universe the week collapsed onto.** 10 of the ten slots of the M30 mature net board are SOL+XRP, so the pairwise-different-universes rule admits exactly one; #2 is the best mature multi-coin council outside that universe on the quality boards (off the **WR** board: +$528.4 on BTC+ETH, R2 0.860, 9.0% maxDD, 0 unfunded entries). Read the cards as two universes deep, not three.*

***On the calibrate axis 1 of the 2 M30 cards are calibrated cells**, where no published WK council is a calibrated cell -- the same axis, read on two windows (Calibrate axis below). Card #1 pays for its +$1,351.8 with a **69.8% drawdown** on 145 closed trades and 1 unfunded entry, card #2 with 9.0% on 26 and 0. The M30 mature pool is 2,989 configs against 2,652 on WK, because four weeks clear the >=20-trade gate that one week does not.*

**OPEN IN SANDBOX**   **marketmania.ai/s/cw7-m301**   **marketmania.ai/s/cw7-m302**

*Each link opens this exact frozen window; results visible without sign-in (embedded share signature). The codes are minted by the publish script before this file goes live; every card's full frozen query signature is in the open-data JSON next to its simulation index.*

![M30 published councils equity](chart_equity_m30.png)

*Daily settled equity for card #1 above ($1,000 start, 10x leverage, frozen window; #1 red -- fixed rank colors across all issues). A curve running past the window end is an open-at-cutoff trade settling at its horizon (the 2-3 day luft); its last point reconciles to the card's net to the cent. There is no contrast line on this window: the net-board maximum is the same simulation as card #1 (i=3061, 145 closed, above the >=20 maturity gate). The solo, Calmar, WR and smoothness series stay in the open data (equity.cards) -- the PDF carries one equity chart per window. Source: equity_night.json.*

---

## Stage-2 grid: where the defaults sit (M30)

![M30 stage-2 grid net PnL histogram](chart_hist_m30.png)

*Net PnL across all 4,176 stage-2 grid sims, M30 window (Aug 24-Sep 21, 2026), universe-aware stake throughout. Grid median **-$83.5**, 25.7% positive against 57.0% on the week (issue #6's month: median -$0.1, 30.3% positive). The spike at $0 is mostly cells whose entry filters produced no trades (738 zero-trade sims on this window). Range -$853.0 to +$1,351.8; 169 cells finished above +$500, against 960 on the week. Red marker = best solo default at table stake; green = the published cards. Source: grid_hist_night.json.*

---

## Walk-forward: the standing OOS verdict

Every config this series has ever published -- all 43 cards from issues #1, #2, #3, #4, #5, #6 and Monthly Config Watch #1 -- replayed VERBATIM (same knobs, same stake as printed, no re-tuning) on both fresh windows, plus the four hand-tuned presets running live on the platform, listed first. The monthly issue's 3 cards are judged by the weekly rule, like every other row. The replays and their origin re-runs cost 78 simulations in the weekly collect and 125 in the night collect, which replays every row on both of its own windows.

**Verdict rule, unchanged since issue #2.** **Survived** = net-positive on BOTH fresh windows; **Faded** = net-negative on at least one. The live presets and the monthly cards are judged by the same rule. New WK is pure out-of-sample for every row -- the week begins exactly at issue #6's right edge (Sep 14). New M30 (Aug 24-Sep 21, 2026) shares 21 of its 28 days with issue #6's own M30 window (shared stretch Aug 24-Sep 13), 14 of its 28 days with issue #5's own M30 window (shared stretch Aug 24-Sep 6), 7 of its 28 days with issue #4's own M30 window (shared stretch Aug 24-30) and 8 of its 28 days with Monthly Config Watch #1's calendar-August window (shared stretch Aug 24-31), so for those four cohorts that column is decay context rather than pure out-of-sample: the same asymmetry issues #3, #4, #5 and #6 disclosed for their own second columns.

![issue #6 cards: home number vs fresh-week replay](chart_oos_cw6.png)

*Issue #6's 5 published cards: the number its own issue printed (grey, in-sample on its own window) against the same configuration replayed verbatim on this week (green, out-of-sample). 1 of the 5 is net-negative on the fresh week; the 2 M30 cards were measured at home on a four-week window and the 3 WK cards were on a one-week window, so the grey bars are not all the same kind of number. Source: oos_cw6_wk.json + issue #6 open data.*

| Config | Family | Coins | Home | New WK (n) | New M30 (n) | Verdict |
|---|---|---|---|---|---|---|
| **preset 4h4** | fable, qwen, deepseek, opus | ETH+SOL | -- live | +$661.8 (18) | -$571.0 (19) \+ | Faded |
| **preset bc2** | gemini, gpt, deepseek | SOL+XRP | -- live | -$225.1 (7) \* \+ | -$195.3 (23) \+ | Faded |
| **preset sol** | deepseek, fable, opus | SOL [conc] | -- live | +$233.4 (15) | +$777.8 (45) | **Survived** |
| **preset xrp** | gemini, gpt, deepseek | XRP [conc] | -- live | +$42.2 (4) \* | -$118.9 (3) \* \+ | Faded |
| cw6-m301 | opus, deepseek, gemini | SOL+XRP | +$1,793.8 | +$992.3 (51) | +$967.7 (160) \+ | **Survived** |
| cw6-m302 | gpt, deepseek, gemini, grok | BTC+ETH | +$547.8 | -$234.2 (9) \* \+ | +$313.5 (41) | Faded |
| cw6-s1 | grok | SOL+XRP | +$404.8 | +$727.4 (45) | -$580.4 (28) \+ | Faded |
| cw6-wk1 | opus, gpt, grok | SOL+XRP | +$609.5 | +$935.6 (58) | +$711.7 (135) \+ | **Survived** |
| cw6-wk2 | fable, opus, deepseek | BTC+ETH | +$135.4 | +$447.8 (40) | +$161.3 (82) \+ | **Survived** |
| cw5-m301 | fable, opus, gpt, gemini | SOL+XRP | +$1,717.1 | -$118.3 (6) \* \+ | -$645.0 (21) \+ | Faded |
| cw5-s1 | gemini | BNB+XRP | +$178.7 | +$494.1 (59) \+ | +$244.0 (130) \+ | **Survived** |
| cw5-wk1 | gpt, gemini | SOL+XRP | +$510.4 | +$547.4 (79) | +$464.7 (205) \+ | **Survived** |
| cw5-wk2 | fable, opus, gpt, gemini, grok | BTC+ETH | +$183.5 | +$352.0 (50) | +$142.4 (98) \+ | **Survived** |
| cw4-m301 | gpt, deepseek, gemini | SOL+XRP | +$1,676.0 | -$123.2 (7) \* \+ | -$138.3 (37) \+ | Faded |
| cw4-m302 | fable, gpt, deepseek, gemini | BNB+XRP | +$1,155.2 | +$687.0 (10) | +$3.6 (24) \+ | **Survived** |
| cw4-s1 | deepseek | BTC+ETH+SOL+BNB+XRP | +$325.0 | +$34.0 (19) \+ | -$118.5 (74) \+ | Faded |
| cw4-wk1 | opus, deepseek, gemini, qwen | BTC+ETH+SOL+BNB+XRP | +$399.4 | +$144.2 (22) \+ | +$153.6 (77) | **Survived** |
| cw4-wk2 | fable, deepseek, qwen | BNB+XRP | +$128.2 | +$425.0 (36) | +$563.9 (100) | **Survived** |
| cwm1-m1 | fable, gemini | SOL+XRP | +$1,618.4 | -$373.2 (7) \* \+ | -$374.4 (31) \+ | Faded |
| cwm1-m2 | fable, gpt, deepseek, gemini, qwen | ETH+SOL+BNB+XRP | +$756.6 | -$43.0 (15) \+ | -$221.0 (44) \+ | Faded |
| cwm1-s1 | gemini | SOL+XRP | +$1,721.9 | +$2.3 (9) \* \+ | -$91.2 (42) \+ | Faded |
| cw3-m301 | opus, deepseek, gemini | SOL+XRP | +$2,137.6 | +$836.7 (45) | -$599.6 (26) \+ | Faded |
| cw3-m302 | fable, opus, deepseek, gemini, qwen | BTC+ETH | +$412.3 | +$88.3 (6) \* | -$102.4 (18) \+ | Faded |
| cw3-s1 | gemini | SOL+XRP | +$1,604.3 | -$644.3 (11) \+ | -$576.2 (64) \+ | Faded |
| cw3-s2 | gemini | SOL+XRP | +$1,449.0 | +$2.3 (9) \* \+ | -$91.2 (42) \+ | Faded |
| cw3-wk1 | opus, deepseek, gemini | SOL+XRP | +$2,175.1 | +$836.7 (45) | -$599.6 (26) \+ | Faded |
| cw3-wk2 | opus, deepseek, gemini, grok | BTC+ETH | +$837.6 | +$500.6 (33) \+ | -$646.8 (14) \+ | Faded |
| cw2-30d1 | opus, gpt, deepseek, gemini, qwen | BTC+ETH+SOL+BNB+XRP [conc] | +$744.7 | -$271.2 (8) \* \+ | +$332.8 (36) \+ | Faded |
| cw2-30d2 | opus, deepseek, gemini | ETH+SOL+BNB+XRP [conc] | +$655.1 | +$514.6 (13) \+ | -$114.8 (34) \+ | Faded |
| cw2-30d3 | opus, deepseek, gemini | BNB+XRP | +$489.1 | +$561.8 (10) | -$156.6 (29) \+ | Faded |
| cw2-30ds1 | gemini | BNB+XRP | +$383.8 | -$209.3 (7) \* \+ | -$586.6 (20) \+ | Faded |
| cw2-7d1 | fable, opus, deepseek, gemini | SOL+XRP | +$299.1 | +$779.1 (39) | -$595.1 (25) \+ | Faded |
| cw2-7d2 | gpt, deepseek, gemini, qwen | BNB+XRP | +$223.7 | +$476.4 (10) | -$35.8 (23) \+ | Faded |
| cw2-pt1 | fable, deepseek, grok [remap] | SOL | +$698.3 | +$93.2 (19) | +$400.8 (58) | **Survived** |
| cw2-pt2 | opus | XRP | +$543.7 | -$560.7 (3) \* \+ | -$651.8 (4) \* \+ | Faded |
| cw1-30d1 | fable, deepseek, opus, qwen | ETH+SOL | +$868.4 | +$750.2 (19) | +$184.2 (45) \+ | **Survived** |
| cw1-30d2 | gemini, gpt, deepseek | SOL+XRP+ETH | +$590.6 | -$119.9 (13) \+ | -$162.3 (44) \+ | Faded |
| cw1-30d3 | fable, deepseek, opus, qwen | BTC+ETH+SOL | +$580.5 | +$653.1 (28) | -$10.9 (65) \+ | Faded |
| cw1-30ds1 | opus | SOL+XRP | +$442.8 | +$1,089.1 (41) | +$777.6 (101) \+ | **Survived** |
| cw1-30ds2 | grok [remap] | SOL+XRP | +$262.9 | +$674.8 (32) | +$145.6 (74) \+ | **Survived** |
| cw1-30ds3 | deepseek | XRP+BNB | +$119.0 | +$849.7 (38) | -$550.5 (27) \+ | Faded |
| cw1-7d1 | fable, deepseek, opus, gpt | SOL+XRP | +$439.2 | +$590.2 (48) | +$759.7 (137) \+ | **Survived** |
| cw1-7d2 | fable, qwen, gemini, opus | XRP+BNB | +$429.5 | +$395.2 (30) \+ | -$621.7 (15) \+ | Faded |
| cw1-7d3 | gpt, opus | BNB+XRP+SOL | +$398.6 | +$559.2 (47) | -$450.5 (92) \+ | Faded |
| cw1-7ds1 | gemini | XRP+BNB | +$417.8 | +$235.0 (67) | -$569.2 (23) \+ | Faded |
| cw1-7ds2 | gpt | XRP+BNB | +$366.3 | +$622.5 (39) | -$98.2 (89) \+ | Faded |
| cw1-7ds3 | qwen | XRP+BNB | +$267.0 | +$412.2 (37) \+ | -$591.7 (20) \+ | Faded |

*n in parentheses; \* = fewer than 10 closed trades (insufficient cell); \+ = non-zero skipped_nofunds (entries the engine could not fund -- the printed economics are then not the economics that ran). Home = the number the card's origin issue printed; the live presets have no home issue. [conc] = stake above this issue's cap for its universe size; [remap] = grok-4.5 replayed on the grok-4.6 lineage. Verdict rule: Survived = net-positive on both fresh windows, Faded = net-negative on at least one. Survivors are tinted green.*

**Counted:** 14 of 43 past cards **Survive** both fresh windows (issue #6 on its two windows: 6 of 38) -- by cohort 3 of 5 issue-6 cards, 3 of 4 issue-5 cards, 3 of 5 issue-4 cards, 0 of 3 Monthly-1 cards, 0 of 6 issue-3 cards, 1 of 8 issue-2 cards, 4 of 12 issue-1 cards. Taken a window at a time: 33 of 43 are positive on the week and 16 of 43 on the four weeks, and the two columns do not nest: a row can clear one and fail the other, so both windows remove cards. The survival curve by generation, each on its own first, second, third, fourth, fifth and sixth fresh week, reads: issue-1 50.0% -> 91.7% -> 8.3% -> 16.7% -> 25.0% -> **91.7%** now (11 of 12); issue-2 100.0% -> 50.0% -> 12.5% -> 25.0% -> **62.5%** now (5 of 8); issue-3 33.3% -> 33.3% -> 33.3% -> **83.3%** now (5 of 6); issue-4 20.0% -> 0.0% -> **80.0%** now (4 of 5); Monthly-1 66.7% -> 0.0% -> **33.3%** now (1 of 3); issue-5 25.0% -> **75.0%** now (3 of 4); issue-6 **80.0%** now (4 of 5). Live presets: 1 of 4 Survive, 3 of 4 positive on the week, 1 of 4 on the month.

*The four live presets lead the table by standing rule. On the week: preset 4h4 +$661.8 on 18 closed, preset bc2 -$225.1 on 7 closed, preset sol +$233.4 on 15 closed, preset xrp +$42.2 on 4 closed -- 3 of the four net-positive. On the four-week window: preset 4h4 -$571.0 on 19 closed, preset bc2 -$195.3 on 23 closed, preset sol +$777.8 on 45 closed, preset xrp -$118.9 on 3 closed -- 1 of four. 1 Survives both. **sol** and **xrp** are marked [conc]: both run a $900 stake on a single coin, twice the $450 cap the stake table sets for a 1-coin universe, so their printed economics assume funding this issue's own rule would not grant. Passports for all four are in the open data (presets_recon_wk.json and presets_recon_night.json).*

***Origin re-runs (drift baseline).** Re-simulating the issue-6 cards on their OWN origin windows today lands within -$217.3..$0.0 of the printed issue-6 numbers and the issue-5 cards within -$497.3..-$99.0 of theirs; 3 of the 9 re-runs reproduce to the cent. The widest deviation is cw5-m301 (+$1,717.1 printed, +$1,219.8 today) -- today's engine reads a forecast history that keeps growing, which is the one thing a verbatim replay cannot freeze. Snapshot principle: no past issue is restated; the full drift table is in the open data (oos_walk_forward.drift_baseline).*

**Honest read.** The fresh search and its own back catalogue moved in the same direction this week. On the fresh week 33 of 43 past cards are net-positive -- issue #6 counted 8 of 38 on its own week -- and the ceiling of the fresh search moved to +$1,467.8 from +$609.5. On the fresh four weeks 16 of 43 are net-positive. Same replay machinery, same frozen knobs; what changed is the window and the size of the board. Three things keep this from being a statement about search. The board is not a fixed sample: it grows every issue and nothing is ever pruned, so a share computed on 43 cards is not comparable with one computed on 38 without saying which generations joined -- 5 of the 43 rows are new this issue. The month column is not clean evidence for four of the seven cohorts: it shares 21 of its 28 days with issue #6's own M30 window (shared stretch Aug 24-Sep 13), 14 of its 28 days with issue #5's own M30 window (shared stretch Aug 24-Sep 6), 7 of its 28 days with issue #4's own M30 window (shared stretch Aug 24-30) and 8 of its 28 days with Monthly Config Watch #1's calendar-August window (shared stretch Aug 24-31), so those cards are partly being scored on their own data. And the size of the in-sample number is not what decides the give-back on this board: the row that lost most on the week is cw3-s1 (-$644.3 on the week against +$1,604.3 printed at home), while the largest home number on the board, cw3-wk1 at +$2,175.1, made +$836.7. Six fresh weeks of verdicts now: brutal, the opposite, split by window, the fourth, the fifth, and this one. That is a series of weathers, not a strategy.

---

## Calibrate axis: does TP/SL calibration help a config?

Matched pairs -- identical knobs, calibration OFF vs ON -- measured on net PnL. The table is the **whole stage-2 grid** of each window, and the **Both** row pools the two windows' pairs and nothing else. The board-leader slice is a second, separate measurement (14 WK pairs and 12 M30 pairs whose OFF or ON leg reached the top of a board); it is reported in the paragraph below and in the open data, and the two are never pooled together.

| Window | Pairs | Calibrate wins | Mean delta | Best delta | Worst delta |
|---|---|---|---|---|---|
| WK | 1,452 | 25.8% | -$111.9 | +$771.2 | -$962.5 |
| M30 | 1,453 | 65.9% | +$253.6 | +$1,772.5 | -$1,158.2 |
| **Both** | **2,905** | **45.9%** | **+$70.9** | **+$1,772.5** | **-$1,158.2** |

**The two windows answer the axis, and the whole-grid reading and the board-leader slice are separate measurements.** On the week calibration improves 25.8% of the 1,452 matched pairs (mean -$111.9, median -$77.6) and 20.4% of the 706 mature pairs (mean -$230.3); on the four weeks it improves 65.9% of 1,453 (mean +$253.6, median +$176.0) and 73.6% of the 726 mature pairs (mean +$284.5). Pooled over both windows the axis improves 45.9% of 2,905 pairs with a mean of +$70.9 -- a number that describes neither window on its own. The board-leader slices are separate: 0.0% of the 14 WK slice pairs improve (mean -$722.7, median -$782.0, OFF legs already at a median of +$1,425.3), and 0.0% of the 12 M30 slice pairs (mean -$421.4, OFF legs at +$1,272.9). Issue #6 measured a 14-pair WK slice (0 improving) and a 13-pair M30 slice (2). Where the axis reached the cards is visible above: 0 of the 2 published WK councils are calibrate=ON cells and 11 of the 62 WK finalists run it, while 1 of the 2 M30 cards are and 25 of the 51 M30 finalists do.

*The open-data pack also carries a 400-pair export (calibrate_effect) per collect. Its delta distribution is not the population's (median -$415.0 against the WK grid's -$77.6), so no claim in this section is computed from it; it is published for inspection only.*

---

## Secondary analysis -- solo-model top

**Sidebar to the council narrative.** Solo configs run inside the same pipeline and face the same gates: >=2 coins, the maturity gate for the window (>=15 closed on WK, >=20 on M30), one config per model.

### Solo top -- WK (Sep 14-21, 2026)

| # | Model | Coins | FH/TF | Net | Net % | maxDD | Trades | WR |
|---|---|---|---|---|---|---|---|---|
| **#1** | gemini | SOL+XRP | 4h/1h | **+$1,022.0** | +102.2% | 5.7% | 61 | 75.4% |
| **#2** | opus | SOL+XRP | 4h/1h | **+$990.4** | +99.0% | 15.0% | 53 | 83.0% |

| # | Ladder | BE | Shifts | Conf filter | Source | Entry | Positional | Stake |
|---|---|---|---|---|---|---|---|---|
| #1 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | median | 0/0 | reenter | $375 |
| #2 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | median | 0/0 | reenter | $450 |

***2 solo cards on this window.** The two qualifying solos do not run the same stake: #1 (gemini) runs $375, a value of the 2-coin stake sweep; #2 (opus) runs the 2-coin table stake of $450. The same knobs as #1 at the $450 table stake (i=4161) make +$1,021.8 on 57 closed with 4 unfunded entries -- one of this window's margin-skip alert rows. No single-coin row outranks them on the mature net board this issue; single-coin results are excluded from the cards by the >=2-coin rule either way and appear under Per-ticker bests. Source: summary_wk.json top_mature_by_window_branch.*

### Solo top -- M30 (Aug 24-Sep 21, 2026)

| # | Model | Coins | FH/TF | Net | Net % | maxDD | Trades | WR |
|---|---|---|---|---|---|---|---|---|
| **#1** | gemini | SOL+XRP | 4h/1h | **+$934.9** | +93.5% | 54.7% | 155 \* | 61.9% |
| **#2** | gpt | SOL+XRP | 4h/4h,1h | **+$918.6** | +91.9% | 35.5% | 172 \* | 67.4% |

| # | Ladder | BE | Shifts | Conf filter | Source | Entry | Positional | Stake |
|---|---|---|---|---|---|---|---|---|
| #1 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | median | 0/0 | reenter | $450 |
| #2 | 50:40,100:60 | on | TP -10 / SL +75 | >=60 (each) | median | 1/0 | reenter | $450 |

***Two solo cards on M30.** The mature solo board carries two models, gemini and gpt, both on SOL+XRP, so the one-config-per-model rule admits two. Unfunded entries on the cards: #1 42, #2 73 (flagged \* on Trades); the best row on the board with 0 skips is gemini on SOL+XRP at +$847.8 on 198 closed, a calibrate=ON cell. Source: top_mature_by_window_branch in the night collect.*

**OPEN IN SANDBOX**   **marketmania.ai/s/cw7-s1**

*Each link opens this exact frozen window; results visible without sign-in (embedded share signature). The codes are minted by the publish script before this file goes live; every card's full frozen query signature is in the open-data JSON next to its simulation index.*

**Council vs solo.** On the week the best solo sits below the best council (+$1,022.0 vs +$1,467.8 on the mature board, 61 closed against 60, 5.7% drawdown against 7.7%). On the four weeks the best solo sits below the best council (+$934.9 vs +$1,351.8, 155 closed against 145, 54.7% against 69.8%). Same in-sample caveats apply to both sides.

---

## Quality boards

Net PnL stays the primary board (continuity with issues #1-#6); four selection views sit on top of it at zero additional simulations. **Calmar-like** = net / max(maxDD, 1.0), positive net only; **WR board** = win rate among mature configs (>=15 closed on WK, >=20 on M30); **Smoothness** = R2 of a linear fit through the daily closed-equity series, mature rows with net > 0 and at least 5 days, ties broken by ulcer index; **Quality composite** = mean rank over net, Calmar-like, win rate and R2, lower is better. Top-3 of each board on each window below; all 15 rows of all eight boards are in the open-data JSON.

| Board | Config | Coins | FH/TF | Net | WR | maxDD | Cls | R2 |
|---|---|---|---|---|---|---|---|---|
| WK Calmar | fable, gpt, gemini, grok | SOL+XRP | 4h/1h | +$1,440.8 | 78.2% | 0.0% | 55 | 0.932 |
| WK Calmar | fable, gpt, gemini, qwen | SOL+XRP | 4h/1h | +$1,437.9 | 82.4% | 0.0% | 51 | 0.902 |
| WK Calmar | fable, gpt, gemini, qwen | SOL+XRP | 4h/1h | +$1,408.0 | 82.0% | 0.0% | 50 | 0.893 |
| WK WR | fable, qwen, deepseek, opus | ETH+SOL | 4h/4h | +$936.2 | 100.0% | 0.0% | 16 | 0.967 |
| WK WR | fable, qwen, deepseek, opus | ETH+SOL | 4h/4h | +$780.2 | 100.0% | 0.0% | 16 | 0.967 |
| WK WR | fable, opus, qwen | ETH | 4h/1h | +$638.8 | 100.0% | 0.0% | 23 | 0.978 |
| WK Smoothness | opus, gpt, qwen | ETH | 4h/1h | +$709.0 | 96.4% | 0.0% | 28 | 0.987 |
| WK Smoothness | fable, qwen | XRP | 4h/1h | +$547.0 | 95.5% | 0.0% | 22 | 0.980 |
| WK Smoothness | fable, opus, qwen | ETH | 4h/1h | +$638.8 | 100.0% | 0.0% | 23 | 0.978 |
| WK Quality | fable, qwen, deepseek, opus | ETH+SOL | 4h/4h | +$936.2 | 100.0% | 0.0% | 16 | 0.967 |
| WK Quality | fable, qwen, deepseek, opus | ETH+SOL | 4h/4h | +$780.2 | 100.0% | 0.0% | 16 | 0.967 |
| WK Quality | fable, opus, gpt | BTC+ETH | 4h/4h,1h | +$953.2 | 91.9% | 0.0% | 37 | 0.931 |
| M30 Calmar | deepseek | ETH | 1d/1d | +$656.9 | 90.0% | 2.8% | 20 | 0.922 |
| M30 Calmar | deepseek | ETH | 1d/1d | +$656.9 | 90.0% | 2.8% | 20 | 0.922 |
| M30 Calmar | deepseek | ETH | 1d/1d | +$437.9 | 90.0% | 1.9% | 20 | 0.922 |
| M30 WR | opus, gpt, gemini, qwen [cal] | BTC+ETH | 1d/1d | +$528.4 | 96.2% | 9.0% | 26 | 0.860 |
| M30 WR | opus, gpt, gemini, qwen [cal] | BTC+ETH | 1d/1d | +$440.4 | 96.2% | 7.5% | 26 | 0.860 |
| M30 WR | opus, gpt, gemini, qwen [cal] | BTC+ETH | 1d/1d | +$352.3 | 96.2% | 6.0% | 26 | 0.860 |
| M30 Smoothness | gemini [cal] | SOL | 4h/1h | +$675.0 | 88.5% | 8.5% | 104 | 0.947 |
| M30 Smoothness | gemini [cal] | SOL | 4h/1h | +$675.0 | 88.5% | 8.5% | 104 | 0.947 |
| M30 Smoothness | fable, gpt, gemini, grok [cal] | SOL | 4h/1h | +$586.5 | 91.1% | 13.3% | 79 | 0.932 |
| M30 Quality | gemini [cal] | SOL | 4h/1h | +$675.0 | 88.5% | 8.5% | 104 | 0.947 |
| M30 Quality | gemini [cal] | SOL | 4h/1h | +$675.0 | 88.5% | 8.5% | 104 | 0.947 |
| M30 Quality | deepseek | ETH | 1d/1d | +$656.9 | 90.0% | 2.8% | 20 | 0.922 |

*[cal] = calibrate=ON cell; the board column names the window. The WK Calmar board is degenerate by construction -- its top row sits at 0.0% maxDD, under the 1.0 floor of net / max(maxDD, 1.0), so the score collapses onto net PnL; read it as 'net among configs that barely drew down'. The M30 board's top row sits at 2.8%, above the floor, so that board ranks by the ratio. The Smoothness boards are the ones that can leave a window's dominant universe: the WK board's top row is an ETH config at R2 0.987 on a net of +$709.0, the M30 board's a SOL config at R2 0.947 on +$675.0. That trade -- straightness bought with size -- is exactly what the board exists to expose. The M30 boards are drawn from a mature pool of 2,989 configs against 2,652 on WK, because four weeks clear a >=20-trade gate that one week does not.*

***Stability, and what it is across this issue.** The board is defined as **mature net top-40 on BOTH windows** of the run that computes it. The run that produced it here is the night collect, whose two windows are **WK and M30** -- the same two windows this issue prints, computed in a second execution of the WK plan -- so this issue's board reads **stable across WK and M30**. **16 configurations qualify.** All of them SOL+XRP, 16 councils and 1 clean (zero unfunded entries on both windows); the top three by rank sum are council fable, gpt, gemini, qwen; council opus, gpt, gemini, qwen; council fable, gpt, gemini, qwen. The shared board is narrower than either window's own: all 20 of the 20 rows exported from its WK board are SOL+XRP; all 20 of the 20 rows exported from its M30 board are SOL+XRP, and no row exported from one window's mature board appears on the other's. The count is not comparable with issue #4's 25 (its board was measured across M30 and the calendar month); the boards measured across WK and M30 were issue #6's, at 9, issue #5's, at 0, and issue #3's, at 21. Full list in the open data (stability).*

---

## Hand-tuned baseline vs the grid

Standing section, fourth issue. One hand-tuned configuration is treated as the reference the search has to beat -- not on net PnL, where any large search wins by construction, but on the three properties a configuration is actually tuned for: **win rate, shallow drawdown and a straight equity line**. The reference (referred to below as the **hand-tuned anchor**) is the live preset 1d_3_llm_BC2_v02, a 3-model 1d council on SOL+XRP at the $450 two-coin stake; it is distinct from the pre-search hand-tuned set A/B/C carried in Baselines below. Its replay row is in the walk-forward table above; here it is the yardstick.

**Beat rule (frozen with the criteria):** a config beats the anchor only if it does so on all three at once -- win rate higher, maxDD lower and smoothness R2 higher -- among mature, non-concentrated search rows. Net PnL is not part of the rule; it is reported next to the winners so the cost of the improvement is visible.

**The anchor on each window**

| Window | Config | Coins | FH/TF | Net | WR | maxDD | R2 | Ulcer | Cls |
|---|---|---|---|---|---|---|---|---|---|
| WK | **Hand-tuned anchor (1d_3_llm_BC2_v02)** | SOL+XRP | 1d/1d | -$225.1 | 57.1% | 27.7% | 0.196 | 13.34 | 7 |
| M30 | **Hand-tuned anchor (1d_3_llm_BC2_v02)** | SOL+XRP | 1d/1d | -$195.3 | 65.2% | 69.0% | 0.626 | 29.04 | 23 |

**2,015 of the 2,652 mature WK configs (76.0%) clear all three bars at once, and 233 of the 2,989 mature M30 configs (7.8%).** The anchor's own two windows are different rows: on the week -$225.1 net at 57.1% win rate with a 27.7% drawdown, an R2 of 0.196 and 7 closed trades; on the four weeks -$195.3 at 65.2%, a 69.0% drawdown, an R2 of 0.626 and 23 closed. The week row is thin against the >=15 gate the search rows must pass (the anchor is exempt from that gate by construction: it is the reference, not a candidate). Read against net PnL, all 10 of the 10 best WK winners by net also beat it on net PnL; all 10 of the 10 best M30 winners by net also beat it on net PnL. The share itself moves with the window (76.0% of the WK field, 7.8% of the M30 field), so read it as one field at a time, not as a property of the anchor.

**Top-10 by net among the 2,015 WK configs that beat the anchor on win rate, maxDD and smoothness at once**

| # | Config | Coins | FH/TF | Stake | Net | WR | maxDD | R2 |
|---|---|---|---|---|---|---|---|---|
| #1 | fable, opus, gpt, gemini | SOL+XRP | 4h/1h | $450 | +$1,467.8 | 80.0% | 7.7% | 0.898 |
| #2 | fable, opus, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,459.2 | 86.7% | 8.1% | 0.889 |
| #3 | fable, gpt, gemini, grok | SOL+XRP | 4h/1h | $450 | +$1,440.8 | 78.2% | 0.0% | 0.932 |
| #4 | fable, gpt, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,437.9 | 82.4% | 0.0% | 0.902 |
| #5 | opus, gpt, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,432.0 | 82.4% | 12.1% | 0.847 |
| #6 | gpt, gemini, grok | SOL+XRP | 4h/1h | $450 | +$1,427.6 | 74.6% | 5.5% | 0.913 |
| #7 | fable, opus, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,426.9 | 86.4% | 8.1% | 0.879 |
| #8 | opus, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,423.8 | 84.8% | 13.6% | 0.848 |
| #9 | opus, gemini, qwen | SOL+XRP | 4h/1h | $450 | +$1,423.8 | 84.8% | 13.6% | 0.848 |
| #10 | fable, opus, gpt, gemini | SOL+XRP | 4h/1h | $450 | +$1,422.6 | 79.7% | 7.7% | 0.888 |

**Top-5 by net among the 233 M30 configs that beat the anchor on win rate, maxDD and smoothness at once**

| # | Config | Coins | FH/TF | Stake | Net | WR | maxDD | R2 |
|---|---|---|---|---|---|---|---|---|
| #1 | opus, gpt, gemini [cal] | SOL+XRP | 4h/1h | $450 | +$1,201.6 | 83.6% | 20.4% | 0.857 |
| #2 | opus, gpt, gemini [cal] | SOL+XRP | 4h/1h | $450 | +$1,165.0 | 83.4% | 20.4% | 0.848 |
| #3 | opus, gpt, gemini [cal] | SOL+XRP | 4h/1h | $450 | +$1,113.3 | 86.1% | 28.9% | 0.784 |
| #4 | fable, gpt, gemini [cal] | SOL+XRP | 4h/1h | $450 | +$1,075.4 | 84.6% | 31.2% | 0.726 |
| #5 | gpt, gemini, grok [cal] | SOL+XRP | 4h/1h | $450 | +$1,059.8 | 84.9% | 20.3% | 0.814 |

*Anchor passport (verbatim, as stored): members=gemini-3.1-pro,gpt-5.6-sol,deepseek-v4-pro tokens=SOL,XRP fh=1d tf=1d conf_mode=each min_conf=60 max_sideways=2 max_diff_side=0 tp_sl_source=nearest trade_mode=position same_side=update opposite=reverse no_signal=hold time_stop=3x steps=40:80,100:20 steps_on=1 be=1 tp_shift=30 sl_shift=50 deposit=1000 lev=10 refill=0 stake=450 calibrate=- variant=v0*

---

## Per-ticker bests

Which coin was extractable this window, and by what. Single-coin results are excluded from the council cards by rule (idiosyncratic-coin risk) and reported here, never ranked against the cards. The dedicated per-coin branch contributed 585 + 1,165 sims -- a compacted grid, one best per coin per window; the branch tag in brackets says which kind won.

### WK (Sep 14-21, 2026)

| Coin | Best config (branch) | FH/TF | Net (trades) |
|---|---|---|---|
| BTC | opus, gpt, grok [council] | 4h/1h | +$425.9 (25) |
| ETH | gemini, grok [council] | 4h/4h | +$847.8 (27) |
| SOL | fable, opus, gpt, gemini [council] | 4h/1h | +$803.7 (27) |
| BNB | gpt, deepseek [council] | 1d/1d | +$436.5 (6) \* |
| XRP | gpt, gemini, grok [council] | 4h/1h | +$841.3 (37) |

### M30 (Aug 24-Sep 21, 2026)

| Coin | Best config (branch) | FH/TF | Net (trades) |
|---|---|---|---|
| BTC | deepseek [solo] | 1d/1d | +$138.1 (1) \* |
| ETH | deepseek [solo] | 1d/1d | +$656.9 (20) |
| SOL | fable, opus, gpt, gemini [council] | 4h/1h | +$874.0 (82) |
| BNB | opus, deepseek, grok [council] | 4h/1h | +$521.9 (58) |
| XRP | fable, gpt, gemini, qwen [council] | 4h/1h | +$854.5 (79) |

*n in parentheses; \* = fewer than 10 closed trades (thin -- direction only). 0 of the ten winners run calibrate ON; 5 of the five WK winners and 3 of the five M30 winners are councils.*

**The reads.** On the week ETH leads at +$847.8 and 1 of the five rows is thin -- 6 to 37 closed trades against a 15-trade gate. On the four weeks 1 of the five rows is thin: 1 to 82 closed, SOL leading at +$874.0, and no one coin leads both windows. The 1d horizon wins 1 of the five WK coins and 2 of the five M30 coins. **Winner's curse applies to this whole table**: each row is the maximum of a per-coin grid, biased high by selection alone, and none has an out-of-sample read until next issue.

---

## Baselines

Three independent reference points, none of them search output, re-simulated fresh on both of this issue's frozen windows inside the same two collects (engine 1.1): the engine solo default for each of the 7 models, the default council-of-7, and the pre-search hand-tuned set A/B/C carried since issue #1. The tables show the universe-aware **table-stake** run -- this issue's canonical mode ($200 for the 5-coin defaults; A and B already at their $300 3-coin cap; C moves $400 -> $450 on 2 coins) -- with the best and worst of the seven solo defaults on each window; all eleven rows per window and both stake modes are in the open data.

### WK (Sep 14-21, 2026)

| Baseline | Net PnL | Trades | WR | maxDD |
|---|---|---|---|---|
| Best solo default (claude-fable-5) | +$73.7 \* | 299 | 73% | 1.5% |
| Worst solo default (gemini-3.1-pro) | +$26.1 \* | 166 | 61% | 1.8% |
| Council-of-7 default | +$34.9 \* | 208 | 70% | 1.8% |
| **Hand-tuned config A (pre-search)** | -$213.2 \* | 11 | 55% | 51.7% |
| **Hand-tuned config B (pre-search)** | +$687.0 | 54 | 80% | 4.6% |
| **Hand-tuned config C (pre-search)** | +$750.2 | 19 | 95% | 6.7% |

### M30 (Aug 24-Sep 21, 2026)

| Baseline | Net PnL | Trades | WR | maxDD |
|---|---|---|---|---|
| Best solo default (claude-opus-5) | -$48.9 \* | 699 | 64% | 13.0% |
| Worst solo default (deepseek-v4-pro) | -$122.5 \* | 1179 | 61% | 16.1% |
| Council-of-7 default | -$55.3 \* | 527 | 61% | 8.5% |
| **Hand-tuned config A (pre-search)** | -$700.6 \* | 20 | 50% | 105.8% |
| **Hand-tuned config B (pre-search)** | -$113.1 \* | 118 | 61% | 56.5% |
| **Hand-tuned config C (pre-search)** | +$184.2 \* | 45 | 69% | 60.0% |

*\* = non-zero skipped_nofunds at table stake. On the 5-coin defaults those counts are the stake table doing its job: at $200 a $1,000 deposit funds at most five concurrent positions while the defaults fire hundreds of entries per window (132 to 430 skipped each on the week, 471 to 1,346 on the four weeks). The other five solo defaults run between the two printed rows on each window. **The conclusions do not change between modes**: on WK for every row but one, printed as it stands: solo default deepseek-v4-pro is -$4.4 at the config's own stake and +$27.5 at the table stake, and on M30 for every row but one, printed as it stands: solo default gemini-3.1-pro is +$68.2 at the config's own stake and -$70.8 at the table stake. 2 of the three hand-tuned configs are already at their own table cap, so their two modes are identical on both windows. Source: baselines_wk.json (22 sims) and baselines_night.json (44 sims, of which 22 are the M30 half), inside this issue's two collects.*

**The engine defaults made money on WK and lost money on M30; the hand-tuned set split on both windows.** On WK the seven solo defaults run +$73.7 to +$26.1 at table stake, the council-of-7 +$34.9, and the three hand-tuned configs A -$213.2, B +$687.0, C +$750.2 -- 2 of the three positive. On M30 the defaults run -$48.9 to -$122.5 solo (council-of-7 -$55.3) and the three hand-tuned configs A -$700.6, B -$113.1, C +$184.2 -- 1 of the three positive. The upper half of the standing ordering (search > hand-tuned > defaults) holds on both windows, in-sample as ever -- the fresh search makes +$1,467.8 on the week against a best baseline of +$750.2, and +$1,351.8 on the four weeks against +$184.2. The live presets are a separate row of evidence: 3 of the four are net-positive on the week and 1 of four on the month (walk-forward table above), and 2 of them run stakes above this issue's cap.

### Live presets & preset neighbourhoods

Four hand-tuned configurations trade live on the platform (not search output, distinct from the pre-search A/B/C set): preset bc2 = '1d_3_llm_BC2_v02'; preset 4h4 = '4h_4_llm_v02'; preset sol = 'Sol'; preset xrp = 'XRP'. Their replay rows lead the walk-forward table; passports and window detail are in the open data (presets_recon_wk.json and presets_recon_night.json). The neighbourhood grid asks a narrower question: does a one-knob neighbour beat the preset as configured? Two of the four have a local grid this issue (53 cells per window); the two single-coin presets do not (0 cells -- the local grid is not built for them, so no neighbour claim is made about either).

| Live preset | Window | As configured | Best neighbour in local grid | Cells | Verdict |
|---|---|---|---|---|---|
| preset bc2 | WK | -$225.1 | +$807.8 (4h/1h, SOL+XRP, 38 cls) | 34 | neighbour +$1,032.9 |
| preset bc2 | M30 | -$195.3 | +$244.7 (1d/1d, SOL+XRP, 21 cls) | 34 | neighbour +$440.0 |
| preset 4h4 | WK | +$661.8 | +$936.2 (4h/4h, ETH+SOL, 16 cls) | 19 | neighbour +$274.5 |
| preset 4h4 | M30 | -$571.0 | +$213.6 (1d/1d,4h, ETH+SOL, 9 cls) | 19 | neighbour +$784.7 |

In **4 of the 4** preset-window cells with a grid, a one-knob neighbour beat the configuration as it is actually running. The largest single edit is preset bc2's 1d/1d -> 4h/1h horizon move on WK, which turns -$225.1 into +$807.8. The presets are not at a local optimum -- a testable, low-risk edit, unlike adopting a search champion wholesale. Cell counts are small (34 and 19 per window) and every neighbour is an in-sample maximum of its own little grid.

---

## Patterns: what the finalists look like

Knob modes across each window's unique finalists (mature net top-10 per branch plus all four quality boards, deduplicated by full signature; 62 unique configs on WK, 47 of them councils; 51 on M30, 31 councils).

| Knob | WK finalists mode | M30 finalists mode |
|---|---|---|
| Forecast horizon | 4h (62/62) | 4h (30/51) |
| Ladder shape (steps) | 50:40,100:60 (58/62) | 50:40,100:60 (51/51) |
| Break-even stop (BE) | on (62/62) | on (51/51) |
| Min confidence | 60 (62/62) | 60 (51/51) |
| SL shift | 75 (58/62) | 75 (51/51) |
| TP shift | -10 (58/62) | -10 (51/51) |
| TP/SL source | farthest (54/62) | farthest (39/51) |
| Calibrate ON | 11/62 | 25/51 |
| Universe | SOL+XRP (29/62) | SOL+XRP (20/51) |

The mechanical core is the same on both windows and unchanged for the seventh issue running: 50:40,100:60 ladder, break-even on, conf 60, SL +75 / TP -10, farthest source. The horizon is **4h** on WK (62/62) and **4h** on M30 (30/51). **Calibrate=ON** runs in 11 of 62 WK finalists against 25 of 51 on M30, where issue #6 measured 21/54 on its week and 34/53 on its month. The universe: SOL+XRP takes 29 of the 62 WK finalists while SOL+XRP takes 20 of the 51 M30 finalists. Membership concentrates too -- fable leads the WK council finalists (39 of 47) and gemini the M30 ones (28 of 31). Standing caveats: refinement seeds from the same grid winners, so part of the convergence is search-design echo; and a knob core that produced 6-of-38 survival one issue ago and 14-of-43 this issue is describing the weather, not a recommendation.

---

## Liquidation note

Liquidation is modeled by the engine at the fixed 10x leverage used throughout. Across the published cards, stop-loss placement stays inside the distance that would approach the liquidation threshold; maxDD is printed on every config so realized risk is visible directly. This issue the published cards draw down 7.7% and 0.0% on the week and 69.8% and 9.0% on the month, and the replay table is deeper: 12 replayed rows drew down more than 40% on the week and 38 on the month, 4 and 28 of them past 55%. A high headline net and a survivable path are different claims, and so are a shallow in-sample drawdown and a shallow one next week.

---

## Watch amendments (this issue)

> **Amendment 3 (standing) -- smoothness and a quality composite.** Every simulation stores its daily closed-equity series and the three numbers read off it (n_days, smooth_r2, ulcer); the smoothness board ranks mature rows with net > 0 and at least 5 days by R2 with ulcer as tie-break, and the quality composite ranks by the mean of the net, Calmar-like, win-rate and R2 ranks. Both still cost zero additional simulations and the headline board is still net PnL.

> **Amendment 4 (standing) -- the monthly split.** The weekly Config Watch keeps two windows, the calendar week (WK) and the four calendar weeks that end at the same edge (M30, Aug 24-Sep 21, 2026); the calendar-month board belongs to the Monthly Config Watch line and no calendar-month number is printed in a weekly issue. This issue's night collect computes no calendar-month window at all: Monthly Config Watch #2 covers September and is produced after 01.10.

> **Amendment 5 (standing) -- the standing OOS board grows with every issue.** Every card this series publishes joins the walk-forward table and is never removed: this issue replays **43 cards** -- issue #6's 5 join issue #5's 4, issue #4's 5, Monthly Config Watch #1's 3 and issues #1, #2 and #3's 26 -- plus the four live presets, on both fresh windows. **Monthly cards are judged by the weekly rule** (net-positive on both of this issue's fresh windows), the same rule every other row is judged by; their own calendar-month origin window is the Home column and nothing else. One consequence for the counts: a survival share is computed on a board that changes size every issue, so it is reported with the cohort breakdown beside it and never as a trend on its own. **Stability** is defined inside the run that computes it: this issue's night collect computes WK and M30, so the board reads across WK and M30 -- the issue-#3 definition, the one issues #5 and #6 used as well -- where issue #4's read across M30 and the calendar month. 

All three amendments are standing, provisional and scoped to this series. **Formalization is slated for methodology v1.2**; until then this issue runs on v1.1 plus the Config Watch amendments approved 12.08 (hash e66c7e8c864a2233).

---

## Limitations

- **Two windows, one regime, and they overlap.** M30 (Aug 24-Sep 21, 2026) contains WK (Sep 14-21, 2026) -- 7 of the week's 7 days are inside it -- and it shares 21 of its 28 days with issue #6's own M30 window (shared stretch Aug 24-Sep 13), 14 of its 28 days with issue #5's own M30 window (shared stretch Aug 24-Sep 6), 7 of its 28 days with issue #4's own M30 window (shared stretch Aug 24-30) and 8 of its 28 days with Monthly Config Watch #1's calendar-August window (shared stretch Aug 24-31). The two columns of this issue are therefore not two independent tests: agreement between them is partly arithmetic, and for the issue-6, issue-5, issue-4 and Monthly-1 cohorts the M30 replay column is decay context rather than pure out-of-sample. Do not read them as a cross-validation.
- **The two windows come from two collects, and the two runs of the same week agree on the ceiling, not on the whole grid.** WK is the weekly run (generated 2026-09-22T11:52:38Z), M30 the night run (2026-09-23T03:51:55Z -- boards computed at 03:33:47Z, files written after the baselines pass), which recomputed the same WK week in the same pass on its own grid: the night run ranks the same mature WK top-10 in the same order, so its WK ceiling is the same +$1,467.8 (fable, opus, gpt, gemini on SOL+XRP), but its WK grid median is +$78.8 against the weekly collect's +$79.2, its WK Smoothness leader is R2 0.980 on +$547.0 against R2 0.987 on +$709.0, and 2,010 of 2,653 mature WK rows beat the anchor there against 2,015 of 2,652 here. Every WK number this issue prints is the weekly collect's; the night run's WK half is used only for the cross-window stability board. Same engine, same criteria, same collect code, but not the same execution -- a number is comparable across the two windows only as far as that is.
- **Stability is defined inside the run that computes it.** The board comes from the night run, whose two windows are WK and M30, so it reads across those two -- and 16 configurations are in the mature net top-40 of both. Issue #6 read the same pair one week earlier and held 9, issue #5 the week before that held 0 and issue #3's board (the other one across the same pair) held 21; issue #4's 25 is not comparable at all, having been measured across M30 and the calendar month. A board this size is one observation per issue and no claim is made from a single reading.
- **Model lineage is a splice.** The grok axis runs as grok-4.6 (incl. 4.5-era) -- one lineage across a mid-series model cutover dated 2026-08-24 in the search plan both collects ran to. This issue's WK window (Sep 14-21, 2026) lies entirely after that cutover; the M30 window (Aug 24-Sep 21, 2026) contains it. Issue-1 and issue-2 configs that named grok-4.5 are replayed on the 4.6 lineage ([remap] rows: cw2-pt1, cw1-30ds2), and a remapped replay is not the same simulation the origin issue ran.
- **Concentrated rows are published, not hidden.** 4 rows run stakes above this issue's cap for their universe size and are marked [conc] (preset_sol, preset_xrp, cw2-30d1, cw2-30d2); their printed economics assume funding the current rule would not grant. They stay in the table because removing them would flatter the preset row.
- **The stake rule changes what runs, not only its size.** Capping the stake changes WHICH entries get funded: the seven table-stake solo defaults skip 132 to 430 entries each, and 20 replay rows carry unfunded entries (up to 53 on cw3-s1). Where skipped_nofunds is non-zero the printed economics are not the economics that ran; the rows are flagged, never silently pooled.
- **Snapshot principle.** This report is a frozen snapshot taken at the generated-at timestamps; numbers are not updated retroactively and past issues are not restated. The drift baseline exists precisely because today's engine and a longer forecast history do not reproduce every past number (cw5-m301: +$1,717.1 printed, +$1,219.8 today).
- **Research-to-date counter is pinned.** The counter below is pinned from wave 3 onward and is read at the Sep 21, 2026 16:00 UTC cut-off; issue #6 printed 47,373 resolved and 30 published reports against 52,971 and 34 here. The model line keeps the definition issue #4 introduced (7 tracked, current line-up; earlier versions folded into their successors' lineage), so the four issues' model counts are comparable and issue #3's is not.
- **No calendar month in this issue.** The night collect computes M30 and WK and nothing else, so no claim about a calendar month is made anywhere here. Monthly Config Watch #2 covers September and is produced after 01.10; Monthly Config Watch #1's 3 cards appear in this issue only as replay rows, judged by the weekly rule.
- **Multiple testing.** 16,642 simulations across the two collects (5,559 weekly + 11,083 night); at this scale some winners are expected from chance alone. Antidotes: axes frozen before launch, the walk-forward table, independent baselines, in-sample labeling. No formal correction yet.
- **Winner's curse and thin cells.** Every card, board row and per-ticker best is the maximum of a search, biased high by selection alone. On the week the per-ticker winners closed 6-37 trades and 12 of the 47 replay rows are under 10 closed; on the four weeks 2 replay rows are thin and 1 per-ticker winner is. Flagged with \*, reported for completeness.
- Research output, not financial advice.

---

## Counters & lineage

**Weekly collect (WK window)**

| Stage | Sims |
|---|---|
| Scan (s1) | 476 |
| Systematic grid (s2) | 4,176 |
| Refinement: ladder + stake sweep (s2b+s3b) | 121 |
| Preset neighbourhood grid (s2p) | 53 |
| Per-ticker probes (s5t) | 70 |
| Dedicated per-ticker (s6t+s6tc) | 585 |
| OOS replays (s4cw6+s4cw6o+s4cw5+s4cw5o+s4cw4+s4cw4o+s4cwm1+s4cwm1o+s4cw3+s4cw3o+s4cw2+s4cw2o+s4cw1+s4pre) | 78 |
| **Total** | **5,559** |
| Errors / zero-trade | 0 / 919 |
| Engine | 1.1 |

**Night collect (M30 and a second execution of the WK week; this issue publishes M30 from it, and its WK half feeds the cross-window stability board only)**

| Stage | Sims |
|---|---|
| Scan (s1) | 952 |
| Systematic grid (s2) | 8,352 |
| Refinement: ladder + stake sweep (s2b+s3b) | 243 |
| Preset neighbourhood grid (s2p) | 106 |
| Per-ticker probes (s5t) | 140 |
| Dedicated per-ticker (s6t+s6tc) | 1,165 |
| OOS replays (s4cw6+s4cw6o+s4cw5+s4cw5o+s4cw4+s4cw4o+s4cwm1+s4cwm1o+s4cw3+s4cw3o+s4cw2+s4cw2o+s4cw1+s4pre) | 125 |
| **Total** | **11,083** |
| Errors / zero-trade | 0 / 1,658 |
| Engine | 1.1 |

***Scale note (vs issue #6).** Issue #6 ran 16,111 sims over two windows in two collects; issue #7 runs 16,642 over two windows in two -- 5,559 in the weekly collect (WK) and 11,083 in the night collect (M30 and the same WK week together). The per-window grid is close to unchanged: 4,176 WK cells and 4,176 M30 cells here against 3,952 WK and 4,080 M30 cells there. The per-coin branch is 585 and 1,165 against 560 and 1,150. The replay branch grew in rows (47 configs against 42) and in sims (78 and 125 against 68 and 110), because one more cohort joined and every row is replayed on every window of its collect. The two selection boards still cost zero simulations.*

***No extras pass in either collect.** Equity series, the grid histograms (0 extra sims), the baselines and the calibrate packs are produced by the same two collects that wrote the boards. The baseline runs are verification, outside both search totals: 22 sims on WK (11 configs x 2 stake modes) and 44 in the night collect, of which the 22 M30 rows belong to this issue. 0 equity replays were needed.*

Generated at: 2026-09-22T11:52:38Z (weekly collect) and 2026-09-23T03:51:55Z (night collect)  ·  Methodology v1.1 + Config Watch amendments (approved 12.08), hash e66c7e8c864a2233.

**Snapshot principle:** this report is a frozen snapshot of the search taken at the generated-at timestamps above; numbers are not updated retroactively. Each issue re-runs the pipeline fresh over that issue's windows.

---

## Research to date

> **RESEARCH SNAPSHOT** -- THIS REPORT -- search effort: **16,642 simulations** (0 errors) across two frozen collects, of which 203 walk-forward replays and origin re-runs · 66 verification sims (baseline runs inside the same collects; 0 equity replays needed)

MARKETMANIA RESEARCH TO DATE (as of cutoff): 52,971 directional forecasts resolved since Jul 11 (as of Sep 21, 2026 16:00 UTC cutoff) · 7 models tracked (current line-up; earlier versions folded into their successors' lineage) · 5 assets · 5 horizons · hourly · 34 published reports

*Counter pinned from wave 3 onward and read at the Sep 21, 2026 16:00 UTC cut-off: issue #6 printed 47,373 resolved and 30 published reports. The model line keeps the definition issue #4 introduced (7 tracked, current line-up; earlier versions folded into their successors' lineage), so the four issues' model counts are comparable. Config Watch #7 is the 34th published report.*

> **Issue #7.** The standing walk-forward board reaches 43 cards -- issue #6's 5 join it, and monthly cards are judged by the weekly rule (Amendment 5). Two fresh windows again, the calendar week and the four calendar weeks ending at the same edge, both computed in the night collect as well as the week: the table keeps 14 of 43 past cards where the week alone keeps 33 and the month alone 16. The cross-window stability board is measured on this issue's own two windows (WK and M30) and comes back with **16** configurations. Formalization of the amendments is still slated for methodology v1.2.

---

## What we're testing next

- **The first fresh week of this issue's own cards.** Config Watch #8 replays the 5 cards published here that carry a short link (both WK councils, WK solo #1, both M30 councils; WK solo #2 and both M30 solo cards have no link and are not replayed, as issue #6's two unlinked solo cards were not) on its own fresh week, alongside the 43 already on the board. The number to check against: 80.0% of issue #6's 5 linked cards were net-positive on their first fresh week, which is this issue's week.
- **Whether the WK-and-M30 stability board reads the same twice.** 16 configurations are in the mature net top-40 of both windows of this issue's night run, against 9 on issue #6's board, 0 on issue #5's and 21 on issue #3's, all across the same pair. Config Watch #8's night run computes the same pair one week on, so the question is whether the count moves when the windows do.
- **What the calibrate share reads on a fifth window.** Calibration improves 25.8% of the 1,452 matched WK pairs and 65.9% of the 1,453 M30 pairs this issue, against 49.9% and 48.4% in issue #6. The test is the same two measurements on the next pair of windows -- the whole grid and the board-leader slice, reported separately and never pooled.

---

## Related research

- Consensus Watch #7 -- https://marketmania.ai/research/reports/consensus-watch-2026-09-14.pdf
- Weekly Calibration #7 -- https://marketmania.ai/research/reports/weekly-calibration-2026-09-14.pdf
- Weekly Model Watch #7 -- https://marketmania.ai/research/reports/model-watch-2026-09-14.pdf
- Config Watch #6 -- https://marketmania.ai/research/reports/config-watch-2026-09-09.pdf
- Config Watch #5 -- https://marketmania.ai/research/reports/config-watch-2026-09-02.pdf
- Monthly Config Watch #1 -- https://marketmania.ai/research/reports/config-watch-monthly-2026-08.pdf

*The three weekly reports publish together as one issue each week; Config Watch follows its own cycle. Market-regime figures, where this issue refers to them, come from that wave and are not re-tabulated here. Monthly Config Watch #1 is listed because its 3 cards are replayed in the walk-forward table above.*

---

## Cite this report

```bibtex
@techreport{mm_configwatch_2026w39,
  title        = {Config Watch #7},
  author       = {{MarketMania Research}},
  institution  = {MarketMania},
  year         = {2026},
  month        = sep,
  day          = {23},
  type         = {Weekly Research Report},
  series       = {Config Watch},
  number       = {7},
  note         = {Methodology v1.1 + Config Watch amendments, hash e66c7e8c864a2233},
  url          = {https://marketmania.ai/research/reports/config-watch-2026-09-16.pdf}
}
```
