claude-fable-5
Research profile — published report figures only.
Tracked since Aug 3-9, 2026 UTC, first published 10 Aug 2026.
Live benchmark page →All modelsResearch
Research benchmark — not financial or investment advice; paper trading only.
Where Claude Fable 5 finished in the most recent published window — the figures the issue printed, not a live reading.
Weekly Model Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC · 46.7% is 0.1pp below the all-model figure of 46.8% for the same window
Hit rate follows the benchmark's TP/SL/expiry outcome rules; it is not simply the asset's price direction at the end of the forecast horizon. The weekly and monthly metrics shown here use the 1h, 4h and 1d horizons; stability runs have their own test coverage.
Monthly Model Watch #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
Stated confidence against realised accuracy. The pooled all-model row of the same table is printed next to Claude Fable 5's, because a gap only means something against the population it was measured in.
| Scope | n | Hit rate | Mean stated conf. | Gap | Brier | 95% Wilson |
|---|---|---|---|---|---|---|
| Weekly Calibration #8Claude Fable 5 | 736 | 46.7% | 61.4% | +14.7pp | 0.2692 | [43.2%, 50.3%] |
| Weekly Calibration #8field, all models | 4,958 | 46.8% | 63.0% | +16.1pp | 0.2751 | — |
| Monthly Calibration #1Claude Fable 5 | 2,812 | 46.7% | 60.6% | +13.9pp | 0.2689 | [44.9%, 48.6%] |
| Monthly Calibration #1field, all models | 19,383 | 46.3% | 62.6% | +16.3pp | 0.2784 | — |
Weekly Calibration #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC — Monthly Calibration #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
| Stated confidence bucket | n | Hit rate | Mean stated conf. |
|---|---|---|---|
| 50-60 | 199 | 37.7% | 56.9% |
| 60-70 | 536 | 50.2% | 63.1% |
| 70-80 | 1 | n < 10 | n < 10 |
| 80-100 | 0 | no data | no data |
Buckets as Weekly Calibration #8 printed them; a cell the issue marked insufficient or empty keeps its count and prints no rate.
How much Claude Fable 5's answer moves when nothing else does: identical payloads replayed several times, counted as unanimous sets and as flips. The caption on each tile is the same measurement in the comparison run shipped with the issue.
Stability Index — Run 2 · 10 Aug 2026 · 945 calls, 5 repeats per set, 27 sets per model, 3 tickers · compared against baseline (Jul 31)
Claude Fable 5 split by forecast horizon, as the Calibration issues print it. Only the horizons inside the reports' scoring gate appear here; a cell the builder annotated keeps its annotation.
| Horizon | Weekly n | Weekly hit rate | Monthly n | Monthly hit rate |
|---|---|---|---|---|
| 1H | 464 | 48.1% | 1,776 | 46.5% |
| 4H | 220 | 44.1% | 861 | 46.2% |
| 1D | 52 | 46.2% | 175 | 51.4% |
Weekly Calibration #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC — Monthly Calibration #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
The Model Watch issues print a best-pairs and a worst-pairs table — the model-instrument cells that stood out in the window. Only Claude Fable 5's own rows are listed; a model that made neither list simply has no row there.
| Instrument | List | n | Hit rate | Issue |
|---|---|---|---|---|
| BNB | worst pairs | 114 | 36.0% | Weekly Model Watch #8 |
| SOL | best pairs | 604 | 48.5% | Monthly Model Watch #1 |
| ETH | worst pairs | 529 | 44.0% | Monthly Model Watch #1 |
Weekly Model Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC — Monthly Model Watch #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
Weekly Model Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC
Forecasts are grouped by whether their direction matched or opposed the leave-one-out majority of the other eligible model lines. Claude Fable 5 is scored separately on each group.
| Scope | Agree n | Agree hit rate | Disagree n | Disagree hit rate | Lift | Thin or tied |
|---|---|---|---|---|---|---|
| Consensus Watch #8 | 620 | 47.3% | 2 | 0.0% | +47.3pp | 114 |
| Monthly Consensus Watch #1 | 2,391 | 45.8% | 13 | 53.8% | -8.0pp | 408 |
leave-one-out: matched vs opposed the majority of the other model lines · Consensus Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC — Monthly Consensus Watch #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
What a mechanical paper-trading rule made of Claude Fable 5's forecasts in the same windows. These figures are kept apart from the accuracy and calibration figures above: a model can be well calibrated and unprofitable, or profitable and badly calibrated.
| Scope | Trades | Win rate | Net PnL | Gross PnL | Max drawdown |
|---|---|---|---|---|---|
| Weekly Calibration #8 | 736 | 39.4% | -$45.60 | $28.00 | -$127.29 |
| Monthly Calibration #1 | 2,812 | 36.2% | -$96.55 | $184.65 | -$210.11 |
Weekly Calibration #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC — Monthly Calibration #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC
Every window Claude Fable 5's slot has been published in, oldest first. Model Watch supplies the count, the hit rate, the interval and the rank; Calibration adds Brier, the gap and the mean stated confidence. The two cadences are separate tables because a month and a week are different windows over the same forecasts, and a window an older issue also filed under a superseded id keeps that reading on its own line: no issue published the two as one figure.
| Issue | Window | n | Hit rate | 95% Wilson | Rank | Brier | Gap | Mean stated conf. |
|---|---|---|---|---|---|---|---|---|
| #1 · Model Watch · Calibration | Aug 3-9, 2026 UTC | 642 | 41.4% | [37.7%, 45.3%] | #7 | 0.2811 | +19.2pp | 60.7% |
| #2 · Model Watch · Calibration | Aug 10-16, 2026 UTC | 472 | 41.7% | [37.4%, 46.2%] | #6 | 0.2782 | +18.1pp | 59.8% |
| #3 · Model Watch · Calibration | Aug 17-23, 2026 UTC | 779 | 56.5% | [53.0%, 59.9%] | #2 | 0.2464 | +4.7pp | 61.2% |
| #4 · Model Watch · Calibration | Aug 24-30, 2026 UTC | 651 | 43.0% | [39.3%, 46.8%] | #3 | 0.2799 | +17.9pp | 60.9% |
| #5 · Model Watch · Calibration | Aug 31-Sep 6, 2026 UTC | 628 | 42.8% | [39.0%, 46.7%] | #3 | 0.2795 | +17.9pp | 60.7% |
| #6 · Model Watch · Calibration | Sep 7-13, 2026 UTC | 624 | 36.5% | [32.9%, 40.4%] | #6 | 0.2921 | +23.6pp | 60.1% |
| #7 · Model Watch · Calibration | Sep 14-20, 2026 UTC | 799 | 52.3% | [48.9%, 55.8%] | #2 | 0.2593 | +9.1pp | 61.4% |
| #8 · Model Watch · Calibration | Sep 21-27, 2026 UTC | 736 | 46.7% | [43.2%, 50.3%] | #5 | 0.2692 | +14.7pp | 61.4% |
| Issue | Window | n | Hit rate | 95% Wilson | Rank | Brier | Gap | Mean stated conf. |
|---|---|---|---|---|---|---|---|---|
| #1 · Model Watch · Calibration | Aug 1-31, 2026 UTC | 2,812 | 46.7% | [44.9%, 48.6%] | #3 | 0.2689 | +13.9pp | 60.6% |
Each figure is the one its issue froze.
Quoted from the issues these figures come from.
Methodology v1.1 (2026-08-10) · hash e66c7e8c864a2233 · Methodology · Benchmarks methodology · Dataset