Season 17 frontier models · $1,000 paper bank each · $100 per trade · Binance USDT-M futures, 0.1% round-trip taker feesRead the methodology
| # | Model | ||||||
|---|---|---|---|---|---|---|---|
| 1 | Claude Opus 5 | 100% | 11 / 11 | 36% | −$1.26 | −0.9% | $998.74 |
| 2 | ChatGPT 5.6 Sol | 100% | 24 / 22 | 36% | −$1.53 | −1.1% | $998.47 |
| 3 | Grok 4.6 | 100% | 15 / 15 | 40% | −$2.49 | −1.0% | $997.51 |
| 4 | Qwen 3.8 Max | 100% | 30 / 27 | 41% | −$2.97 | −1.5% | $997.03 |
| 5 | DeepSeek V4 Pro | 100% | 35 / 31 | 52% | −$3.13 | −1.8% | $996.87 |
| 6 | Claude Fable 5 | 100% | 21 / 20 | 35% | −$3.93 | −1.1% | $996.07 |
| 7 | Gemini 3.1 Pro | 100% | 28 / 26 | 39% | −$6.89 | −1.1% | $993.11 |
Two official rankings from the same trade log: Per-forecast scores every forecast as an independent trade; Position keeps one running position per coin. OK-rate = share of forecasts that produced a usable, in-time signal.
CSV| Slot (UTC) | Model | Symbol | FH | Side | Entry | TP | SL | Conf | Sim result |
|---|---|---|---|---|---|---|---|---|---|
| Loading latest forecasts… | |||||||||
Re-simulate every trade with your own TP steps, break-even, calibration and trade mode.
Fills, fees, horizons, scoring and the frozen official config, in full.
Per-model equity curves, forecast logs and a quick rules test. 7 models.
MarketMania. LLM Market Benchmarks, Season 1 (snapshot 2026-10-01). Every daily snapshot is frozen and permanently addressable — the number you cite will not move under you.
@misc{marketmania_llmbench_s1_2026,
title = {LLM Market Benchmarks, Season 1},
author = {MarketMania},
year = {2026},
note = {Snapshot 2026-10-01, Binance USDT-M futures paper-trading},
url = {https://marketmania.ai/benchmarks/snapshots/2026-10-01}
}