Sign up and get 3 free requests with Start plan accessSign up →

Research finding

In the latest weekly window, 99.3% of eligible forecasts matched the other models' majority

Agreeing forecasts had a lower observed hit rate in this window; the disagreement group contained 30 forecasts.

If a model's forecast matches the consensus of the other models, is the forecast more reliable?

Updated 30 Sept 2026 · sources: Consensus Watch #8, Consensus Watch #7, Consensus Watch #6, Consensus Watch #5, Monthly Consensus Watch #1, Consensus Watch #4, Consensus Watch #3, Consensus Watch #2, Consensus Watch #1

All findingsThe benchmarkResearch reports

Answer

In Consensus Watch #8 (Sep 21-27, 2026 UTC), 99.3% of eligible forecasts matched the other models' leave-one-out majority, and agreeing calls scored 14.0pp lower than disagreeing calls (46.0% on n = 4,215 against 60.0% on n = 30). The 99.3% is the share of individual forecasts whose direction matched the leave-one-out majority of the other models; it is not the share of slots on which all 7 models agreed. The agreeing group carries a 95% Wilson interval of [44.5%, 47.5%] and the disagreeing group [42.3%, 75.4%]. A further 713 forecasts were thin or tied — fewer than three directional peers, or a peer split with no majority — and are excluded from both groups rather than forced into one.

Monthly Consensus Watch #1 pools the whole month: 98.7% peer-majority agreement on 16,089 scoreable calls, and agreeing calls scored 1.9pp higher than disagreeing calls (45.5% on n = 15,878 against 43.6% on n = 211). The weekly and the monthly readings do not have to point the same way, and in the published record they do not: the weekly lift reads worse for agreement, the monthly one better.

Across the 8 published weekly windows the sign of the lift changed 4 times, on disagreeing buckets that run from n = 6 to n = 125. The herd is large and steady; the difference between agreeing and disagreeing is not.

Evidence

Agreement against disagreement, leave-one-out

WindowPeer-majority agreementAgreeDisagreeLiftThin or tied
Consensus Watch #899.3%46.0% (n = 4,215)60.0% (n = 30)-14.0pp713
Monthly Consensus Watch #198.7%45.5% (n = 15,878)43.6% (n = 211)+1.9pp3,294

Consensus Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC

Lift is the agreeing hit rate minus the disagreeing hit rate, in percentage points.

Hit rate by size of the agreeing majority — Consensus Watch #8

Peers on the same sidenHit rate
2966.7%
330439.1%
448539.0%
587646.8%
62,54147.8%

Consensus Watch #8 · Sep 21-27, 2026 UTC · cutoff Mon 28 Sept 2026 16:00 UTC

A cell needs at least three directional peers to be scored at all, so the thinnest majorities carry the smallest samples.

Hit rate by size of the agreeing majority — Monthly Consensus Watch #1

Peers on the same sidenHit rate
213860.1%
31,96445.9%
42,63045.1%
54,48245.8%
66,66445.1%

Monthly Consensus Watch #1 · Aug 1-31, 2026 UTC · cutoff Tue 1 Sept 2026 16:00 UTC

A cell needs at least three directional peers to be scored at all, so the thinnest majorities carry the smallest samples.

Across issues

Pooled lift, issue by issue

IssueWindowPeer-majority agreementAgreeDisagreeLift
Consensus Watch #1Aug 3-9, 2026 UTC99.4%40.4% (n = 3,308)63.2% (n = 19)-22.8pp
Consensus Watch #2Aug 10-16, 2026 UTC99.8%43.1% (n = 2,923)50.0% (n = 6)-6.9pp
Consensus Watch #3Aug 17-23, 2026 UTC97.1%54.8% (n = 4,232)37.6% (n = 125)+17.1pp
Consensus Watch #4Aug 24-30, 2026 UTC99.1%41.2% (n = 3,848)47.2% (n = 36)-6.0pp
Consensus Watch #5Aug 31-Sep 6, 2026 UTC99.3%41.6% (n = 3,725)69.2% (n = 26)-27.6pp
Consensus Watch #6Sep 7-13, 2026 UTC99.3%38.8% (n = 3,762)69.2% (n = 26)-30.4pp
Consensus Watch #7Sep 14-20, 2026 UTC98.9%50.6% (n = 4,825)38.5% (n = 52)+12.1pp
Consensus Watch #8Sep 21-27, 2026 UTC99.3%46.0% (n = 4,215)60.0% (n = 30)-14.0pp

The sign of the lift changed 4 times across 8 weekly issues.

Caveats

The issue's own caveats, in its words:

Read with context. The disagree bucket is n=30 this week against 52 in issue #7 and 26 in issue #6; the two Wilson intervals overlap. This is a single week of dependent observations: not evidence about herding in general, and not a claim that the sign will hold next week.
Single week (Sep 21-27, 2026 UTC). Every finding is descriptive for this window only.
Observations are not independent: the same models watch overlapping symbol / FH / TF cells hour after hour. Wilson intervals here are descriptive, not inferential.
Correlational, not causal: all models see the same market data, so 'agreement' and 'hit' can rise together simply because a slot was easy to read. LOO removes self-match bias, not this confound.
Leave-one-out, as the issue defines it: leave-one-out strict majority, >=3 directional peers. A model is excluded from the consensus it is scored against, and a tied peer split is left out rather than forced either way.

Definitions

  • Method — leave-one-out strict majority, >=3 directional peers.
  • Agree — the forecast's direction matches the strict majority of the other model lines in the same cell. Disagree — its direction differs from the strict majority of the other models.
  • Margin — how many of the other lines stood on that majority side.

Agree, disagree and margin, as the issue defines them:

For every model M and every forecast it made, we rebuild that slot's consensus using only the OTHER active models in the same symbol / forecast-horizon (FH) / timeframe (TF) cell -- M's own call never counts toward its own consensus. M is scored agree if its side (long or short) matches the strict majority of those peers, and disagree if it sits alone against that majority; a cell needs at least 3 directional peers to count at all, and tied peer splits are excluded rather than forced either way. LOO peers are genuinely external to the model being scored, so the comparison below is not comparing a model to itself. The grok rows are one line, so grok never counts as two peers in a cell.

Reports and data

IssuePublishedFiles
Consensus Watch #830 Sept 2026PDFMDJSON
Consensus Watch #722 Sept 2026PDFMDJSON
Consensus Watch #615 Sept 2026PDFMDJSON
Consensus Watch #58 Sept 2026PDFMDJSON
Monthly Consensus Watch #14 Sept 2026PDFMDJSON
Consensus Watch #42 Sept 2026PDFMDJSON
Consensus Watch #326 Aug 2026PDFMDJSON
Consensus Watch #219 Aug 2026PDFMDJSON
Consensus Watch #110 Aug 2026PDFMDJSON

Dataset cardPublic APIDaily snapshots

Cite

MarketMania Research (2026). Research reports (weekly, monthly), PDF/MD/JSON. https://marketmania.ai/research — dataset card: https://marketmania.ai/research/dataset

To cite one reading instead, name the issue it came from and the date it was published: an issue is frozen, so a citation to one is stable.

Research benchmark — not financial or investment advice; paper trading only.