Track record

Every forecast, graded and published

We record how each prediction resolved and publish the result, whether it flatters us or not. Sample sizes are shown on every figure. Nothing here is selected after the fact.

Best Picks

The single highest-conviction selection per fixture, across all markets.

71.7%
Hit rate · n = 1,462
Average model probability
73.6%

Value Analysis

Selections where the model's probability diverges most from consensus market pricing.

57.7%
Hit rate · n = 1,805
Average model probability
65.1%
Average probability gap
8.02 pts

Compared to guessing

An accuracy figure means nothing on its own. The only useful question is what a naive rule would have scored on exactly the same matches.

49.9%
Model, match result
43.1%
Always pick the home team
+6.8
Difference

Measured across every graded fixture, not a selected subset. Home advantage is why the naive rule scores as well as it does; the gap above it is what the model contributes. n = 8,372.

Hit rate over time

Each line is a seven-day rolling average of settled selections, with a node marking the days behind it. Days with nothing settled break the line rather than being bridged across.

Best PicksValue Analysis
Hit rate over time
Hit rate over time
DateBest PicksValue Analysis
16.08.202669.4%53.8%
17.08.202670.2%54.3%
18.08.202670.9%55.0%
19.08.202671.9%54.0%
20.08.202670.8%55.2%
21.08.202669.5%55.5%
22.08.202671.2%59.4%
23.08.202672.3%58.9%
24.08.202670.4%59.8%
25.08.202671.6%58.0%
26.08.202672.2%58.2%
27.08.202672.3%57.8%
28.08.202674.4%57.4%
29.08.202675.9%53.0%
30.08.202673.8%54.0%
31.08.202674.7%51.8%
01.09.202673.1%52.6%
02.09.202673.1%53.0%
03.09.202673.4%53.5%
04.09.202672.4%53.8%
05.09.202668.5%56.0%
06.09.202671.0%54.5%
07.09.202670.3%56.6%
08.09.202671.7%57.4%
09.09.202670.9%58.6%
10.09.202671.4%60.0%
11.09.202671.8%61.4%
12.09.202672.4%59.3%
13.09.202670.1%60.5%
14.09.202670.5%60.8%

Are the probabilities honest?

This is the test that matters. We group every match-result forecast by the probability we published, then plot what actually happened. A calibrated model sits on the diagonal.

Perfect calibrationOur model95% interval
How often it actually happened
Are the probabilities honest?
Probability we published
Are the probabilities honest?
Probability bandProbability we publishedHow often it actually happened95% intervalFixtures
0.05-0.2014.4%14.8%13.6% – 16.0%3157
0.20-0.2522.9%22.8%21.6% – 24.2%3962
0.25-0.3027.2%27.5%26.4% – 28.6%5922
0.30-0.3532.3%31.1%29.4% – 32.9%2799
0.35-0.4037.5%36.5%34.6% – 38.4%2448
0.40-0.4542.3%42.7%40.6% – 44.9%2052
0.45-0.5047.4%45.7%43.2% – 48.2%1520
0.50-0.5552.3%54.0%51.1% – 56.9%1119
0.55-0.6057.3%59.8%56.2% – 63.3%729
0.60-0.6562.3%66.6%62.3% – 70.6%497
0.65-0.7067.3%70.7%65.4% – 75.6%304
0.70-0.7572.4%70.2%64.3% – 75.5%252
0.75-0.8077.3%73.6%66.1% – 79.9%155
0.80-0.9585.5%80.0%73.9% – 85.0%200

What this chart coversMatch result across every graded fixture, not just selected picks - a selected subset could not test calibration honestly. Buckets holding too few fixtures to plot are merged into the end points, and point size scales with sample. Vertical bars are 95% intervals: a thin bucket is an uncertain reading, not a precise miss.

Brier score: 0.2015Mean squared error between published probability and outcome. Lower is better; 0.25 is what you would score by predicting 50% on everything.

Why picks land below their stated probability

Both products land below their own stated probabilities. Best Picks carry an average published probability of % and hit %. Value Analysis averages % and hits %. Neither gap is hidden here, because the explanation matters more than the number.

It is a selection effect, not a broken model. Both products rank the model's own estimates and surface the top of the list, and conditioning on the highest outputs preferentially selects the cases where the model happened to read too high. Any system that ranks its own output and publishes the best of it shows this pattern - it is the same reason a fund's strongest backtested trades underperform once they are traded live.

The calibration curve above is the check, because it covers every graded fixture rather than a selected subset. Across the range where most matches sit, our published probabilities track observed frequencies closely, and if anything run slightly low rather than high. That is the opposite of an overconfident model, and it is what tells you the gap above is selection rather than miscalibration.

The practical reading: treat a stated probability as a probability rather than a promise, and expect a ranked shortlist to run a few points under its headline figure.

Which markets the picks came from

A hit rate across a mixed basket is not the same claim as a hit rate on match result alone. These are the six most-used markets in each basket, with the remainder grouped, so you can judge both headline figures for yourself.

Best Picks

  • Home Win or Draw53.9% of picksn = 788
  • Home Win14.2% of picksn = 208
  • Over 1.5 Goals12.2% of picksn = 178
  • Both Teams to Score5.3% of picksn = 78
  • Draw or Away Win5.1% of picksn = 75
  • Away Win4.3% of picksn = 63
  • Other markets (3)5.0% of picksn = 72

Value Analysis

  • Over 2.5 Goals12.4% of picksn = 224
  • Under 2.5 Goals11.0% of picksn = 198
  • Under 3.5 Goals10.6% of picksn = 191
  • Home Win9.6% of picksn = 174
  • Both Teams to Score7.4% of picksn = 133
  • Corners Under 10.56.4% of picksn = 116
  • Other markets (28)42.5% of picksn = 769

How these numbers are produced

Graded automatically, once

Every prediction is settled against the final result after the match finishes. Figures are never revised afterwards to look better, and nothing is removed from the record.

Abandoned fixtures are excluded

Matches that were cancelled, abandoned, walked over or administratively awarded never receive a win or loss. fixtures fall into this category in the period shown, and they are excluded from every figure rather than counted as either outcome.

Small samples are withheld

A percentage drawn from a handful of results is noise presented as fact. Figures below graded predictions are withheld rather than published, and calibration buckets holding fewer than fixtures are merged into the end points or left off the chart.

Past performance is not future performance

Everything here describes what has already happened. It is not a projection, not a target, and not a rate you should expect to personally experience. Football remains genuinely uncertain and no model changes that.

See today's forecasts

Three best picks a day are free, with every live score, fixture and table on the site.