Are the probabilities honest?
This is the real test. Group every prediction by the chance we gave it, then check how often that group actually happened. A well-calibrated model sits close to the diagonal: 30% tips land about 30% of the time, 70% tips about 70%.
| We said | Predictions | Expected | Actually happened | Calibration |
|---|---|---|---|---|
| Under 20% | 102 | 16.0% | 13.7% | -2.2 pts |
| 20–30% | 361 | 25.6% | 27.7% | +2.1 pts |
| 30–40% | 253 | 34.7% | 28.5% | -6.3 pts |
| 40–50% | 376 | 45.3% | 44.9% | -0.3 pts |
| 50–60% | 335 | 54.4% | 55.5% | +1.1 pts |
| 60–70% | 154 | 64.5% | 70.8% | +6.3 pts |
| 70–80% | 121 | 75.1% | 73.6% | -1.6 pts |
| Over 80% | 44 | 83.6% | 84.1% | +0.5 pts |
The marker on each bar is what we predicted; the fill is what happened. Within about four points is good; a persistent gap in one direction means the model is biased there.
By market
Draws deserve special attention. A draw tip at 30% that wins three times in ten is a correct prediction, even though it lost seven times. Judge it on the gap between expected and actual, and on whether backing it at our quoted price made or lost money.
| Market | Settled | Expected | Happened | Brier | Profit at best price |
|---|---|---|---|---|---|
| Home wins | 194 | 42.9% | 38.1% -4.8 pts | 0.209 | −37.12 ROI -19.1% |
| Draws | 194 | 26.1% | 29.9% +3.8 pts | 0.208 | +22.73 ROI 11.7% |
| Away wins | 194 | 31.0% | 32.0% +0.9 pts | 0.192 | −5.31 ROI -2.7% |
| Over 2.5 goals | 194 | 52.7% | 61.3% +8.6 pts | 0.235 | +26.87 ROI 13.9% |
| Under 2.5 goals | 194 | 47.3% | 38.7% -8.6 pts | 0.235 | −39.12 ROI -20.2% |
| Both teams to score | 194 | 54.8% | 54.1% -0.7 pts | 0.241 | −11.73 ROI -6.0% |
How to read this page
Hit rate is not accuracy. If we tip a draw and give it a 30% chance, we are saying it will lose about seven times in ten. Losing seven times in ten is then the correct outcome, not a failure. The only way to judge a probability is to collect many of them and check whether reality matched the number.
Calibration is that check. Take every prediction we rated between 30% and 40%; if roughly 35% of them happened, the numbers are honest. A gap of more than about eight points, repeated across hundreds of predictions, means something is wrong — not bad luck.
The Brier score squares the error on every prediction and averages it. Zero is perfect, 0.25 is a coin flip, and lower is better. It punishes confident mistakes far more than cautious ones, which is exactly what you want from a forecaster.
Profit at best price answers the practical question: if you had backed every one of these at the price we showed, staking one unit each time, where would you be? This is the number that matters most, and it is usually negative for favourites and closer to break-even on draws, because the market systematically overprices short odds.
Predictions are logged before kick-off and never edited afterwards. Past performance is not a guide to future outcomes. 18+.