Every headline number carries a confidence interval, and we track whether our edge is decaying over time instead of quietly hiding it. If the model breaks, this page says so first.
| Tier | Accuracy | 95% CI |
|---|---|---|
| S | 83.7% | 75.6% – 90.7% |
| A+ | 75.1% | 69.2% – 81.1% |
| A | 68.0% | 59.0% – 77.0% |
| B | 66.3% | 63.0% – 69.5% |
| C | 58.3% | 53.4% – 63.1% |
The raw ensemble optimizes which side to pick, not the exact win probability. The pick is a hard 3-model vote while this number is the mean of three probabilities, so the mid buckets win a few points more often than stated (underconfident, not wrong). The confidence we display on picks corrects this with a per-season Platt layer (fit on prior seasons only, clamped so it never flips the pick); this chart shows the raw model underneath. The gap column shows it honestly; whiskers are 95% Wilson intervals.
| Predicted Bucket | Predicted | Actual | Gap | N |
|---|---|---|---|---|
| 40% – 50% | 48.9% | 54.7% | +5.8% | 64 |
| 50% – 60% | 55.0% | 63.7% | +8.6% | 842 |
| 60% – 70% | 64.2% | 68.9% | +4.7% | 528 |
| 70% – 80% | 73.6% | 75.8% | +2.2% | 161 |
| 80% – 90% | 81.9% | 86.7% | +4.8% | 15 |
| Season | Accuracy |
|---|---|
| 2020 | 67.5% |
| 2021 | 62.7% |
| 2022 | 66.2% |
| 2023 | 64.7% |
| 2024 | 69.5% |
| 2025 | 68.3% |