Model Accuracy Tracker

Every prediction this model makes is written to the database before the match is played, then scored against the result. No cherry-picking, and no re-pricing with hindsight.

How CSDB tracks prediction model accuracy

A scheduled job prices every upcoming match with the same model that runs on the predictions pageand stores the probability, the confidence and the model version at that moment. When the match finishes, only the outcome columns are filled in — the probability itself is never rewritten, and shipping a new model version cannot improve the old one's record. Each match is scored once, using the last snapshot taken before kickoff.

What the metrics mean

  • Overall accuracy— percentage of matches where the model's favourite won. A 60%+ accuracy on CS2 BO3s is meaningful, given how high-variance the game is.
  • Current streak — consecutive correct (positive) or incorrect (negative) predictions, useful for spotting hot or cold periods.
  • Calibration— when the model says "70% confidence," does the favourite actually win ~70% of the time? Calibration measures whether the model's confidence scores are honest.
  • Brier score — the squared error between the stated probability and what happened, averaged over every scored match. Lower is better; 0.25 is the score you get by calling everything 50/50, so anything meaningfully below that is the model earning its keep. It only appears here because the probabilities were stored in advance.

What we do and don't include

  • ✅ Every BO1, BO3, and BO5 played by HLTV-eligible teams
  • ✅ Both upset wins and chalk results
  • ❌ Matches where the model didn't have enough data on either team (typically tier-3 newcomers)
  • ❌ Forfeits and walkovers (no actual play)

Why transparency matters

Any model can look good on hand-picked matches. By publishing the full track record, CSDB lets you decide whether the model's edge is real. If you're using the hot picks page or the predictions tool for actual decisions, this page is where you check whether to trust them.

59.4%Accuracy
0.243Brier Score
8.4ppMean Calibration Error
340Matches Scored
v2Model Version

Every probability above was written to the database before its match was played and has not been touched since. Matches from Aug 28 – Sep 9, 2026. 1588 predictions awaiting a result. A Brier score of 0.25 is what you get by calling every match 50/50, so lower is better.

Calibration Check

Does a 70% call win about 70% of the time? Bars are the model's average stated probability against what actually happened, per band. Bands with few matches move around a lot — the count is under each one.

50–55%
56% (139)
55–60%
56% (70)
60–65%
69% (49)
65–70%
68% (41)
70–80%
54% (39)
80%+
100% (2)
Predicted
Actual

By event tier

D-Tier
60%
114/191 correct
C-Tier
56%
65/116 correct
A-Tier
68%
15/22 correct
B-Tier
73%
8/11 correct · too few to read into

By series format

BO3
60%
194/326 correct
BO1
54%
7/13 correct · too few to read into
BO5
100%
1/1 correct · too few to read into

By stated confidence

Low (1–5)
56%
91/163 correct
Medium (6–7)
63%
102/162 correct
High (8–10)
60%
9/15 correct · too few to read into