Receipts · Public Record
This is our report card. Before every game, our model calls who wins and how sure it is — and we lock that call in, stamped with the time, before the game starts. Once it's over, we grade it against what actually happened. Every hit and every miss stays here, in the open. No edits, no hindsight.
| Team | Pre | Aug | Trend |
|---|---|---|---|
| 82% | 82% | ||
| 65% | 65% | ||
| 63% | 63% | ||
| 57% | 57% |
| Team | Jul | Aug | Sep | Trend |
|---|---|---|---|---|
| 100% | 100% | 100% | ||
| 100% | 100% | 100% | ||
| 100% | 100% | 100% | ||
| 100% | 100% | 100% |
| Team | Pre | Trend |
|---|---|---|
| 91% | ||
| 85% | ||
| 82% | ||
| 80% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 44% | 49% | 49% | ||
| 29% | 33% | 33% | ||
| 6% | 8% | 8% | ||
| 12% | 5% | 6% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 54% | 61% | 62% | ||
| 24% | 28% | 28% | ||
| 5% | 4% | 4% | ||
| 9% | 3% | 3% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 46% | 50% | 49% | ||
| 18% | 15% | 15% | ||
| 10% | 9% | 12% | ||
| 9% | 11% | 10% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 82% | 85% | 85% | ||
| 8% | 8% | 8% | ||
| 2% | 2% | 2% | ||
| 3% | 2% | 2% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 46% | 40% | 40% | ||
| 19% | 20% | 20% | ||
| 8% | 8% | 8% | ||
| 6% | 7% | 7% |
| Team | Pre | Aug | Sep | Trend |
|---|---|---|---|---|
| 72% | 78% | 77% | ||
| 66% | 71% | 70% | ||
| 51% | 59% | 60% | ||
| 28% | 41% | 47% |
| Team | Pre | Trend |
|---|---|---|
| 22% | ||
| 17% | ||
| 11% | ||
| 10% |
| Team | Jul | Aug | Sep | Trend |
|---|---|---|---|---|
| 45% | 39% | 39% | ||
| 12% | 13% | 13% | ||
| 11% | 12% | 12% | ||
| 7% | 11% | 11% |
| Team | Jul | Aug | Trend |
|---|---|---|---|
| 97% | 98% | ||
| 2% | 1% | ||
| 1% | 1% | ||
| 0% | 0% |
Three outcomes, not two. Picking one of home, draw or away is a harder question than picking a winner, so these hit rates are not comparable to our football or baseball numbers. The Brier is summed over all three outcomes and runs 0–2, where the two-way sports' runs 0–1.
A draw is a miss, not a push. If we called a home win and it finished level, the hit rate counts that against us in full. The Brier is the fairer read of the same match: it scores every number we published, so rating the draw 30% and being wrong costs far less than rating it 8% and being wrong.
We rarely call a draw. Under this model the draw peaks around 29% while the two win chances split the rest, so it is almost never the single likeliest result. Every drawn match is therefore an automatic miss, which is what caps the hit rates below. The number that actually tests the draw model is on each league's card: what we rated those matches to draw, against how many did.
Results so far 10 home · 4 draw · 6 away|Draws we rated them 23%, 20% finished level
home / draw / away %, then our call
Results so far 16 home · 7 draw · 7 away|Draws we rated them 25%, 23% finished level
home / draw / away %, then our call
Results so far 8 home · 1 draw · 11 away|Draws we rated them 26%, 5% finished level
home / draw / away %, then our call
Results so far 7 home · 2 draw · 0 away|Draws we rated them 22%, 22% finished level
home / draw / away %, then our call
Results so far 6 home · 7 draw · 5 away|Draws we rated them 24%, 39% finished level
home / draw / away %, then our call
Results so far 15 home · 12 draw · 9 away|Draws we rated them 27%, 33% finished level
home / draw / away %, then our call · showing 12 of 22 locked
Before every game, the model locks in a win probability. Once the game is final, we score it — no edits, no hindsight. Here's how the calls have held up, sport by sport.
Games called
676
locked before first pitch
Now final
652
played & scored
Right side
58%
our pick won 376 of 652
Brier score
0.245
must beat 0.249 — a model that knows nothing
predicted vs actual
When the model gives a team, say, a 60% chance — do teams like that actually win about 60% of the time? Each row is how often our picks in that range really won.
actual predictedWell-calibrated → the marker sits at the edge of the bar.
Reading a row: the percentage is our chance that the named team — the home team — wins. The bolded team is our pick. Every line is locked before first pitch and graded automatically once the game is final. SP is what tonight's starters are worth, and it always names the team it helps: “SP PHI +5” means the pitching matchup moves the game five points to Philadelphia. Each starter is measured against his own club's rotation rather than the league, so an ace on an excellent staff can show a small number, or none at all. A row with no SP tag was locked before the starters were posted, so it's team-level only.
MLB home-win probabilities come from a log5 matchup on each club's regressed Pythagorean rating plus a fixed home-field edge — the same engine that runs the season simulation — then adjusted for the announced starting pitchers. Each starter is rated on FIP (the earned-run rate implied by his strikeouts, walks and home runs, which strips out defense and batted-ball luck), using only his innings as a starter, regressed 70 innings toward the league's rotation baseline, and measured against his own rotation's norm, since the club's run prevention already assumes an average start from it. That gap is credited 65% to the starter and capped at 10 points. A quarter of the rotation's own FIP-versus-runs gap is applied to the club's expected runs allowed on the same logic. Every weight was fit out of sample, on games played after the stats used to predict them — and the bullpen was tested the same way and left at zero because it didn't pay.
Games called
49
locked before kickoff
Now final
6
played & scored
Right side
83%
our pick won 5 of 6
Brier score
0.159
must beat 0.222 — a model that knows nothing
predicted vs actual
When the model gives a team, say, a 60% chance — do teams like that actually win about 60% of the time? Each row is how often our picks in that range really won.
actual predictedWell-calibrated → the marker sits at the edge of the bar.
Reading a row: the percentage is our chance that the named team — the home team — wins. The bolded team is our pick. Every line is locked before kickoff and graded automatically once the game is final.
CFB home-win probabilities come from each team's margin — its power rating minus the opponent's, plus a home-field edge — mapped through the model's margin spread. The same ratings that run the season sim.
Every prediction is written before the game and cannot be edited. Brier score is the mean squared error against the 0/1 outcome, lower is better; we judge it against the home-field base rate rather than a coin flip, because the base rate is free. Nothing on this page is backdated.