MLB · 2026

Model Results

MLB baseball market snapshot

Last successful check: 2026-09-21T23:27:24.883957Z. Enable JavaScript to view current quote availability.

Every market has an independent outcome check. These results measure probability accuracy, not betting profits.

What was tested

3,178 completed games supply the history. Model weights use games through 2026-07-22; model selection and probability calibration use the next period through 2026-08-21. The untouched regular-season test runs 2026-08-22 to 2026-09-20.

Rolling player and team features update before each test date using earlier dates only. Live features now include completed games through 2026-09-20. Same-day results never enter a forecast.

Regular-season holdout

MarketGamesModel Brier ↓Reference Brier ↓Relative gainCalibration gap ↓Status
Pitcher strikeouts4050.15450.1709+9.6%1.6 ppPassed initial checks
Pitcher outs4050.18490.2006+7.8%5.3 ppPassed initial checks
Batter hits4050.15000.1521+1.4%0.9 ppPassed initial checks
Total bases4050.18610.1886+1.3%1.4 ppPassed initial checks
Home runs4050.10190.1031+1.2%1.4 ppPassed initial checks
RBIs4050.15110.1522+0.7%1.5 ppPassed initial checks
Game totals4050.23250.2313-0.5%5.5 ppResearch only
Moneylines4050.24830.2500+0.7%5.7 ppPassed initial checks
Run lines4050.22970.2301+0.2%4.9 ppPassed initial checks

2025 postseason audit

MarketGamesModel Brier ↓Reference Brier ↓Relative gainCalibration gap ↓Status
Pitcher strikeouts460.17710.2101+15.7%5.2 ppPassed initial checks
Pitcher outs460.22100.2308+4.2%11.8 ppPassed initial checks
Batter hits470.14830.1504+1.4%1.4 ppPassed initial checks
Total bases470.18340.1862+1.5%1.8 ppPassed initial checks
Home runs470.10520.1071+1.8%1.8 ppPassed initial checks
RBIs470.13740.1402+1.9%2.3 ppPassed initial checks
Game totals470.21600.2551+15.3%4.6 ppPassed initial checks
Moneylines470.24620.2500+1.5%4.7 ppPassed initial checks
Run lines470.23220.2346+1.0%3.2 ppPassed initial checks

How to read these numbers

Brier score is the average squared difference between a predicted probability and the result (0 or 1). Lower is better. The reference uses the training period’s overall outcome distribution, without player or matchup information. Relative gain compares the two scores; a negative gain means the model was worse. A small positive gain is not proof of a reliable advantage.

The calibration gap is the weighted difference between predicted and observed outcomes in ten probability buckets. Tests use fixed line grids, with several thresholds per forecast. Games are the number of distinct games; thresholds from the same game are correlated and are not independent trials. We have not established statistical significance.

Initial eligibility requires a lower Brier score than the reference, at least 150 regular-season games and a calibration gap no greater than 6 percentage points. Postseason checks require 30 games and at most 12 points. Those broader postseason limits reflect the small sample; they do not make playoff forecasts equally reliable.

The postseason audit used a separate earlier model: training through 2025-09-07, calibration through 2025-09-28. It never trained on the later regular-season holdout or on those playoff results. This is a limited historical stress test, not a validation of the current weights on future playoffs.

What this does not establish

We do not have timestamped historical book prices for this audit. There is no claimed historical return, win-rate-at-offered-odds, or proof that these forecasts beat the sportsbook market. Current picks are experimental estimates. A forward record begins with the archived forecast snapshots.

Model version: mlb-v1.0-ffe831160a077d5d-2026-09-20. Download the full validation report.