SSILVERMINECollege basketball stats
LIVE BOARD
2026–27 SEASON
Experiment 01 / Learning through the season

Does another week
make a better forecast?

A frozen preseason model and a weekly updating challenger, compared on the same 5,734 games from 2025–26. Inspect the errors, the probabilities and the training behind each prediction.

Retrospective result / Full comparison

0.66 points less margin error.

The weekly model’s MAE was 9.61, compared with 10.26 for the preseason model. Winner accuracy was 70.0% versus 67.7%.

This is an experiment, not a live betting record. It uses currently published historical data and does not replace the 2026–27 preseason model. Rosters, injuries and market prices are absent.

Weekly minus preseason MAE: -0.66 points. Approximate 95% week-block bootstrap range: -0.80 to -0.50. Resamples 23 UTC weeks; repeated teams across weeks can still be dependent.

Across dated transitions

Does the update travel?

Independent holdouts · same field rules

Each row calibrates on the prior season, freezes that mapping, and scores the following season. The 2024, 2025 and 2026 tests stay separate so a strong year cannot hide a weak transition.

Test seasonCalibrated onGamesPreseason MAEWeekly MAEWeekly winner %Weekly fits
2023242022235,69510.229.4070.6%23
2024252023245,70110.109.4171.1%23
2025262024255,73410.269.6170.0%23

Margin MAE is points. Weekly fits are Monday snapshots; they are evidence of temporal replay, not a guarantee of future edge.

The 2023–24 row replays cached 2022 schedule and team-box releases as its prior-season training layer; it does not change production D1 data or current forecasts.

Transition index: all transitions ↗ · 202324 evidence ↗ · 202425 evidence ↗ · 202526 evidence ↗

Research track / Dated roster features

Does continuity explain the next season?

basketball-roster-challenger-v1 · generated Sep 12, 2026

This separate ridge challenger combines prior net efficiency with exact-ID roster continuity, represented prior minutes and listed-player counts. It is evaluated chronologically and stays outside the production forecast until more dated transitions are available.

10.49Held-out MAE · 202526
10.99Prior-net baseline MAE
0.50Points improved vs baseline
1,5792026–27 scenario games
Test seasonTraining seasonsTeamsMAEBaseline MAEImprovement
2024252023–242865.616.480.87 pts
2025262023–24, 2024–252895.876.550.68 pts

Historical transitions use the NCAA roster release and have no publisher Box BPM, so their scores are not directly comparable with the current ESPN-derived production challenger. 345 teams have current roster features; the 2026–27 scenario is a research prompt and does not change win probabilities, uncertainty or ledger registrations.

Roster releases are source snapshots without a verified pre-season publication clock. The challenger is research-only and does not replace the primary forecast or enter the prospective ledger. Only two historical roster transitions are available for fitting; the held-out evaluation is one season and is not a guarantee of future performance. Roster listings do not establish eligibility, availability, transfer reason, injury status or depth-chart role. Publisher Box BPM is unavailable for some source IDs; rows without prior BPM coverage are withheld from the challenger rather than imputed. The scenario changes the primary margin by the learned team-strength delta but does not recalibrate win probability or uncertainty.

Open roster-aware game planning ↗

What the models were allowed to know

Move forward. Never peek ahead.

01 / Before 2024–25

Establish the field

Fit 2023–24 efficiency and tempo. Freeze program membership before the next season; ten games in the latest fitting year are required.

02 / During 2024–25

Calibrate probabilities

Generate weekly predictions, then use 5,701 games to fit the challenger’s probability mapping and 80% margin range. The preseason model has its own calibration.

03 / During 2025–26

Replay the next season

Freeze calibration. Each Monday, refit using earlier completed games whose starts precede Sunday 00:00 UTC. Compare with a preseason fit that stays fixed all year.

The 24-hour start buffer reduces overlap with unfinished games; it is not proof of historical data availability. Source corrections are not rolled back. Earlier 2025–26 results may enter later weekly fits, but never their own forecast.

Explore the same-game comparison

Where does the difference appear?

All filters apply to both models

Loading the historical comparison…

Account for the missing games.

Completed schedule records6,300
Usable paired box scores6,298
Same-game comparison5,734
Outside the frozen program field564

Both methods exclude the same out-of-field games. The source is not a certified Division I membership list. These counts describe this source edition.

Use the prospective ledger to distinguish forecasts actually registered before games from historical experiments.

Reproduce the comparison.

Download game predictions, all weekly coefficients and training-game IDs, the calibration sample, and the file hash manifest.

Experiment edition: Sep 12, 2026. Dataset edition: Sep 12, 2026.

87187136f46b007234115ef7e60d4f4f7a2cf3682d817faec3fb72dbb7d39a4d

Interpretation & provenance

What this experiment cannot establish.

Retrospective replay using current source releases; historical revisions and availability timestamps are not reconstructed.

Weekly fits include only completed records with starts before Monday 00:00 UTC minus 24 hours; exact historical final-publication times are unavailable.

The 2023 calibration transition uses the cached 2022 schedule/team-box releases only for replay; it is retained outside the production D1 warehouse and is not a current forecast input.

2023–25 rolling predictions calibrate the challenger; those calibration results are not independent test performance.

2025–26 games enter later weekly fits only after the cutoff buffer. No game enters its own prediction or any earlier week's fit.

The preseason team's field is frozen before each season. New programs outside it are excluded from both methods.

No roster, availability, injury, recruiting or bookmaker inputs. This experiment does not replace live preseason forecasts or enter the prospective ledger.

The week-block bootstrap describes sampling variation within this one season. It is not a guarantee across future seasons or protection against shared-team dependence between weeks.

Fixed penalties, yearly weights and update cadence; no parameter search was performed. This is a new exploratory comparison on a season already used for the published baseline evaluation.

Temporal evaluation and calibration references: scikit-learn’s time-series evaluation guidance and probability calibration documentation. Source data: SportsDataverse, labeled CC BY 4.0 by its publisher. Source receipts and download URLs are in the summary. Read the production model notebook →