The model estimates one thing: the probability that a given starting hitter hits a home run in a given game. It outputs a calibrated probability, never a yes/no call. About 12% of starter-games end in a home run; a pick at 20% is a strong call, not a certainty.
Three groups, deliberately described as groups rather than as a feature list: talent (career and season contact quality — barrel rate, exit velocity, hard-hit rate), form (recent 30-game and 7-day deviations from that bat's own norm), and matchup (the starting pitcher's home-run and contact profile, the park's handedness-specific home-run factor, and weather). Predictions are anchored to each bat's own trailing baseline, so the model is always answering 'how is tonight different for this hitter' rather than 'who is good'.
Nine pre-registered tests, eight of them dead. Each had its decision rule written down before the result was known, so a null could not be argued away afterwards. Publishing the failures is the honest half of publishing the record.
| Idea | Verdict |
|---|---|
| Weather (air density, wind to centre) | The one that lived. Passed its gate and is in the candidate model now, awaiting a scheduled review. |
| Batter age | Dead. Added nothing measurable; the run-to-run noise of the training itself exceeded the effect. |
| Bullpen exposure | Dead. No replication on any cut, and the only significant result was the placebo — a sign of confounding. |
| Batter profile x pitcher profile | Dead. The construction was contaminated by batter quality; the signal did not survive removing it. |
| Spray angle x park shape | Dead. The variant that looked best predicted the opposing team's home runs — it was reading the park, not the hitter. |
| Pulled barrels | Dead. It appeared to work, then a control showed it was plain barrel rate wearing a direction label. |
| Expected plate appearances | Dead. Team plate appearances turned out to be barely forecastable, and batting order already carries the part that is. |
| Day-game-after-night fatigue | Dead before it was even tested properly — the raw data pointed the wrong way. |
| Starting-pitcher recent workload | Dead. One season looked convincing; the next was exactly zero. |
Boards are never edited after the fact. A pick that was published stays published, exactly as it was posted, whether it hit or missed — that is what makes the record on this site worth anything. If a scoring error is ever found that changed a number a reader saw here, the correction is stated on the affected night's own page, alongside the original, rather than quietly folded into a total.
The data directory carries the same picks and grades as CSV. Every figure on this site is computed from those rows by the same code that produces the nightly scorecard — there is no second pipeline that could quietly disagree.