Sparrow Data

Sparrow Data

Graded prediction archive

Methodology

The model estimates one thing: the probability that a given starting hitter hits a home run in a given game. It outputs a calibrated probability, never a yes/no call. About 12% of starter-games end in a home run; a pick at 20% is a strong call, not a certainty.

What goes into it

Three groups, deliberately described as groups rather than as a feature list: talent (career and season contact quality — barrel rate, exit velocity, hard-hit rate), form (recent 30-game and 7-day deviations from that bat's own norm), and matchup (the starting pitcher's home-run and contact profile, the park's handedness-specific home-run factor, and weather). Predictions are anchored to each bat's own trailing baseline, so the model is always answering 'how is tonight different for this hitter' rather than 'who is good'.

The rules we hold ourselves to

What we tried and threw away

Nine pre-registered tests, eight of them dead. Each had its decision rule written down before the result was known, so a null could not be argued away afterwards. Publishing the failures is the honest half of publishing the record.

IdeaVerdict
Weather (air density, wind to centre)The one that lived. Passed its gate and is in the candidate model now, awaiting a scheduled review.
Batter ageDead. Added nothing measurable; the run-to-run noise of the training itself exceeded the effect.
Bullpen exposureDead. No replication on any cut, and the only significant result was the placebo — a sign of confounding.
Batter profile x pitcher profileDead. The construction was contaminated by batter quality; the signal did not survive removing it.
Spray angle x park shapeDead. The variant that looked best predicted the opposing team's home runs — it was reading the park, not the hitter.
Pulled barrelsDead. It appeared to work, then a control showed it was plain barrel rate wearing a direction label.
Expected plate appearancesDead. Team plate appearances turned out to be barely forecastable, and batting order already carries the part that is.
Day-game-after-night fatigueDead before it was even tested properly — the raw data pointed the wrong way.
Starting-pitcher recent workloadDead. One season looked convincing; the next was exactly zero.

Known limitations, stated plainly

If we get something wrong

Boards are never edited after the fact. A pick that was published stays published, exactly as it was posted, whether it hit or missed — that is what makes the record on this site worth anything. If a scoring error is ever found that changed a number a reader saw here, the correction is stated on the affected night's own page, alongside the original, rather than quietly folded into a total.

Auditing this yourself

The data directory carries the same picks and grades as CSV. Every figure on this site is computed from those rows by the same code that produces the nightly scorecard — there is no second pipeline that could quietly disagree.

Sparrow Data — every pick published before first pitch and graded the next morning. Archive covers 2026-07-09 through 2026-08-16. Nothing here is ever edited or removed after the fact. Methodology · Raw data