PROPHET 9
We measure whether a hitter’s underlying skill has diverged from his box score — and we publish every call.Four-season public record →
How we see it early

The box score is the last place a breakout shows up.

What it tells you

We surface players whose underlying skill has pulled ahead of — or fallen behind — their box-score results, and we say where that points in a stat you already use. For a starter: “his skill points to an ERA around 3.5, while his results sit at 4.0 — the skill is ahead.” For a hitter: “his skill points to an OPS around .840, ahead of his current .760.” (Relievers get a directional read, not a number — their signal is real but too noisy for a precise figure, and we won’t fake one.)

When it tends to happen

Over the coming weeks. Across four seasons of held-out history, players showing this pattern tended to see their results move toward their skill over roughly the next month. We never put a date on one player — we give you the direction and the window the pattern has actually played out over.

How we do it

Advanced machine-learning and statistical modeling, validated on history the models never saw. The rigor is the point; the recipe stays in the vault — we publish what we found, never how we find it.

How sure we are

From the track record, never a promise. Across 83 held-out weeks over four seasons, the players our signal flagged went on to beat comparable players at the same production level — not the unflagged field — by +.065 wOBA on the highest-conviction signals, stepping down to +.028 across all flagged players. And every published call is graded in the open as games finish — you watch the record fill in real time. And it is not recency dressed up: measured head-to-head, betting on a hitter's own recent form is the worst of four predictors — worse than assuming no change at all. What we will never give you is a per-player percentage; the honest numbers are the aggregate edge and the public ledger.

Every swing carries information the box score hasn’t absorbed yet. Prophet 9 measures a player’s underlying performance signals over the windows where they become reliable, compares them with his surface line, and flags the gap. When the underlying leads, improvement follows. We proved that relationship on four seasons of held-out history before we published a single call — and our track record stays public so you can keep auditing it.

The picture to hold in your head is below — drawn from today’s live lead signal, not a sketch. A player’s underlying signals — the quality of what he’s actually doing on the field — move first. His surface line follows, late, because surface stats need sample to catch up. The shaded area between the two readings is what we measure, every night, for every player.

SEASON →SIGNAL FLAGS HEREINPUTS & RESULTS ALIGNEDSKILL-IMPLIED ERAACTUAL ERA
Today’s lead signal · Gavin Williams, P
The proof came before the product
+.065wOBA lift on the highest-conviction signals, over comparable players at the same production level
+.028
across all flagged players — the ladder's floor
83
scored out-of-sample weeks, 2022–2025
2022–2025
four seasons, every CI excluding zero

Before publishing a single call we validated the signal out of sample, and the finding is a ladder rather than a number: the highest-conviction signals separate from comparable players at the same production level by +.065 wOBA, stepping down to +.035 for breakout-versus-regression and +.028 across all flagged players — the effect scales with the size of the gap. Across four seasons the top tier beat a matched control by +19.5 to +32.5 percentage points, every interval excluding zero. 2026 is under watch at +8.9pp with an interval crossing zero, and is never run as a headline. The methodology is ours; the result is public.

The projection engine — a separate product

The figures above are the breakout & regression signal’s record. The projections below are a different product on a different engine, and their held-out accuracy has not been published — the slot exists, the figure does not. The signal’s validation record.

Pattern Spotter

Reads the most recent form aggressively. Quickest to believe a genuine surge — and the first to be fooled by a hot streak, which is why it never publishes alone.

Consensus — what we publish

The blended view. Where the two readers agree, it speaks plainly; where they split, the spread is reported to subscribers as information in its own right.

Cautious Forecaster

Shrinks toward the player’s established level. Slowest to chase noise — and the last to recognize a real transformation, which is why it never publishes alone either.

What we won’t tell you

Which signals we read, over which windows, with which thresholds. That’s the edge, and publishing it would erase it for everyone who pays for it. What we publish instead is everything you need to judge us: the validation, the intervals, and a permanent public record of every call.

Trust the audit, not the adjectives.

What we do tell you is whether his skill is leading or lagging his results — and where his stats are heading because of it. You won’t see which inputs we read, over which windows, or where the lines are drawn — that’s the part that stays in the vault.

The three commitments
Underlying vs. surface

Two readings of every player: what the inputs say, and what the box score says. The gap between them is the signal.

Validated before published

Four seasons out of sample, measured against comparable players at the same production level: +.028 wOBA, rising to +.065 at the highest conviction. The proof came before the product.

Calibrated, auditable

Every projection ships with an honest interval, and every call lands on a public ledger — hits and misses alike.

Audit the record yourself →