How every graded card did, by Diamond Radar score
Cards Radar scored 90–100 were upgraded 78% of the time — 3.2× more often than a card picked at random.
The bar to beat: 1 in 4 cards (25% of 6,313) got upgraded this season no matter what Radar scored them. Any band above 25% beat chance.
Scored by Radar models v1, v5 and v11.
| Score band | Cards | Upgraded | Hit rate | vs chance |
|---|---|---|---|---|
| 90–100 | 65 | 51 | 78% | 3.2× |
| 80–89 | 96 | 76 | 79% | 3.2× |
| 70–79 | 166 | 116 | 70% | 2.8× |
| 60–69 | 269 | 160 | 59% | 2.4× |
| 50–59 | 485 | 243 | 50% | 2.0× |
| 40–49 | 765 | 308 | 40% | 1.6× |
| 30–39 | 974 | 271 | 28% | 1.1× |
| 20–29 | 1,112 | 219 | 20% | 0.8× |
| 10–19 | 1,128 | 88 | 8% | 0.3× |
| 0–9 | 1,253 | 22 | 2% | 0.1× |
If you only take the top of the board
The table grades every card Radar scored. These count only the highest-ranked rows of each update’s board, which is how the board is actually read.
- Top 5: 76% upgraded, against 25% by chance — 3.0× better. 19 of 25 across 5 updates, ranging 60%–100%.
- Top 10: 78% upgraded, against 25% by chance — 3.1× better. 39 of 50 across 5 updates, ranging 60%–100%.
- Top 20: 79% upgraded, against 25% by chance — 3.1× better. 79 of 100 across 5 updates, ranging 70%–100%.
Model v11: of the 162 cards it scored 70 or above in its 1 graded update, 72% were upgraded, against 21% by chance: 3.5× better.
The band table pools every graded update this season, scored on model versions v1, v5 and v11. The Model v11 line uses only v11’s graded updates.
What actually happened, by signal strength
Every scored card sorted into 10 signal bands, and the share of each band that San Diego Studio actually upgraded. The dashed line is chance: the share of all scored cards that upgraded, whatever Radar scored them. Bars above it beat chance. Scored by Radar models v1, v5 and v11.
Share of the band that upgradedChance 25%
The record
Mixed methods5
Windows scored
May 8 – Sep 11
Diamond Radar scores players who may be upgraded at the next roster update — it doesn't guarantee upgrades. Each update is graded once San Diego Studio publishes it: every scored card counts, in the score band it held before the update.
By roster update window
20 windows with no major SDS changes hidden.
What a “call” is
A call is a prediction Diamond Radar makes before San Diego Studio publishes a roster update — not an explanation written afterwards. Radar flags a card as an upgrade candidate while the update is still unannounced, and that flag is timestamped and frozen. When the update lands, the card either went up or it didn’t.
That ordering is the whole point, and it is what makes this page checkable. Anyone can explain a rating change after the fact. The scores graded above were on record first, and every scored update below shows them next to what San Diego Studio actually did.
How Radar decides a card is a candidate
The core idea is a gap. Every rated attribute on a card maps to something the real player actually does on a baseball field — contact and power map to how he hits, a pitcher’s ratings map to what he gets out of hitters. Radar reads the player’s real MLB production and compares it against what his card currently claims. A player producing far above his card’s ratings has an upward gap, and a persistent gap is what a roster update tends to correct.
A single hot week is noise, so the gap is measured over several time horizons at once — the last few days, the current update window, the past month, the season — and blended, weighted toward recent form. Horizons with no data are dropped and the rest re-weighted between them, so a player who missed a month is judged on what exists rather than penalised for the gap in the record.
Each horizon carries a confidence from sample size, and the bar adapts to the horizon rather than being fixed: a typical hitter accumulates a few dozen plate appearances in a week and several hundred across a season, and starters and relievers are held to their own workloads. A part-time player’s hot stretch counts for less than an everyday player’s, because it should.
The threshold is also tier-aware. High-overall cards have less headroom — a 95 has fewer attributes that can plausibly move — so their natural gaps are smaller, and a signal that means nothing on a Bronze card can be significant on a Diamond. A single fixed cutoff would flag low-rated players constantly and elite ones almost never. Radar predicts a direction and a size, and it can flag cards trending the other way too.
How to read this page
The table grades every card Radar scored, not only the ones it would have flagged. When an update lands, each scored card is sorted into its Diamond Radar score band, and a band’s hit rate is the share of its cards San Diego Studio upgraded. Every band is read against chance: the share of all scored cards that were upgraded, whatever Radar scored them. A band well above chance is the score doing its job; a low band near the bottom is the score being right about the cards it rates low.
The table pools every graded update this season, and those updates were scored on different model versions: model versions v1, v5 and v11. It is the season’s record, not any one model’s.
The one-line figure under the table counts only the latest model version’s graded updates (model v11), and names it, so it always describes a single model. It is taken at a Radar score of 70 or above, the cards at the top of the board.
The chart under the table shows the same bands as bars, with chance as a dashed line. It is a count of outcomes, not a claim about them.
Nothing here is held back for a paid tier. A number can still be missing — either there is not yet enough behind it to report honestly, or it was not recorded for that stretch — and the page says which, in its place, rather than filling the gap with a figure we would not stand behind.
Roster updates are a human decision at San Diego Studio, and real-life production is one input into that decision, not the whole of it. A player can rake for a month and get nothing; another can be adjusted for reasons no model can see from a box score.
There is also a ceiling, and it is worth stating plainly: the strongest signal Radar produces does not reach certainty. No honest reading of this page should suggest a card is guaranteed, and anyone promising that is guessing.
Why most roster updates aren’t on this page
A roster update is San Diego Studio re-rating players in MLB The Show based on real-world performance. They arrive regularly, but most carry no meaningful attribute changes — nothing to predict and nothing to score, so no scorecard exists for them and none is invented here. Of the 25 update windows on record, 5 produced changes worth scoring and 20 carried none.
That is also why this record grows in months rather than weeks, and why a scored update is worth reading in full: each one is a batch of predictions resolved at a single moment, against a set of ratings that had not moved in weeks.
Every scored roster update
One page per update — the cards Diamond Radar called before San Diego Studio published it, and what they actually got.