How Do Prospect Grades Translate to Future Outcomes?

Reggie Hildred-USA TODAY Sports

Hello, and welcome to Prospect Week! (Well, closer to Prospect Fortnight — as you can probably tell from the navigation widget above, the fun will continue well into next week, including the launch of our Top 100.) I’m not your regular host – that’d be Eric Longenhagen – but not to worry, you’ll get all the Eric you can handle as he and the team break down all things minor leagues, college baseball, and MLB draft. I’m just here to set the stage, and in support of that goal, I have some research to present on prospect grades and eventual major league equivalency.

When reading coverage of the minor leagues, I often find myself wondering what it all means. The Future Value scale does a great job of capturing the essence of a prospect in a single number, but it doesn’t translate neatly to what you see when you watch a big league game. Craig Edwards previously investigated how prospect grades have translated into surplus value, but I wanted to update things from an on-field value perspective. Rather than look at what it would cost to replace prospect production in free agency, I decided to measure the distribution in potential outcomes at each Future Value tier.

To do that, I first gathered my data. I took our prospect lists from four seasons, 2019-22, and looked at all of the prospects with a grade of 45 FV or higher. I separated them into two groups — hitters and pitchers — then took projections for every player in baseball three years down the line. For example, I paired the 2019 prospect list with 2022 projections and the 2022 prospect list with 2025 projections. In this way, I came up with a future expectation for each player.

I chose to use projections for one key reason: They let us get to an answer more quickly. In Craig’s previous study, he looked at results over the next nine years of major league play. I don’t have that kind of time – I’m trying to use recent prospect grades to get at the way our team analyzes the game today. If I used that methodology, the last year of prospect lists I could use would be 2015, in Kiley McDaniel’s first term as FanGraphs’ prospect analyst.

Another benefit of using projections is that they’re naturally resistant to the sample-size-related issues that always crop up in exercises like this. A few injuries, one weird season, a relatively small prospect cohort, and you could be looking at some strange results. Should we knock a prospect if his playing time got blocked, or if his team gamed his service time? I don’t think so, and projections let us ignore all that. I normalized all batters to a 600 plate appearance projection and all pitchers to a 200 innings pitched projection.

I decided to break future outcomes down into tiers. More specifically, I grouped WAR outcomes as follows. I counted everything below 0.5 WAR per season as a “washout,” including those players who didn’t have major league projections three years later. Given that we project pretty much everyone, that’s mostly players who had either officially retired or never appeared in full-season ball. I graded results between 0.5 and 1.5 WAR as “backup.” I classified seasons between 1.5 and 2.5 WAR as “regular,” as in a major league regular. Finally, 2.5-4 WAR merited an “above average” mark, while 4-plus WAR got a grade of “star.” You could set these breakpoints differently without too much argument from me; they’re just a convenient way of showing the distribution. There’s nothing particularly magical about the cutoff lines, but you have to pick something to display the data, and a simple average of WAR projections probably isn’t right.

You Aren't a FanGraphs Member
It looks like you aren't yet a FanGraphs Member (or aren't logged in). We aren't mad, just disappointed.
We get it. You want to read this article. But before we let you get back to it, we'd like to point out a few of the good reasons why you should become a Member.
1. Ad Free viewing! We won't bug you with this ad, or any other.
2. Unlimited articles! Non-Members only get to read 10 free articles a month. Members never get cut off.
3. Dark mode and Classic mode!
4. Custom player page dashboards! Choose the player cards you want, in the order you want them.
5. One-click data exports! Export our projections and leaderboards for your personal projects.
6. Remove the photos on the home page! (Honestly, this doesn't sound so great to us, but some people wanted it, and we like to give our Members what they want.)
7. Even more Steamer projections! We have handedness, percentile, and context neutral projections available for Members only.
8. Get FanGraphs Walk-Off, a customized year end review! Find out exactly how you used FanGraphs this year, and how that compares to other Members. Don't be a victim of FOMO.
9. A weekly mailbag column, exclusively for Members.
10. Help support FanGraphs and our entire staff! Our Members provide us with critical resources to improve the site and deliver new features!
We hope you'll consider a Membership today, for yourself or as a gift! And we realize this has been an awfully long sales pitch, so we've also removed all the other ads in this article. We didn't want to overdo it.

With that said, let’s get to the results. My sample included 685 hitters from 45-80 FV. Allowing for some noise at the top end due to small sample size, the distribution looks exactly like you’d hope:

Hitter Outcome Likelihood by FV
FV Washed Out Backup Regular Above Average Star Count
45 51% 25% 17% 6% 1% 295
45+ 52% 18% 19% 11% 1% 91
50 23% 24% 30% 21% 2% 197
55 17% 17% 30% 31% 6% 54
60 14% 12% 19% 38% 17% 42
65 0% 33% 33% 0% 33% 3
70 0% 0% 0% 0% 100% 2
80 0% 0% 0% 0% 100% 1
Note: Projections from three years after the player appeared on a prospect list

Consider the 55 FV line for an explanation. Of the players we graded as 55 FV prospects, 17% look washed three years later – Jeter Downs, a 2020 55 FV, for example. Another 17% have proven to be backup-caliber, like 2022 55 FV Curtis Mead, or 2019 55 FV Taylor Trammell if you don’t think Mead’s trajectory is set just yet. Continuing down the line, 30% look like big league regulars – 2021 55 FV Alek Thomas, perhaps. A full 31% appear to be above-average major league contributors three years later, like 2019 55 FV Sean Murphy or 2021 55 FV Royce Lewis. Finally, 6% project as stars three years later – Jackson Merrill, a 55 FV in 2022, feels appropriate as an example.

Two things immediately jump out to me when looking at this data. First, the “above average” and “star” columns increase at every tier break, and the “washout” column decreases at every tier break. In other words, the better a player’s grade, the more likely they are to be excellent, while the worse their grade, the more likely they are to bust. That’s a great sign for the reliability of our grades; they’re doing what they purport to do, essentially.

Second, each row feels logically consistent. The 45 FV prospects are most likely to bust, next-most-likely to end up as backups, and so on. The 45+ FVs look like the 45 FVs, only with a better top end; their chances of ending up above average are meaningfully better. The 50 FVs are a grab bag; their outcomes vary widely, and plenty of those outcomes involve being a viable major leaguer. By the time you hit the 55 and 60 FV prospects, you’re looking at players who end up as above-average contributors a lot of the time. The gap between 55 and 60 seems clear, too; the 60 FVs are far more likely to turn into stars, more or less. Finally, there are only six data points above 60 FV, so that’s mostly a stab in the dark.

This outcome pleases me greatly. Looking at that chart correlates strongly with how I already perceived the grades. For a refresher, roughly 30 prospects in a given year grade out as a 55 FV or above, give or take a few. Something like three quarters of those tend to be hitters. That means that in a given year, 20-ish prospects look like good bets to deliver average-regular-or-better performance. The rest of the Top 100? They’re riskier, with a greater chance of ending up in a part-time role and a meaningfully lower chance of becoming a star. But don’t mistake likelihood for certainty – plenty of 55 and 60 FVs still end up at or below replacement level, and 45 FVs turn into stars sometimes. Projecting prospect performance is hard!

How should you use this table? I like to think of Future Value in terms of outcome distributions, and I think that this does a good job of it. Should a team prefer to receive two 50 FV prospects in a trade, or a 55 FV and a 45 FV? You can add up the outcome distributions and get an idea of what each combination of prospects looks like. Here are the summed probabilities of those two groups:

Two Similar Sets of Prospects, Grouped
Group Washed Out Backup Regular Above Average Star
Two 50 FVs 46% 49% 60% 42% 4%
One 55, One 45 68% 42% 47% 37% 6%

Another way of saying that: If you go with the two-player package that has the 55 and 45 FV prospects, you’re looking at a higher chance of developing a star. You’re also looking at a greater chance of ending up with at least one complete miss, and therefore lower odds of ending up with two contributors. Adding isn’t exactly the right way to handle this, but it’s a good shorthand for quick comparisons. If you want to get more in depth, I built this little calculator, which lets you answer a simple question: For a given set of prospects, what are the odds of ending up with at least X major leaguers of Y quality or better? You can make a copy of this sheet, define X and Y for yourself, and get an answer. In our case, the odds of ending up with at least one above-average player (or better) are 40.7% for the two 50s and 41.4% for the 45/55 split. The odds of ending up with two players who are at least big league regulars? That’d be 28.1% for the two 50 FVs, and 16.1% for the 45/55 pairing. Odds of at least one star? That’s 4% for the two 50 FVs and 6% for the 45/55 group. In other words, the total value is similar, but the shape is meaningfully different.

For example, you’d have to add together a ton of 50 FV prospects to get as high of a chance of finding a star as you would from one 60 FV. On the other hand, if you have three 50 FVs, the odds of ending up with at least a solid contributor are quite high. Meanwhile, even 60 FV prospects end up as backups or worse around a quarter of the time. That description of the relative risks and rewards makes more sense to me than converting players into some nebulous surplus value. Prospects are all about possibility, so representing them that way tracks analytically for me.

Take another look at the beautiful cascade of probabilities in that table of outcomes for hitting prospects, because we’re about to get meaningfully less pretty. Let’s talk about pitching prospects. Here, the outcomes are less predictable:

Pitcher Outcome Likelihood by FV
FV Washed Out Backup Regular Above Average Star Count
45 53% 26% 16% 5% 0% 230
45+ 38% 24% 25% 13% 0% 68
50 27% 27% 24% 20% 2% 96
55 17% 20% 37% 27% 0% 30
60 17% 33% 25% 25% 0% 12
65 0% 0% 0% 100% 0% 1
70 0% 0% 100% 0% 0% 1
Note: Projections from three years after the player appeared on a prospect list

I have tons of takeaways here. First, there are substantially fewer pitching prospects ranked, particularly as 50 FVs and above. Clearly, that’s a good decision by the prospect team, because even the highest-ranked pitchers turn into backups at a reasonable clip. Pitching prospects just turn into major league pitchers in a less predictable way, or so it would appear from the data.

Second, there are fewer stars among the pitchers than the hitters. That’s true if you look at 2025 projections, too. There are only six pitchers projected for 4 WAR or higher, while 42 hitters meet that cutoff. It’s also true if you look at the results on the field in 2024; 36 hitters and 12 pitchers (22 by RA9-WAR) eclipsed the four-win mark. You should feel free to apply some modifiers to your view of pitcher value if you think that WAR treats them differently than hitters, but within the framework, the relative paucity of truly outstanding outcomes is noticeable.

Another thing worth mentioning here is that pitchers don’t develop the same way that hitters do. Sometimes one new pitch or an offseason of velocity training leads to a sudden change in talent level in a way that just doesn’t happen as frequently with hitters. Tarik Skubal was unmemorable in his major league debut (29 starts, with a 4.34 ERA and 5.09 FIP). Then he made just 36 (very good) starts over the next two years due to injuries. Then he was the best pitcher in baseball in 2024. Good luck projecting that trajectory. Perhaps three-year-out windows of pitcher performance just aren’t enough thanks to the way they continue to develop even after reaching the majors.

There’s one other limitation of measuring pitchers this way: I don’t have a good method for dealing with the differential between reliever and starter valuation. Normalizing relievers to 200 innings pitched doesn’t make a ton of sense, but handling them on their own also feels strange, and I don’t have a good way of converting reliever WAR to the backup/regular/star scale that I’m using here. A 3-WAR reliever wouldn’t be an above-average player, they’d be the best reliever in baseball. I settled for putting them up to 200 innings and letting that over-allocaiton of playing time handle the different measures of success. For example, a reliever projected for 3.6 WAR in 200 innings would check in around 1.2 for a full season of bullpen work. That’s a very good relief pitcher projection; only 20 players meet that bar in our 2025 Depth Charts projections.

In other words, the tier names still mostly work for relievers, but you should apply your own relative positional value adjustments just like normal. A star reliever is less valuable than a star outfielder. A star starting pitcher might be more valuable than a star outfielder, depending on the degree of luminosity, but that one’s much closer. This outcome table can guide you in terms of what a player might turn into. It can’t tell you how to value each of those outcomes, because that’s context-specific and open to interpretation.

This study isn’t meant to be the definitive word on what prospects are “worth.” Grades aren’t innate things, they’re just our team’s best attempt at capturing the relative upside and risk of yet-to-debut players. Being a 60 FV prospect doesn’t make you 17% likely to turn into a star; rather, our team is trying to identify players with s relatively good chance of stardom by throwing a big FV on them. And teams aren’t beholden to our grades, either. They might have better (or worse!) internal prospect evaluation systems.

With those caveats in mind, I still find this extremely useful in my own consumption of minor league content. The usual language you hear when people discuss prospect trades – are they on a Top 100, where do they rank on a team list, what grade are they – can feel arcane, impenetrable even. Breaking it down in terms of likelihood of outcome just works better for me, and I hope that it also provides valuable information to you when you’re reading the team’s excellent breakdown of all things prospect-related this week.





Ben is a writer at FanGraphs. He can be found on Bluesky @benclemens.

112 Comments
Oldest
Newest Most Voted
Cool Lester SmoothMember since 2020
1 year ago

Absolutely incredible piece, Ben!

soddingjunkmailMember since 2016
1 year ago

Yeah, this is the stuff I come to Fangraphs for. Thanks Ben.

darren
1 year ago

Agreed!

Broken BatMember since 2020
1 year ago

Great stuff. Was wondering Ben how you treated a player during the time period where the FV started below 45 but in the next year or more increased to FV45 or better? Were these “risers” added to data? If not, Trying to populate to see what this group looked Iike vs. initial FV45+ guys might also be interesting. Keep your research coming. Very enjoyable.

PC1970Member since 2024
1 year ago

Good stuff & fun article- Couple questions:

  • Just verifying that players are on the list multiple times, I.E., if someone is a 45 one year, a 55 the next, then a 50 in year 3, they count for all 3 buckets? I’d assume so, since the add’l information attained changes their rating & likelihood of becoming a star (or a washout).
  • It also makes sense that the higher rankings are more likely to be a star. Those are usually given out to players that have succeeded in at least High A/AA, which is a huge indication of talent, ability to make adjustments, etc. OR are Top 5 draft picks, which is an indication of high level amateur skill, esp if it’s a college player.
  • Who is the 70 starter that is “just” a regular? Mackenzie Gore?
Jorge FabregasMember since 2016
1 year ago
Reply to  PC1970

Seems like it would have to be Gore, as he’s the only 70FV pitcher in the sample. And if the projections were taken 4 years later instead of 3, then he would’ve been in the above-average group.

bubblesMember since 2024
1 year ago

Prospects week is here!

Jorge FabregasMember since 2016
1 year ago

Very interesting. One takeaway–from what I recall, Longenhagen has some of the harshest prospect grades in the industry. And yet, for 50 and below, I wonder if they should be shaded down half of a grade–as the median outcome for 50FV hitters seems to be the very low-end of the “regular” group.

A Salty ScientistMember since 2024
1 year ago
Reply to  Jorge Fabregas

Agreed. And also wondering if we should be thinking most about modal and cumulative outcomes. A 55 has at least a 67% chance of being a regular or better vs 46% for a 50.

tdmocMember since 2023
1 year ago
Reply to  Jorge Fabregas

He didn’t use to (look at his 2017-2018 grades and compare to 2024) , although I find that simply a reflection of how often prospects actually pan out and how other publications may feel obligated to generate excitement (this is obviously speculation, of course).

Jorge FabregasMember since 2016
1 year ago
Reply to  tdmoc

That seems likely. IIRC, MLB Pipeline gives every player on the top 100 at least a 50, and they probably give out 50s beyond that.

Actually I just checked and they give out 55s to the bottom of the 100! And I don’t think this is because their analysts are inferior (it seems like Jim Callis is considered to be the best in the industry by his peers), so it must be that there’s grade inflation built in.

Cool Lester SmoothMember since 2020
1 year ago
Reply to  Jorge Fabregas

MLB essentially give everyone 5 more than FG would.

sadtromboneMember since 2020
1 year ago
Reply to  tdmoc

Eric has become much more conservative with grades over time (and to a lesser extent, Kiley McDaniel at ESPN).

He and Kiley did a ton of work here a few years ago (I think some of which wound up in Future Value as well) that established a connection between the WAR distribution and the FV grades. Then they wound up aligning the grades with the distribution. This means that there are some things that rarely happen, like almost no pitchers get an FV65 or FV70, because they are so rare.

cowdiscipleMember since 2016
1 year ago
Reply to  sadtrombone

Either that or we just suck at identifying them! FV40 on Cole Ragans, for example, could’ve been a FV65 or FV70 in hindsight.

Last edited 1 year ago by cowdisciple
sadtromboneMember since 2020
1 year ago
Reply to  Jorge Fabregas

IMO this seems pretty accurate to me. An FV50 is supposed to be a regular, and the fact that it’s in that bucket is reassuring. There’s always going to be some disconnect between the individual grades and the overall distribution.

slamcactusMember since 2024
1 year ago
Reply to  sadtrombone

Part of that is that the FV grade will always be a blunt tool, albeit a necessary one if the exercise involves putting everyone on a uniform scale.

Normally we think of probabilities along a bell curve. But I bet if you sat down with scouts to talk about the range of outcomes for say…2022 Elly de la Cruz, it wouldn’t look anything like a bell curve. With huge bust potential but a near-certainty that if he made his skill-set work, he’d be a star and very little possibility he’d settle in as just ok, it would look more like a Bactrian camel (that’s the two-humped one, right?). For a boom-or-bust guy like that, it’s hard to convey a lot of meaning by assigning a number like 50 or 55FV.

But, of course, we also want to know about those guys. So typically they get lumped in with the guys who have much more traditional quartile/decile projections, and 50FV becomes more of a hedge against the downside than a statement about the most likely/median outcome.

slamcactusMember since 2024
1 year ago
Reply to  Jorge Fabregas

2 things: 1) if there were a problem in terms of definition vs results, wouldn’t that reflect an issue with the individual grades, rather than with the grading system itself? and

2) Um…isn’t that a very good result? A 50FV grade is, roughly speaking, an opinion that a given prospect has a median projection as a regular, someone who produces in the 1.6-2.4 WAR range (Ben used 1.5-2.5, but 1.6-2.4 is how Eric/Kiley originally mapped it out).

If 50% of FV hitters are turning into regulars or better, and 50% aren’t, isn’t that a pretty good indication that Eric’s doing a really good job of assigning those grades? In statistical terms, that means half of the guys are playing at or above their median projections, and half aren’t. I’m not a data scientist, but to my layperson’s understanding when you’re working with models based on probability, about half performing at or above the median and half performing below it is pretty much exactly what you’re going for, no?

Last edited 1 year ago by slamcactus
Brian ReinhartMember since 2016
1 year ago

I (sincerely) love that Prospect Week has escaped the containment of a single week.

dbannonMember since 2022
1 year ago

This is a great piece. I think it would overwhelm the analysis, but I’m curious how the added layer of risk plays in… I know FV numbers discount risk, but a risky 55 and a safe 50 presumably have different outcomes, and I’m curious how different teams quantify/qualify this

Thom with an HMember since 2017
1 year ago
Reply to  dbannon

Maybe it’s just a different way of saying the same thing, but I wonder how age affects things. Does a 23-year-old player with a 55 FV have less variance than a 19-year-old 55 player? I would assume so, but “risk,” despite being far less granular than age, might be a better variable as it subjectively incorporates age, injury, and traits/flaws.

jrp1918Member since 2024
1 year ago

Great article. Those washout rates need to be stapled to every trade article when fans lose their minds over giving up 45/50 FV guys to get an actual major leaguer

baseballfan115Member since 2020
1 year ago

Thank you! This is amazing.

How do you think the numbers might change if you stretched it to 5 year forecasts, giving pitchers a better chance to develop?

Also, this might be rephrasing the question asked by another commentator, but did age come up as a factor on “hit” rate? As in the college hitters given a 50 were more likely to end up solid regulars than the younger prospects given the same grade?

MikeSMember since 2020
1 year ago

TINSTAAPP is greatly misunderstood, but I think the second half of this post explains the concept pretty well.

Especially this line:

Pitching prospects just turn into major league pitchers in a less predictable way

Last edited 1 year ago by MikeS
David (Dave) Roller
1 year ago

Very Interesting. It seems like the last 3 levels of pitchers and hitter have too small of a sample size. Would you consider consolidating the bottom 3? Other wise if next year you get the rare player with an extremely high fv who only becomes a back up or gasp a washout it would skew the data.

guyarrigoniMember since 2024
1 year ago

Am I reading this wrong or could one conclusion be that we are giving too much weight to “deep” systems and not enough to “top heavy” ones?

OkraMember since 2016
1 year ago
Reply to  guyarrigoni

I had the same thought. The farm system rankings here on fangraphs do address this by giving a lower $ value to each lower FV grade. However, I think 40 and 35 FV prospects are still a bit overvalued. For example, Arizona currently has 20 prospects with a FV of 35+ which equates to $10M in value. A 45+ hitter is valued at $8M. Would any team be willing to give up their 45+ hitter for 20 35+ lottery tickets? I doubt it. I wouldn’t.

For sure teams fall in love with a random 40/35 prospect and request them as a trade throw-in; but as a whole they are not super valuable. For that reason i don’t really think including anything below a 40+ is very helpful in farm rankings.

NatsNationMember since 2020
1 year ago
Reply to  Okra

It seems to me that there is context that applies in the lower end of the FV ratings in a system. In particular, a lot of starters with relief risk end up in the 35+ range (I think off the top of my head). Their WAR ceiling is clearly limited, but they can be the pipeline that a real team needs to staff a playoff quality bullpen?

I think that there are plenty of scenarios where I trade a 45FV (DH only slugger in his late 20s) for 20 35+FV “lottery picks” (reliever risk, skill sets that align with my systems development strengths, etc).

I like the Fangraphs existing system ranking approach. It allows me to see the most data and draw my own conclusions.

I totally agree with your comments from a Fantasy perspective.

bookbookMember since 2024
1 year ago
Reply to  Okra

You’re arguing that there’s no way 20 35+ FV picks could outvalue the 31% chance 1 45+ FV player has of being a regular or better? Couldn’t the 35+ FV guys average a 2% chance of making regular or better? I think they could. Roster scarcity. At this extreme, does become a problem, but in a frictionless universe….

OkraMember since 2016
1 year ago
Reply to  bookbook

Yeah exactly – you can only roster so many guys in your farm system and still have them development properly, right? So due to the roster scarcity I think large sums of 35+ prospects begin to lose value.

TJMember since 2020
1 year ago
Reply to  guyarrigoni

I think the conclusion is that we have a way to quantify such systems’ quality. (Or rather, these probability distributions are a second way, in addition to surplus value.)

slamcactusMember since 2024
1 year ago
Reply to  guyarrigoni

I ring this bell as often as I can. In most years the 50FV tier extends from prospect ~30 or so to about prospect ~130, and I’d bet there are at least another 150 or so 45/45+ guys.

Put another way, the difference between prospect ~35 and prospect ~280 is generally a half-grade, and qualified evaluators can and do have half-grade differences of opinion on guys all the time.

I’d rather have 2 60s or 3 55s and nobody else in the 50FV tier than 6 50s.

slamcactusMember since 2024
1 year ago
Reply to  slamcactus

Obviously it’s not quite so simple, as there are tiers within tiers, and I bet most people would see a bigger difference than that between a “high 50” and a “low 45.”

But it’s simple enough that you can say pretty comfortably that not a ton of expected future value separates a prospect ranked 80th or so and a prospect ranked 200th (if rankings went that far, which most don’t).

treebeardedMember since 2019
1 year ago

Love this measure of standard deviation. Question – did this table include Wander Franco’s 80 FV ranking?

SaltyChipsMember since 2021
1 year ago
Reply to  treebearded

He’s the one 80 FV batter that made the table I believe. I’m not sure the last time there was an 80 grade prospect before him but my memory is poor.

3cardmontyMember since 2025
1 year ago
Reply to  SaltyChips

He was the first one

OlanMember since 2020
1 year ago
Reply to  3cardmonty

So sad

pmart1995Member since 2016
1 year ago
Reply to  Olan

What were Puig and Jose Fernandez’s FV’s? I remember them being very high, but maybe not 80.

Matt
1 year ago

Very interesting way to look at prospect projections vs MLB player projections of those prospects.

It looks like a 55 FV is roughly 3x more likely to be a starter (or better) than a 45 FV while a 45 FV is 3x more likely to washout. But I’d be very interested in seeing how the numbers workout with a larger sample, or over a longer timeframe.

Perhaps I missed it, but what projection system did you use?

JameyMember since 2020
1 year ago

One thing I think is important translating the usual language of prospects is that a Top 100 prospect is a 50 FV or better prospect (and, if referred to in those terms, probably is a 50 FV because people at the top end of the scale tend to be referred to more precisely).

OddBall Herrera
1 year ago

Something’s not right here, or I’m not understanding the methodology. Lux last showed on the 2020 prospect report as a 70 FV. That means we should be looking at his 2023 projections, which say 2 WAR. How is there a 100% star rate on the 70 FV prospects?

OddBall Herrera
1 year ago
Reply to  Ben Clemens

I see, so he was a 60 in the 2019 mid season, moved up to 70 2020 beginning of the season but dropped off the list by mid season so he probably went into your numbers as a 2019 60 (can’t get to that google doc at the moment to validate)

Not sure enough guys follow that particular trajectory to mess with the numbers in terms of evaluating them against the player’s *peak* FV, though I would bet most of them who do are higher FV players.

That makes sense though, thank you for the response!

ScoreboardMember since 2016
1 year ago

I’ve wanted this assessment for so long. Thank you for bringing this eval to us fortunate and grateful readers.

booondMember since 2019
1 year ago

Great article.

The one caveat is the lack of data. Not Ben’s problem though going backward a bit more may have smoothed out some numbers. However, at the highest levels you’d still have issues with tiny samples.

melyacht
1 year ago

I find it mildly interesting that roughly 50% of the 50 FV prospects in this sample became big league regulars or better. (53% of hitters and 46% of pitchers.)

The GuruMember since 2026
1 year ago

This is the probalem with systems like zips…..50% chance theyre war total is 0. Thats why he overvalues prospects.

GeorgeMember since 2021
1 year ago
Reply to  The Guru

FV is a probability distribution so finding that this probability distribution has 50% below the median expectation is not insane at all. Also its not ZIPS. That is an entirely different thing.

The GuruMember since 2026
1 year ago
Reply to  George

I know that. I’m saying zips shouldnt be using reversion to means on propsects. That is antiquated method and why very few are even looking at zips these days outside of casual fans. Teams are way past that method. I don’t know how he’s going to hang 4 war on a AAA guy and not laugh….he need to be using some system like this.

si.or.noMember since 2017
1 year ago
Reply to  The Guru

> I don’t know how he’s going to hang 4 war on a AAA guy and not laugh

Wut? Most of the star projections are for players _already_ in the big leagues, but who were prospects 3 years before.

carterMember since 2020
1 year ago

I am curious what the percents are for hitters who rank below 40, or at 40.

scottsjunk1981Member since 2021
1 year ago
Reply to  Ben Clemens

But couldn’t you pretty simply label all the players without projections as “washouts”? I mean, getting injured and retiring is a different kind of washout than becoming an org guy, but it still means they aren’t generating any WAR from now on.

And I don’t know if I agree about the mean value being less relevant when the probability of success is low. It seems like it matters a lot whether it’s 3% or 6%, and if you’re able to find cohorts – like certain sets of tool grades – that meaningfully over or under perform their FV peers, I’d think that would be incredibly helpful.

CosmoMember since 2024
1 year ago

Last year Cowser and Butler both had 45+ grades. Are they still viewed in that manner? Do either of them now have star or above average futures?

tdmocMember since 2023
1 year ago

Regarding pitcher grading and outcomes, I frequently think about this off-the-cuff Eric monologue from an EW episode last year:

I tend to lean on […] the stuff that plays, like, you look at the WAR leaderboard, and it’s big, strapping dudes who have the prototype body and delivery. Their delivery is almost always beautiful and fluid and really athletic. It is Zack Wheeler, it is Spencer Strider, it is Gerrit Cole, it’s Sandy Alcántara, it’s Dylan Cease. Mitch Keller figures it out, you know what I mean? Do you have the body, do you have the delivery, do you have feel for spin, all that stuff over time tends to be what allows these guys to bubble to the surface and be Logan Gilbert and George Kirby and be Charlie Morton even if it takes until his 30s to do it. […] Mick Abel didn’t have a great year. Well, he’s 6’5″ and has elite arm speed, his delivery looks like Gerrit Cole’s […] over time, that guy is just going to be an absolute monster […] so I’m not going to overreact to him walking a bunch of guys. That was true of Sandy Alcántara […] it was true of Corbin Burnes, it was true of Dylan Cease. Stay on the guys where the *look* is right when it comes to pitching.

si.or.noMember since 2017
1 year ago
Reply to  tdmoc

So do we need a LOOK column on the pitching side of the prospect board?

tdmocMember since 2023
1 year ago
Reply to  si.or.no

I’m not sure how religiously updated this page is, but we already do?

Last edited 1 year ago by tdmoc
rwperu34Member since 2025
1 year ago

I like the methodology. An offshoot of this methodology would be to take the projection at age 27 ir 28, which in many ways is what scouts are projecting.

Matt GreeneMember since 2025
1 year ago

Prospect week creep = 10 days of prospect week

(I’m here for it) (unlike christmas decorations on nov 1)

OtisMember since 2019
1 year ago

Wow, this is amazing. Thank you, Ben! And of course thank you to Eric as well for his incredibly valuable prospect assessments.

Bruce McClure
1 year ago

Ben, this is amazing work. Kudos!

jtricheyMember since 2021
1 year ago

I may be wrong, but aren’t prospect grades part of the ZIPS projections? Therefore wouldn’t that hurt the data a bit because one variable already affects the other? Am I wrong about this?

GeorgeMember since 2021
1 year ago
Reply to  jtrichey

I am decently sure that ZIPS is FV-agnostic.

Thom with an HMember since 2017
1 year ago

I’d love to know if there are any trends within position groups. Are shortstops and center fielders more likely to live up to their FV than catchers or first basemen? It’s possible the differences are already baked into their FV number, and I realize breaking it up by position makes the sample sizes even smaller.
But it’d be cool to know if, let’s say, it’s easier to identify middle infielders than others, and corner outfielders are a struggle (or whatever it might actually be).

JoshuaMember since 2017
1 year ago
Reply to  Thom with an H

Interesting thought. I remember some years ago Baseball Prospectus was advancing a theory of SBPODE: second base prospects often don’t evolve. The natural following question being whether there’s something special about second base prospects, or whether SBPODE is just a result of looking to closely at PODE.

NATS FanMember since 2018
1 year ago
Reply to  Thom with an H

I’d bet different franchises are better or worse at certain positions rather than positions themselves. Although, minor league starting shortstops often end up almost anywhere on the field in the end.

Roger McDowell Hot Foot
1 year ago

This is fascinating to look at, but it’s very unclear how much signal there is in all the noise. The sample sizes are small, the projections are projections and not actual measured performance any more than the prospect grades… it’d be interesting to revisit this same pool of players in, say, 5 years and look at what happened to see if anything meaningful can be prised out of it. As it is, it’s hard to look at that data and conclude anything at all.

Jason BMember since 2017
1 year ago

In that case, stay tuned for:

“How’s My Driving: 2018 Top 100 Audit”

coming up later in the week!

Roger McDowell Hot Foot
1 year ago
Reply to  Jason B

Based on my proprietary projection system, that article grades out as exactly what I’d have predicted based on this one.

scottsjunk1981Member since 2021
1 year ago

I know you run into sample size issues pretty quickly, but I’d be curious to know splits – at least within the FV50 and lower tiers – by max level, age, projected position, and underlying tool grades, among other things.

I’d think you’d be able to populate one dimensional splits, even if full xtabs are a nonstarter.

TommyfastballMember since 2016
1 year ago

Sorry, not sorry, for the snark, but maybe the top-100 list would be a little better if you released it at the trade deadline next August?

CosmoMember since 2024
1 year ago
Reply to  Tommyfastball

FanGraphs’ way of doing things is a known quantity. You have several other good choices if this one doesn’t suffice.

3cardmontyMember since 2025
1 year ago
Reply to  Tommyfastball

There’s a major update released before the trade deadline

TommyfastballMember since 2016
1 year ago
Reply to  Tommyfastball

This was a great article. Thank you. I wonder if you could do the reverse…take the current major leaguers in the buckets (or use guys with 2-4 years experience) and run the distribution of prospect grades. I know you’ll have multiple prospect grades for many players, but you’ll figure it out….

Lil NubberMember since 2026
1 year ago

Forrest Whitley’s prospect outcome is “All-Star?” What?

Lil NubberMember since 2026
1 year ago
Reply to  Ben Clemens

Just looking at the Board, Whitley was a 65 in the 2019 report (60 was the 2019 updated) and the only 65 pitcher in the sample is listed as an All-Star. Is there a different 65 pitcher I’m missing?

Lil NubberMember since 2026
1 year ago
Reply to  Ben Clemens

I’m confused then — Whitley is listed as a 65 on the 2019 report — why do the outcomes only include Perez?

dbminnMember since 2026
1 year ago

Great work, Ben. I wish FanGraphs made season opening projections from past years available to subscribers.

Last edited 1 year ago by dbminn
dbminnMember since 2026
1 year ago
Reply to  Ben Clemens

woo hoo!

fangraphsreaderbutwokeMember since 2025
1 year ago

Would love to see this experiment run by GM/team too. I intuitively _know_ Ben Cherington has had higher rates of washout hitters than most other GMs, for example.

mr.met89Member since 2024
1 year ago

Who the hell was an 80 grade prospect???

KinanikMember since 2016
1 year ago
Reply to  mr.met89

Wander Franco.

evo34Member since 2023
1 year ago

Nicely done. Wondering if next time you could look at how indiv skill grades (raw/game power, present and future, etc.) correlate to MLB success.

Greg SimonsMember since 2016
1 year ago

This is excellent, Ben!

“depending on the degree of luminosity” is an absolutely wonderful turn of phrase!

pepper69funMember since 2020
1 year ago

Some prospects are a 50 due to high probability of a modest floor and are lumped in with 50’s with much higher ceiling and higher risk. If we are using this for fantasy, then historically, which type of 50 is best?

TrigauxMember since 2016
1 year ago

I’d love to see this methodology applied to different ranking systems — is Law, BP, Fangraphs, etc. better at predicting future performance?

KyleMember since 2024
1 year ago

Is it intentional that the headings for pitchers and hitters are different?

KyleMember since 2024
1 year ago

I tried to take this and combine with the Farm system rankings to see which teams would have what expected count in each category (only including 45 and up). I just multiplied the outcome % by the count in each category (e.g. 50 FV Bat) for each team and summed them up. Sorry if the formatting sucks.

+----+-----+-------+-----------+--------+---------+--------------+------+
| Rk | Org | Count | WashedOut | Backup | Regular | AboveAverage | Star |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 1  | TBR | 11    | 4.28      | 2.37   | 2.21    | 1.74         | 0.42 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 2  | LAD | 14    | 5.72      | 3.25   | 3.02    | 1.86         | 0.19 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 3  | BOS | 13    | 5.46      | 2.92   | 2.62    | 1.71         | 0.32 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 4  | CHW | 9     | 2.62      | 2.05   | 2.44    | 1.77         | 0.15 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 5  | CHC | 12    | 4.53      | 2.92   | 2.76    | 1.64         | 0.16 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 6  | BAL | 7     | 2.65      | 1.5    | 1.42    | 1.17         | 0.28 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 7  | NYM | 9     | 3.44      | 2.11   | 2.13    | 1.24         | 0.1  |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 8  | NYY | 9     | 3.53      | 2.07   | 1.96    | 1.31         | 0.16 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 9  | SEA | 8     | 2.71      | 1.89   | 2.02    | 1.27         | 0.12 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 10 | DET | 10    | 3.99      | 2.38   | 2.28    | 1.28         | 0.09 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 11 | CLE | 10    | 4.26      | 2.28   | 2.11    | 1.23         | 0.15 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 12 | ARI | 8     | 3.32      | 1.71   | 1.63    | 1.12         | 0.24 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 13 | PHI | 6     | 1.92      | 1.58   | 1.43    | 0.99         | 0.08 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 14 | WSN | 7     | 2.71      | 1.57   | 1.42    | 1.07         | 0.24 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 15 | COL | 11    | 4.8       | 2.63   | 2.35    | 1.18         | 0.06 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 16 | ATL | 10    | 3.89      | 2.47   | 2.28    | 1.29         | 0.08 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 17 | MIL | 8     | 3.23      | 1.88   | 1.82    | 1.02         | 0.07 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 18 | MIN | 7     | 3.02      | 1.62   | 1.44    | 0.84         | 0.1  |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 19 | PIT | 12    | 5.29      | 2.87   | 2.53    | 1.25         | 0.08 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 20 | MIA | 9     | 3.85      | 2.11   | 1.92    | 1.06         | 0.08 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 21 | STL | 6     | 2.08      | 1.56   | 1.35    | 0.92         | 0.09 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 22 | TEX | 7     | 3.03      | 1.71   | 1.37    | 0.79         | 0.11 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 23 | TOR | 8     | 3.4       | 2.07   | 1.59    | 0.87         | 0.07 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 24 | CIN | 10    | 4.56      | 2.49   | 1.91    | 0.97         | 0.08 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 25 | OAK | 5     | 2.07      | 1.28   | 1.03    | 0.57         | 0.05 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 26 | SDP | 3     | 0.9       | 0.62   | 0.65    | 0.64         | 0.19 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 27 | KCR | 4     | 1.79      | 0.93   | 0.82    | 0.43         | 0.04 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 28 | SFG | 6     | 2.58      | 1.53   | 1.2     | 0.63         | 0.06 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 29 | HOU | 4     | 1.76      | 0.99   | 0.81    | 0.39         | 0.05 |
+----+-----+-------+-----------+--------+---------+--------------+------+
| 30 | LAA | 5     | 2.18      | 1.25   | 1.04    | 0.5          | 0.03 |
+----+-----+-------+-----------+--------+---------+--------------+------+
KyleMember since 2024
1 year ago
Reply to  Kyle

Caveat: I’m just doing this for fun and realize that SSS would really throw off the top tiers. But I like the idea here that instead of just a straight ranking you could show the likelihoods of different types of players for a farm system. Food for thought.

3cardmontyMember since 2025
1 year ago

Long live Prospect Fortnight

NATS FanMember since 2018
1 year ago

Wasn’t Moncada a fv70? I do not consider him a star. So he does not fit your data. Unless one season of 5 WAR in 9 seasons is all you need using this model. he has 13 war in those 9 seasons. How am I wrong?

NATS FanMember since 2018
1 year ago
Reply to  NATS Fan

oh i guess he is outside the data sample.

pmart1995Member since 2016
1 year ago

Echoing other posts, this is excellent! One thing I’m curious about is the predictiveness of other tool grades, both for overall future value and for the tool in question. Put another way, which tool grade has the highest correlation with future value (e.g., power for hitter, fastball or command for pitcher)? And how predictive are power grades for future HR or ISO, and speed grades for SB?

DanMember since 2018
1 year ago

Great article, Ben.

Thomas BertonMember since 2020
1 year ago

Really fascinating article Ben! One note – the columns are labeled differently for Position Players and Pitchers. Is there a reason for that (I’m assuming just a change during editing that got missed)

MannybeingMannyMember since 2016
1 year ago

Ben, this is way more entertaining of an audit than the ones at my corporate job.
Really excellent piece and extremely useful to the whole community.

One question i’m curious about: What % of mlb players that have made their debuts over the past x seasons have actually even had prospect projections? And what percentage falls within each prospect grade (or no prospect grade)? From that, then potentially doing a similar kind of analysis as you did above with WAR to then break that out further to see the actual value of that MLB time?

I’m really just curious about how many mlb players of recent never had prospect projections :).

Thanks a bunch Ben, really enjoyed the article!

Slacker GeorgeMember since 2016
1 year ago

This is great work. Prospect grades are evaluations, not height measurements, so a 2024 55+ might be different than a 2020 55+, as the evaluator’s skills may “improve” and the grades become better predictors. Breaking down by year may throw too much variance due to smaller sample sizes, so I don’t think it is possible to tease out any more nuance doing so.

darren
1 year ago

You make a really interesting point about how difficult it is to evaluate starters vs. relievers, but I disagree with how you handled the relievers. Relievers have far less value than starters do, which is why they are paid far less and most organizations will put considerable effort into making their best pitchers starters before resorting to making them relievers. For them to top out somewhere around average regular makes perfect sense.

The idea of ‘normalizing’ relievers to 200 IP is especially problematic. First, most relievers have demonstrated that, due to either stuff or durability, they cannot handle a full-time role. Second, relievers have leverage baked into their WAR. When normalize their WAR to 200 IP, you’re not just giving them credit for innings they are almost certainly incapable of pitching, you’re giving them triple credit for leverage shouldn’t be included. (The 200 IP marker is similarly, though far less, problematic for starters.)

In the end, the result is distortion of their value as prospects compared to other pitchers. It assumes teams would be just as happy with a 60 FV starter turning into a 1.3 WAR reliever as they would be with him turning into a 4 WAR starter.

wlx17Member since 2025
1 year ago

who was the 80 grade prospect?

deanmachine5488Member since 2019
1 year ago

This is a great piece. Interesting just for the sake of baseball and I’m going to be using this calculator when I think about trades in my dynasty league! I’ve always felt folks in my league over-value prospects and this kind of backs that up.

wandererkentMember since 2025
1 year ago

Great foundational peice to build on analyses for years to come.