Why I Voted for Michael Fulmer
These days, some people are hard at work trying to understand one another. As votes rolled in and results were released, segments of the population were taken aback. In certain corners, the mood has been celebratory, triumphant. Elsewhere there has been fear, disappointment, and, more than anything else, confusion. “How could anyone make that choice?” many have wailed. “How could so many people overlook all the evidence?” I won’t pretend to be more than I am. I know that I am but a single man in a rising, roiling ocean of souls. But I can try to defend at least my own decision. For 2016 American League Rookie of the Year, I voted for Michael Fulmer ahead of Gary Sanchez.
Unsurprisingly, I was far from alone. Here are the results, straight from the BBWAA. Fulmer won! He’s the newest AL RoY, picking up 26 of 30 first-place votes. Sanchez got the other four first-place votes, so he finished in second. Tyler Naquin wound up a distant third. The results mirror my own ballot, which went Fulmer — Sanchez — Naquin. This was my first-ever year in the BBWAA and so it was my first-ever ballot, which meant I thought about this award more than I usually do. Even though I voted for the landslide winner, I feel almost obligated to explain myself.
Fulmer pitched a full major-league season, and he ended with a low ERA. It would’ve been very easy to just stop thinking there. And that sentence is why Fulmer was the obvious favorite, before the voting results were made public. It was hard to imagine Sanchez generating enough momentum, given that him winning would’ve been essentially unprecedented. I did try to give Sanchez a shot. The idea of him making history was appealing to me. I just talked and thought myself out of it.
Any sensible conversation about Fulmer had to focus on his ERA. More specifically, one needed to address the fact that Fulmer had a 3.06 ERA, but a 3.76 FIP. Put differently, by one WAR, Fulmer beat Sanchez, 4.6 to 3.2. By another WAR, Fulmer lost to Sanchez, 3.2 to 3.0. That, despite the advantage of playing a full year. That’s basically what opened the door to Sanchez in the first place. By one measure, he was as valuable in two months as Fulmer was in five.
In the middle of September, I took a look at Fulmer’s case for the award. I couldn’t find much evidence that Fulmer could sustain the same level of run prevention. He doesn’t seem like someone who’s going to run a huge ERA – FIP difference. Now, there are two things about that. One, awards look back. What’s done is done, and one could argue that Fulmer earned his results, by the process of generating his results. Two, let’s say you want to mentally regress Fulmer’s numbers. You want to believe far more strongly in his FIP. Why limit that to Fulmer? Why not apply the same thinking to Sanchez?
Fulmer, you could say, outdid how you’d expect him to perform. We make it easy to think in these terms, because we have the numbers listed on every pitcher’s player page. We don’t do the same for hitters, but that doesn’t mean similar principles don’t apply. To just get right to it, I’ve prepared a plot, using data pulled from Baseball Savant. The simple analysis here mirrors something I did when investigating Bryce Harper some months ago. On the x-axis, you see 2016 average exit velocity for balls hit into the air. On the y-axis, you see each player’s air-ball slugging percentage. I set a minimum of 50 batted air-balls.
Despite the sample sizes here, and despite the simplicity of the model, the shown relationship is strong. And I’ve clearly highlighted the point corresponding to Gary Sanchez. One takeaway, for sure, is that Sanchez generated upper-level exit velocities. But his performance on air balls was extreme, and disproportionately so. His air-ball slugging percentage was 1.569. Based on the best-fit line, one would’ve expected it to be 1.051. That would’ve dropped Sanchez’s overall slugging percentage from .657 to .471. That’s doing things very simply, and unfairly simply, but Sanchez topped his “expected” slugging by 518 points. No one else was even close to him.
The argument for Sanchez rests on how incredible he was within a two-month window. You can get behind the idea of supporting a two-month catcher who slugged .657. It’s not nearly so easy to get behind a two-month catcher who slugged .471. That’s still a very good player and a very good rookie performance, but that no longer seems like enough to outweigh a mostly full year of starting pitching. Just as FIP takes some of the shine off Fulmer’s ERA, the above analysis takes some of the shine off Sanchez’s slugging, or OPS, or wRC+, or whatever. It’s not as easy to get there, but it’s valid. If you’re going to give Sanchez credit for his results, you should mostly give Fulmer credit for his results. If, instead, you want to regress what Fulmer did, you have to do it to Sanchez, too. Otherwise you’d be tipping the scales.
There is one thing that still gives me a little trouble. On a rate basis, I believe Sanchez was better than Fulmer was. And, moving forward, I believe Sanchez is going to be the more valuable player. He’s a pretty good defensive catcher with a hell of an arm and a hell of a powerful bat. Fulmer seems like a No. 3 starter, maybe a No. 2. Sanchez feels like he could be a perennial All-Star. It’s caused me to question what, exactly, the Rookie of the Year is, and I still don’t have a satisfying answer. I gave my first-place vote to a player who I think is inferior to the player I gave my second-place vote. It’s weird. I wish that the rules were somehow more explanatory.
But, again, awards, by their nature, look back. When you look back, you try to minimize your speculation. When I think about Sanchez being the better player, some of that is just my own scouting. But I suppose the award should recognize the player who had the best major-league season. As such, playing time has to matter for something. It can’t all be about rate stats. Fulmer was good enough, for a long-enough time. Sanchez was close, and I wanted to vote for him, I really did. I thought it would be a thrill. I just couldn’t get myself all the way there, and I hope that my explanation is satisfactory.
Jeff made Lookout Landing a thing, but he does not still write there about the Mariners. He does write here, sometimes about the Mariners, but usually not.

You put up a good showing your first year kid. Aren’t you the least bit concerned that your vote mirrored the great unwashed sports community?! And, yes, I mean that with humor.
More seriously, it obviously is a difficult choice. If I had a vote, which I don’t (so the sporting world is safe), I would have leaned towards Sanchez. It’s the historic nature of what Sanchez accomplished that would have pulled me over the line, and he had the higher fWAR. History and more value by one key measure. We’ll see another Fulmer-like season, probably as soon as next year, and many more over the years. It was a nice season. I’m not sure we’ll see what Sanchez did for quite some time. Historically, I believe there’s a better chance that when we look back on both Sanchez’s and Fulmer’s careers, many will likely decide Sanchez should have won.
Think about it. If you’re really battling the decision between these two players, which player would you rather have moving forward? I’ll take Sanchez. I think history will too.
The point about Sanchez’s season being historic makes sense to me. As for ‘which player would you rather have going forward, though, how far should you take that logic? Given how often we’re wrong about players, I don’t feel all that comfortable basing the ROY on how we expect careers to play out. Otherwise, the award would always go to the best prospect who debuted and didn’t bomb. I’d rather focus on who performed this year, and let the future play out before rewarding guys for it.
It’s occurred to me that it’s possible to win ROY more than once. I’m not sure what the minimum number of games or PA is required to constitute a rookie season, but Sanchez illustrates that it’s conceivable that a player could fail to reach this minimum and still win the award, making him eligible the following season as well. Trout played 40 games in 2011, not much fewer than Sanchez this year, and was still able to win ROY the following season.
Technically, Edinson Volquez could have won ROY four times
I’m sure a BBWAA member voted for him, following up on his Gold Glove vote last week for Rafael Palmeiro.
Just for context, the PA cutoff is 130 (a little over 1 month). Barry Bonds in Aug 2004 hit 11 home runs, walked 38 times (wtf), and put up a 278 wRC+ for what was one of the best offensive months ever.
He tallied 2.7 war while playing solid if unspectacular LF defense. Prorate to 130 PA and it becomes 3.1.
So, yeah. “Possible”
The cut off is 130 AB’s, not PA’s.
http://mlb.mlb.com/mlb/official_info/about_mlb/rules_regulations.jsp
Fascinating! So, we just need a rookie who can match Bonds’ 30% walk-rate.
It is possible, but Sanchez eclipsed the rookie limits at some point in September. And while it’s possible, it’s almost likely impossible. If Sanchez couldn’t win with his two blistering months, it’s unlikely another player would be able to do so with even fewer games and plate appearances. Then again, never say never. I suppose it could happen in a year when there’s just no real competition.
The most “likely” scenario for a player to receive RoY votes in multiple years is for a pitcher, not a position player. The ceiling for maintaining rookie eligibility for a pitcher is less than 45 days on the 25 man roster (i.e., before roster expansion) and less than 50 innings pitched. So …
Imagine a rubber armed Andrew Miller type reliever who comes up the last week of July and and goes 5-0 while recording 20 saves in as many opportunities over the last two months of the season while pitching just slightly less than 50 innings.
If there is someone with access to BB-Ref’s database around here — would be a pretty easy query to identify players at least receiving votes in multiple years for RoY.
You’re really going above and beyond here, Jeff. In a world where someone LEFT SANCHEZ OFF THE BALLOT, merely voting Fulmer over Sanchez requires no explanation.
Damn. Donald Trump is president and Jeff Sullivan is very wrong about something.
Who is the dot all the way on the left? The one whose “fly balls” seemed to be wiffle balls?
That’s Billy Burns. And amazingly, it’s not just fly balls but all balls hit in the air, including line drives.
I would have voted for Sanchez based on the fact I’d be willing to trade Fulmer to get Sanchez in ye olde proverbial vacuum.
I think that’s the wrong question though. I agree with Jeff that the award looks backwards, so the question really should be which player’s performance this past season would you have traded for? That is, for just this past season, would you rather have had Fulmer or Sanchez on your team?
Nice try, Jeff. And you think you’re going to get away with not having to explain why you voted for Corey Seager….
I am of course kidding. Or maybe you didn’t vote for NL ROTY. Anyways, good article, man!
You’ve modelled the relationship between air ball slugging percentage and exit velocity as linear, but we know from first principles that the relationship cannot be linear. At the very least, exit velocity can theoretically increase forever, while slugging percentage can never be greater than 4.00, so we know that as exit velocity increases the relationship will asymptote. In the more realistic range of possibilities it is very likely not to be linear either, though. It is likely that the part of the graph you have highlighted is curvilinear with slugging percentage increasing faster than exit velocity.
If the relationship is curvilinear then Gary Sanchez is much less of an outlier than your linear model makes it appear.
Sure but look at the data, that nonlinearity moves his modeled SLG from about 1.050 to 1.150 — nice change, but only covering 1/5 of the gap to his 1.550.
I hope Gary Sanchez can Make The Yankees Great Again. I would like to see them build a wall to get rid of the mets.
Could you give me a web site that randomly picks one baseball stat, scatterplots it against another, and labels the outliers? I’d click all day.
ERA gives you credit for other peoples’ defense though. If you’re looking at some luck-adjusted “what they should have done” model shouldn’t you look at xFIP for Fullmer?
So basically by your metric you chose the guy with a 4.00 ERA vs. the catcher with a .500 slugging % and a rocket arm. Never choose the ordinary over the extraordinary!
I honestly do not understand the fascination with Fielding Independent Pitching, and using it to evaluate a pitchers actual performance. I get that it is supposedly a better predictor of future performance than ERA+, but it is not a measurement of actual performance.
FIP is a prediction of performance based upon only 4 variables, and in the case of Michael Fulmer those 4 variables accounted for 30.76% of the plate appearances against, and only 27.67% of the outs recorded by Fulmer.
The prediction for FIP is based upon the “average” relationship of these 4 variables to actual ERA, for all pitchers for a multitude of seasons. There is no context as to variability in ERA to these 4 variables, and as such we have no ability to contextualize the range of possible outcomes that would be at least probable. All this tells you is that the average pitcher, who in fact does not exist, would have an ERA somewhat close,(but it is pretty vague just how close) to this calculated FIP value, based upon the number of home runs, walks, HBP, and strikeouts he recorded. In other words it tells us very little about the actual pitcher in question.
Now I tend to avoid this kind of thing because one can always find a ridiculous exception to almost anything, but indulge me. In 2016 Justin Verlander had a FIP of 3.48 and an ERA of 3.04, and Shane Greene has a FIP of 3.13 and an ERA of 5.32. Verlander averaged 1.2 more strikeouts, and 1 fewer walk, and hit HBP/9 was 1/2 the rate of Shane Greene. Verlander had a better hit/9 rate, a far better BABIP, and a lower opponent OPS. There was one FIP variable where Greene was better; HR rate, where he allowed 0.4 HR/9 and Verlander allowed 1.2 HR/9.
By nature people look at FIP and assume that it is a better measure of performance than ERA, because FIP takes out the impact of the defense behind the pitcher. Of course that is a misnomer, because the 69.24% of PA that did not end in one the 3 true outcomes + HBP, are in no way completely dependent upon the defense, but are in fact heavily influenced by the pitcher. We know from just 2 seasons of data that exit velocity and launch angle have a substantial influence on hit ball outcome, and that even if the ball is in play, the defense often has little to no influence on the outcome. Yet FIP completely ignores all of those outcomes, assuming that because the defense was in some measure involved (be it minuscule to significant or somewhere in between) it is not relevant to a pitchers performance.
Shane Greene pitched with the same defense as Verlander, with the exception of the greater number defensive replacement innings behind Greene. Verlander was better than Greene in every possible measure, except Home Runs allowed, and that single data point according to FIP, says that Greene clearly out performed Verlander.
This is utter nonsense. As I said we can all find exceptions, but if you have a measurement that has such an obvious and absurd conflict, then it’s value as a good evaluation measurement must be called into question. Please excuse my hyperbole, as I am struggling for a more accurate metaphor, but unless one can show that this kind of an outcome is something that could be described as a “Black Swan”, one must place little to no value in FIP as a worthwhile measurement of actual performance.
Fullmer was the clear choice,fangraph faithful have listed the reasons quite nicely.