What (New) Statcast Data Tell Us About Pitcher BABIP
For the past few days, I’d been searching for a baseball topic to write about. It usually takes less time, but we’re in that calm (if not monotonous) period between the All-Star break and the trade deadline. Ideas are scarcer. Maybe I’d settle on an article with a simple premise?
So I committed myself to tackling pitcher BABIP. (Good going, Justin!)
The notion that pitchers have no control over what happens to a ball in play ushered in a golden age of baseball research, and findings from back then still influence how we view the game today. But over time, we realized that exceptions do exist; for example, Clayton Kershaw consistently allows a below-average BABIP, most likely because he’s a phenomenal pitcher. In addition, certain pitchers have a knack for inducing weak contact in the form of pop-ups or grounders. Exactly how those batted balls impacted BABIP remained a mystery, but you could no longer brush off the metric as total noise.
Years later, Statcast data became available for public use. Even so, research on pitcher BABIP remained far and few between; it’s a daunting subject! I did use two articles as inspiration, however. The first is from FanGraphs user rplunkett97 on our community research page. Dating back to 2017, it mainly discusses a linear model with several variables (BB/9, GB%, Team UZR, and more) used to produce an expected BABIP for each pitcher. The second is courtesy of Alex Chamberlain, also from the same year, who used a mixture of Hard-hit and Barrel rate to create his own version of xBABIP.
We’re still in the Statcast era, but there have been a couple of changes since then that could help us better evaluate pitcher BABIP. Statcast is now provided by Hawk-Eye, which does a better job at recording batted balls that in previous seasons might have flown under the radar or been mislabeled. As proof of this being the case, below is the yearly rate of batted balls that are either above 40 degrees or below -20:

Hitters and pitchers aren’t producing or allowing more of these balls; Hawk-Eye is just more adept at measuring extreme angles.
Those launch angle ranges were part of my article about Cristian Javier and Framber Valdez as examples of pitchers who have control over how much extreme contact they allow. Moreover, such contact pertains to BABIP in important ways. This isn’t a new concept, but it’s helpful to remind ourselves that not all groundballs are created equal:
| LA Range | No. of BBE | BABIP |
|---|---|---|
| -5 to 10 | 14,889 | .421 |
| -20 to -5 | 9,968 | .160 |
| Below -20 | 8,326 | .116 |
The lower the launch angle, the lower the BABIP allowed. Some of the super-negative angles do produce hits depending on how they were struck and the batter’s speed (see: the Rays), but I imagine they’re offset by the fact that low angles limit exit velocity. Also, what’s surprising is how drastically BABIP rises once a batted ball is even slightly elevated. It’s possible a fair amount of noise within BABIP is due to pitchers’ inability to control the rate of bad grounders they allow.
As for batted balls over 40 degrees, they rarely become hits; when they do, they’re mostly home runs that don’t count as balls in play. And since pitchers can choose to allow them (or not), including high fly balls and pop-ups into a BABIP model is a no-brainer.
But when Alex Chamberlain examined pop-up rate (PU%) in his article, he found that it “correlates very weakly with BABIP, producing an R^2 of 0.10.” My semi-educated guess is that the pre-Statcast definition of a pop-up was much narrower and thus made little difference between pitchers on a rate basis. Even MLB.com defines a popup as a ball hit 50 degrees or higher, which, if we’re basing it off the relationship between BABIP and launch angle, is also a bit stingy. Though the exact number could be debated, a lower limit of 40 degrees seems ideal.
In the same spirit, I used exact launch angles to adjust the definition of a line drive. According to MLB.com, line drives are balls within the 10–25 degree range, but I found that shifting each end back by five degrees best captured the relationship between BABIP and launch angle. Plus, I’ve always found it weird that a nine-degree rocket would count as a groundball, so this change is a fantastic reconciliation.
Lastly, we now have access to Outs Above Average (OAA) data by pitcher, meaning we can see who was harmed or aided by his teammates. I know the point of BABIP is that the defense behind him is not something we should hold a pitcher accountable for, but I wanted to see how much of the variance in BABIP select variables can explain. Speaking of variables, here’s a list of the ones I chose below:
- Percent of BBE below -20 degrees (GB%)
- Percent of BBE between 5 and 20 degrees (LD%)
- Percent of BBE above 40 degrees (PU%)
- Outs Above Average when pitching
I then fed these four variables into a linear model, which did the hard work of measuring how they correlated to each pitcher’s BABIP in 2021 as of July 19. Here are the long-awaited results:

The sample consists of 138 of 148 pitchers this season with at least 50 innings pitched; the remaining 10 weren’t included in a database I use to translate MLB IDs into FanGraphs ones. I thought of calculating their xBABIPs manually but figured they wouldn’t make much of a difference on the results.
Regardless, it’s cool to see that four variables alone can account for nearly half the variance in BABIP among pitchers with at most 110–120 innings under their belts. Even if you toss out the defensive component, you still end up with a robust adjusted R^2 of 0.329. Individually, Statcast-adjusted line drive rate has the strongest correlation with BABIP (R^2 = 0.211), and groundball rate has the weakest (R^2 = 0.002, basically nonexistent). Within the context of a model, however, I suspect it works in tandem with defensive data, and removing it weakens both the R^2 value and the RMSE.
What about predicting future BABIP? This is a question worth answering. How, for example, does Kershaw author season after season of low BABIP? Can our model provide an answer? My instincts said no, given that two of the variables (LD%, OAA) are highly volatile year-to-year, and a third (GB%) correlates weakly to same-season BABIP.
That leaves us with PU%. To start, I compiled pitchers with both (a) at least 30 innings pitched in 2020 and (b) at least 50 innings pitched in 2021; that returned a sample of 104. Then I looked at how a pitcher’s PU% in 2020 correlated to his BABIP in 2021 and how that compared to his BABIP the previous year. The results:
- 2020 PU% to 2021 BABIP: R^2 = 0.051
- 2020 BABIP to 2021 BABIP: R^2 = 0.034
Womp, womp.
Yep, the rate that’s not only sticky year-to-year but also quite indicative of same-season BABIP is barely better at estimating future BABIP. There’s the caveat that we’re dealing with micro samples here, but even with more innings, it’s doubtful PU% will reach a point of moderate correlation. Unfortunately, this is a dead end.
Still, the descriptive power of xBABIP provides us with a few insights. Take John Means as an example. He has a microscopic .201 BABIP since last season, and while he did possess an abnormally low line drive rate in 2020, it’s rebounded in 2021, but without a corresponding change to BABIP. His secret? Means has gained a few ticks of ride and velocity on his fastball, allowing him to induce pop-ups (and whiffs) up in the zone. Although there’s virtually no chance that his current BABIP of .192 lasts for an entire season, the regression-friendly xBABIP nevertheless puts it at .241. Barring drastic decline, Means looks like he could outpitch his peripherals for years to come.
I ideally would have multiple full seasons to work with, but I don’t feel comfortable using years prior to 2020 since there’s been a change in tracking systems. Predicting BABIP is still an extremely tall task, but more accurate and granular Statcast data allow us to break down the metric’s components within a season. As more batted balls are compiled under the gaze of Hawk-Eye, it’ll be interesting to see if the details of xBABIP change. For now, this will do. (P.S. BABIP data for the aforementioned 138 pitchers can be found here.)
All statistics are current through games of July 19.
Justin is an undergraduate student at Washington University in St. Louis studying statistics and writing.
This progressive narrative of science and data unraveling its own flawed ideas is crazy. So, people who know nothing state that BABIP is all luck. Then, those people take credit for unraveling the mystery that they were wrong. Thank you Statcast for this valuable gift! I like that the cooperate sponsor needs to get credit for an entire revolution of junk. I am sure that the new technology is super reliable although they will need a sponsor. All this kind of analysis leads to is digging real deep in arbitrary metrics in search of generalizations that don’t hold. This is such a weird tone but typical of Fangraphs.
Imagine what might happen if analyst abandon linear models.
Always thus.
Old people: things fall down
Newton: actually, things are attracted to each other by a force based on their relative mass
Einstein: actually, gravitational effects are the result of spacetime curvature, not forces
The latter is a more detailed understanding that explains more phenomena, but Einstein doesn’t make Newtonian physics irrelevant within non-relativistic motion anymore than ‘things fall down’ is inaccurate when the frame of reference is on Earth.
Intuitively I wonder if horizontal angle should be given more attention in expected BABIP models. These kinds of exercises always seem to focus on exit velocity and launch (vertical) angle, while horizontal (spray) angle is completely ignored.
Is a .46 r squared good?
it’s…moderately good?
r^2 is always between 0 and 1, with 0 being the independent variable explains absolutely none of the variation in the dependent variable, and 1 being that the IV explains all of the variation in the DV. So, 0.46 is a moderate correlation.
I have been reading Fangraphs articles for many years now and I always see the same types of errors in every “analysis”. Too much “group think” happening with analysis focusing on raw data of velocity and angles. My grandfather was a top pitching and switch hitting prospect for the Yankees in the late 1930s, early 1940s prior to an arm injury ended his career. After a stint in the army during the war, he returned home, had a family and then spent more than a decade as an American Legion coach so he could stay connected to the game and coaches several dozen players that went on to play major college and minor league ball. His most successful player at the MLB level was Jim Kaat. My grandfather taught me to play the game and I was drafted by the Cardinals as a senior in high school after being named one of the top 100 prep players in the country. I chose military and law enforcement over baseball but I still follow the game. So I am going to share some of the things my grandfather taught me in regards to pitching as well as my own first hand experience and observations as a baseball obsessed fan for 40+ years now.
First, velocity in the least important of the three factors in pitching. Location is the most important and break/movement is second. A pitch thrown at the bottom of the zone that is breaking down will generate far more groundballs, which are converted into more outs. Historically, the BABIP on groundballs is around .210 (though it might be slightly different now since it’s been 4 years since I ran the numbers).
A pitch at the very top of the zone, the “shadow” zone generates more pop-ups. The historical BABIP on those pitches was under .200 (.189 actually) when I ran the numbers in 2017.
Pitches on the inside of the zone, on the “black” breaking inward result in weak contact and usually go opposite field. Pitches on the outside “black” breaking away usually result in slightly stronger contact but also usually go opposite field.
A pitch in the middle of the zone is most likely going to get crushed regardless of the itch velocity or movement. Pitchers do NOT want the ball in the middle of the zone. It’s just very bad for them.
Every “analysis” I see here makes the same type of mistakes — too much focus on raw data like exit velocity, angles and BABIP results. No one ever looks at pitch location data and the runs the numbers on pitch type, movement and handedness of pitchers and hitters.
Pitchers that have low BABIPS are good at generating weak contact. They do that by staying out of the middle of the zone, consistently work the “shadow” zone, mix their pitches well to set hitters up for their next pitch, and have a mix optimized for inducing weak contact and groundballs.
Every “analysis” that fails to consider pitch location and movement omits the two biggest factors pitchers have in determining their success.
The problem is this type of data is hard to get ahold of. Location, command (how often a pitcher can hit a spot (location) that he was aiming at), spin mirroring, and sequencing are all very important factors in pitch value results. This is the new frontier that is ripe for Fangraphs analysts to tackle. Sometimes we do see articles on this topic but I agree we could use more!
Launch angle and EV do sort of act as proxies for those things. A pitcher who keeps the ball down is going to get more high-negative-angle balls, for example. But I agree with you; how pitch location and movement relate to BABIP is a wide-open area right now. And, seeing at how dismally the Braves have done with their velocity-first approach to drafting pitchers, I agree with you on the importance of the three factors.
If it’s so easy to “run the numbers” and find out how to predict BABIP, why don’t you do it yourself? CommunityGraphs is open for business.
Out of curiosity, who is the data point at the very top of the graph?
Pretty sure it’s Matt Harvey 🙁