The Math Behind Intentionally Walking Juan Soto

At my old job, my boss occasionally held idea sessions. He wanted everyone to participate, and the point wasn’t to come up with something actionable, just to brainstorm. No suggestion was too ridiculous – sure, it might get picked apart in discussion, but the whole point was to suggest weird stuff and see what came out of it. Still, I can safely say that none of those judgment-free-zone ideas sounded quite as zany to me as “let’s intentionally walk the guy in front of Aaron Judge.”
That didn’t stop Stephen Vogt on Tuesday night. With runners on second and third base and one out in the bottom of the second inning, Vogt didn’t let Juan Soto hit. He put up four fingers to send Soto to first. His reward? A bases-loaded encounter with Judge, the best hitter in baseball. Obviously Vogt had a reason for his decision. I ran the math to see how well that reason agrees with theory.
In a vacuum, it’s pretty clear why this intentional walk was bad: It loaded the bases with only one out, increasing the chance of a big inning, and it did so with the presumptive American League MVP at the plate. But there were two reasons to do it. First, it took the bat out of Soto’s hands, and Soto is himself a phenomenal hitter, particularly against righties. Second, it created the chance for an inning-ending double play, which would have been a huge boon to the Guardians’ chances (they already trailed by two). If you squint, you can kind of see it; maybe these two choices are equal. It didn’t matter in Game 2, because the Yankees won going away, but if the Guardians come back to win the series, they’ll be facing New York’s best hitters in important spots again, so what Vogt chose to do Tuesday night might help us guess what he’ll do in the future.
I decided to work backwards to figure out what I would have done in this situation. First, I took projections for Judge and Cade Smith, who was summoned from the bullpen for this matchup. I adjusted both of those projections based on regressed platoon splits. Smith doesn’t have a ton of major league data to work with, but he’s pitched to an observed reverse split, and I forecast him to have roughly neutral platoon matchups going forward. Judge is about 5% better against lefties than righties, roughly average for a right-handed hitter.
With those projections and a modified log5 model, I came up with a grid of modeled outcomes. That looks like this:
| Outcome | Likelihood |
|---|---|
| HR | 4.9% |
| 2B | 3.7% |
| 3B | 0.0% |
| 1B | 10.5% |
| BB+HBP | 20.4% |
| K | 34.6% |
| FO | 12.9% |
| GO | 12.9% |
From there, I calculated what each of those outcomes would do for Cleveland’s win probability. For example, a home run would make it 6-0 Yankees, and more or less end the game (5.4% Guardians win probability). A strikeout would make things much better (23.5% win probability). A groundout could either score a run or produce a double play, with roughly equal likelihood. Here’s what that looks like with all the win probability numbers filled in:
| Outcome | Likelihood | CLE Win% |
|---|---|---|
| HR | 4.9% | 5.4% |
| 2B | 3.7% | 8.0% |
| 3B | 0.0% | 6.5% |
| 1B | 10.5% | 11.6% |
| BB+HBP | 20.4% | 12.7% |
| K | 34.6% | 23.5% |
| FO | 12.9% | 20.7% |
| GO | 12.9% | 22.8% |
In aggregate, I get an 18.1% chance of Cleveland winning the game when Smith stepped in to face Judge. That’s largely because of a good chance of striking Judge out; most of the ball in play outcomes are pretty rotten for the Guardians. In reality, Judge hit a sacrifice fly, which was almost exactly the average outcome – our game odds gave Cleveland a 17.7% chance of winning after the plate appearance.
That’s a rough spot for the Guardians, obviously. But we’re not comparing it to 50% – they were already trailing and already had the dangerous part of the Yankees lineup coming up with runners on base. Things were already bad. To compare it to the alternative of pitching to Soto, I had to do some similar analysis. First, I created a matchup grid for Soto against Smith. I picked Smith instead of a lefty because I can’t imagine Vogt would want to use a worse reliever in such a big spot; Smith was the only Guardians reliever warming up, so this seems like a good bet. In any case, here’s the outcome likelihood grid for Soto against Smith:
| Outcome | Likelihood |
|---|---|
| HR | 4.1% |
| 2B | 3.9% |
| 3B | 0.0% |
| 1B | 13.2% |
| BB+HBP | 21.7% |
| K | 25.5% |
| FO | 15.8% |
| GO | 15.8% |
From there, I just did some plug-and-play math. For each potential Soto outcome, I adjusted the base/out state, then used my grid of Judge’s potential outcomes from up above to further progress the game state. For example, after a Soto strikeout, I re-ran the Judge numbers with runners on second and third and two outs. After a Soto double, I re-ran the Judge numbers with a runner on second, one out, and a four-run deficit. I did this for all of Soto’s possible outcomes so that I could figure out how likely the Guardians would be to win in each case.
Some of these were easy — an unintentional walk is the same as an intentional walk, for example. Some are tricky – a groundout doesn’t always score the runner, so I’m guessing there. Maybe Smith would pitch Soto differently based on the base being open; maybe he’d pitch Judge differently based on what happened in Soto’s at-bat. These are just generalizations, with plenty of margin for error. But still, it’s worth doing the math, so I did.
Here’s the result of all that math, the likelihood of the Guardians winning, accounting for the fact that a Smith/Judge matchup will follow Soto’s at-bat:
| Outcome | Likelihood | CLE Win% |
|---|---|---|
| HR | 4.1% | 8.3% |
| 2B | 3.9% | 11.1% |
| 3B | 0.0% | 10.1% |
| 1B | 13.2% | 12.2% |
| BB+HBP | 21.7% | 18.1% |
| K | 25.5% | 24.4% |
| FO | 15.8% | 21.3% |
| GO | 15.8% | 20.3% |
Those are pretty intuitive results: If Soto gets a hit, the Guardians are worse off than if they’d just walked him. If he makes an out, they’re better off than if they’d walked him. Thank you, I’ll be here all week. But the really interesting part is that when you sum all of those up, you get a 19.1% chance of winning the game, a full percentage point better than the projected win probability after an intentional walk.
That feels strange, because Soto’s projected outcomes are strong there. We’re talking about a .267/.425/.428 batting line, a near-.400 wOBA. Intuitively, walking someone with results that good to set up a double play feels like a smart idea. The problem is that many of Soto’s best outcomes are walks. That’s what’s so tough about him: Plenty of plate appearances that would normally end in a strikeout or weak contact become walks thanks to his elite batting eye. That makes Soto a nightmare in spots where a baserunner is valuable. But when the pitcher’s alternative is intentionally walking him, those unintentional walks simply don’t matter. If you face Soto and walk him, no big deal! That was your backup plan in the first place.
All that said, the total increase in Cleveland’s win probability isn’t outrageous. One percentage point of win probability isn’t nothing, but it’s hardly a massive effect. It’s definitely small enough that other factors could make intentionally walking him a good decision, even if the math disagrees in a vacuum. For example, the Yankees were pretty likely to win anyway. Vogt also had something going for him that I can’t quantify for this exercise: Soto didn’t face Smith, so he isn’t building up a mental catalog of his pitches. The Guardians plan on using their bullpen a lot, so preserving a little mystery there has value. Down 2-0 in the series, there isn’t much room for error; every matchup between Smith and Soto is going to be meaningful. If you think that Smith is going to face Soto multiple times with the game on the line, maybe the added value offsets what the Guardians lost by forgoing a chance to get Soto out.
As is always the case when you’re splitting something so minutely, there’s a chance my math is wrong here. A little error could go a long way given the tiny effects we’re searching for. Additionally, the composition of both the Yankees and Guardians leans in favor of taking a risk; the Guardians are underdogs, so I generally like strategies that increase variance for them. But even with that caveat, I’m pretty surprised by the outcome. My initial guess was that thanks to Soto’s proclivity to walk even when you don’t want him to, pitching to him would be a far better idea than walking him. But his power and platoon splits, along with the fact that Smith is such an elite reliever that his outcome grid against Judge isn’t abysmal, make the math work, or at least make it close. So good job, Stephen Vogt. You did something that sounds completely crazy – issuing an intentional walk to load the bases for literally Aaron Judge – and I can’t even say for certain that it was a bad decision.
Ben is a writer at FanGraphs. He can be found on Bluesky @benclemens.
Probabilities aside, I think it’s always a good idea to try it with specific players, until it doesn’t work. Depends on the pitcher matchups with the two players. Yes it puts pressure on the pitcher, but you wouldn’t do it if you were worried about his control; but if you are comfortable there and with the matchup, it puts more pressure on the batter, and keep doing it until it doesnt work. But with the variation in batters and pitchers and game situations, I don’t think you could ever make blanket statement about likelihood of success.
I apologize for being off-topic, but I was wondering when you expect the Depth Chart projections for 2025 to be released. I remember that Zips was ready mid-Nov last year, and since Depth Charts is a mixture of projection systems, I know it will at least be later than that.
Steamer is available mid-November usually, but ZiPS isn’t usually fully released until late January, so I imagine that true Depth Charts wouldn’t be until even after that
I think mostly Vogt might have felt that he was only giving one of the Yankees great hitters a chance to beat him instead of both of them. If he pitches to Soto, he has to get both Soto and Judge out (barring unlikely events like strange double plays, pick offs, etc) to avoid allowing a big inning. If he walks Soto, he has to get Judge and Austin Wells out, maybe just Judge if he gets lucky and gets a GIDP.
Also, if Soto homers you are down 5 and if Judge homers after Soto walks, you are down 6. There isn’t much difference in Cleveland’s chances being down 5 or 6, so look at what is your best chance of allowing zero or one, even if it increases the odds of giving up 4 instead of 3. It’s sort of the opposite strategy of bunting to play for one run and reducing the chance of a big inning when you are hitting. He is increasing the chance of a bigger inning, but that doesn’t hurt him as much as allowing a single run helps him.
I think it’s this
The best chance to survive the inning is facing Judge and Wells, not Judge and Soto.
And Soto’s extra run doesn’t really hurt.
The flaw in this logic is that if he gets Soto out, he can walk Judge and face Wells. So he is still choosing to face Judge over Soto and the only advantages he gets are still really the handedness splits and the DP potential.
I was wondering about this, so I’m really happy somebody did the math!
The results are closer than I would have thought. My only quibble is forecasting a neutral split for Smith. My knowledge of regressing platoon splits is rusty, but I wouldn’t expect a career TBF against LHH of fewer than 150 to have much of an effect on a platoon split projection. i.e. I’d assume you’d just get something very close to your average split. There’s also nothing in Smith’s arm angle or pitch mix to indicate he might have a true talent reverse or neutral split.
If we do project a standard platoon split for Smith, that might just swing the math in favor of Vogt’s decision being the right one.
Yeah, I’m with you on Smith’s career splits. I used the method from The Book, which is maybe a little dated but still works pretty well as a first approximation. Take 700 PA worth of a league average platoon split, add in the player’s platoon split (weighted by PA’s against lefties), and take an average. That works out to a nearly dead even split (0.6% better against righties) for Smith.
It’s very unlikely that his .284 BABIP allowed against lefties and .319 against righties will persist. On the other hand, he ran a swinging strike rate of 15.9% against lefties and 14.3% against righties – seems somewhat real. He’s basically fastball/splitter to lefties, and those are his two best pitches IMO, so I can certainly understand why it’s the case. Like I said, I don’t believe he’s a giant reverse splits guy true talent, but even isn’t a crazy assumption.
Interesting! I might need to rethink what kind of pitchers can potentially have reverse splits. I appreciate the reply, Ben. I really enjoyed the article and thinking about this.
In the moment, I thought it was the right call.
One technical element I wonder about has to do with whether the win% for GB results accounts for the idea that bases loaded gives an infield defense maximal options for getting even one out, and might increase the scenarios for double plays. On the flip side, this could just be “conventional” baseball wisdom and it’s not actually true, or the increase in options is offset by the stress/chaos tax that seems to afflict fielders in high-pressure situations — maybe more errors are made in these situations, too.
From a single game perspective, it is, in my opinion, a terrible idea. From the series perspective, it actually starts to make sense in a way that I don’t think Vogt would ever publicly admit.
Those 15-30% percent win probabilities are almost a worst-case scenario for a team so heavily reliant on a small circle in their bullpen. You’re still pretty significantly unlikely to win, but there is enough of a chance that it’s hard to justify not using the Leverage Dudes. Using the Leverage Dudes in a loss is a pretty significant negative outcome for the series overall, both due to adding usage to those arms and the familiarity effect.
They were already bringing in one of the Dudes, but if you can keep more of them in your pocket, you’re better off in G3+.
The other argument might be that Soto has been hitting well, while Judge is not. Or that you might get Judge too amped up by IBBing to get to him. Mind you, I don’t think either of these are useful/correct arguments, but I could imagine that would be justification some could use here.
Excellent article. It’s very interesting to me that it’s this close even if we ignore that Judge was ice cold. I know, it’s 5 games. But maybe there’s a small mechanical issue that is harder to identify because he’s facing so many different pitchers. Or a small injury. If there’s any reason to believe that Judge isn’t the same guy that he was in the regular season then walking Soto is the right move.