Guides · 6 min read

What WAR means in baseball

One number meant to summarise a player’s entire value — how it’s built, what the numbers mean, and where it should be trusted.

The question WAR answers

Wins above replacement tries to answer one deliberately broad question: if this player vanished and the team replaced him with a freely available minor-league call-up, how many wins would it lose?

That comparison point — replacement level — is the key idea, and it’s what distinguishes WAR from most other statistics. It isn’t comparing a player to the league average; it’s comparing him to the realistic worst case, the player any team could sign for nothing. A replacement-level team would win roughly 48 games in a 162-game season.

Everything a player does that raises the team above that baseline — hitting, baserunning, fielding, pitching, and the positional difficulty of where he plays — is converted into runs, and then runs are converted into wins.

How WAR is built

WAR is an assembly, not a formula. For a position player it sums roughly five components:

  • Batting — runs created above average, park- and league-adjusted.
  • Baserunning — steals, taking extra bases, avoiding outs on the paths.
  • Fielding — runs saved or cost relative to an average defender at that position.
  • Positional adjustment — a credit for playing a demanding position. A shortstop and a first baseman with identical bats are not equally valuable, because shortstop is far harder to fill.
  • Replacement adjustment — the constant that shifts the baseline from average down to replacement level.

Those add up to a run total, which is divided by the runs-per-win figure — around ten — to produce wins.

Pitcher WAR works the same way but starts from run prevention, with the versions differing sharply on whether to use actual runs allowed or a fielding-independent estimate.

Reading the scale

The numbers are meaningless without the benchmarks, which are roughly stable across versions.

MVP candidate MVP candidate: 8 WAR — a handful of players per season 8 WAR Superstar Superstar: 6 WAR 6 WAR All-Star All-Star: 5 WAR 5 WAR Solid regular Solid regular: 3 WAR 3 WAR Role player Role player: 1.5 WAR 1.5 WAR Replacement level Replacement level: 0.2 WAR — freely available 0.2 WAR
Approximate WAR benchmarks for a full season. These are conventions, not hard cutoffs.

Why the versions disagree

There is no single official WAR, and the two best-known versions — fWAR from FanGraphs and bWAR from Baseball-Reference — routinely differ on the same player, sometimes by two wins or more.

The main disagreement is pitching. FanGraphs builds pitcher WAR primarily from fielding-independent outcomes — strikeouts, walks, home runs — on the reasoning that a pitcher controls those and not what happens once a ball is in play. Baseball-Reference uses runs actually allowed, adjusted for the quality of the defense behind him. A pitcher with a great defense or unusual luck on balls in play can look very different under the two.

They also use different fielding metrics, and defensive measurement is the least settled area in the sport.

The practical advice: never compare an fWAR to a bWAR. Pick one version and stay inside it, and treat differences under about one win as noise in either.

Where WAR should and shouldn’t be trusted

WAR is genuinely good at what it was built for — a rough, one-number summary for comparing dissimilar players across positions, and for valuing contracts, since the market has a fairly stable price per win.

It is much weaker in a few places. Single-season defensive metrics are noisy, so a big WAR swing often reflects a fielding estimate rather than a real change in the player. Catchers are poorly served, since framing and game-calling are only partly captured. And small samples mislead — a month of WAR means almost nothing.

The honest way to use it: as a starting point that tells you roughly which conversation a player belongs in, not as a final verdict that settles it. A 0.4 gap between two players is not a difference.

Impact, game by game, in TwentySeven

WAR is a season-long accounting of value. TwentySeven does the game-scale version: open any game and you get player impact rankings built on win probability, showing who actually moved the needle that night and which five high-leverage plays decided it.

The two are complements. WAR tells you who has been valuable across a season; WPA and impact rankings tell you who won this game.

TwentySeven icon

Put this into practice with TwentySeven

TwentySeven brings player impact rankings in every game right into your practice — so you can put this into action, not just read about it.

Frequently asked questions

What is WAR in baseball? +
Wins above replacement estimates how many more wins a player is worth than a freely available replacement-level substitute. It combines batting, baserunning, fielding, and positional difficulty into runs, then converts runs into wins at roughly ten runs per win.
What is a good WAR for a season? +
Around 2 marks a solid regular, 5 is roughly All-Star level, 6 is a superstar season, and 8 or more is MVP territory and reached by only a handful of players a year. Zero means replacement level, the production a team could get for nothing.
Why do FanGraphs and Baseball-Reference WAR differ? +
Mostly over pitching. FanGraphs builds pitcher WAR from fielding-independent outcomes like strikeouts, walks, and home runs, while Baseball-Reference uses runs actually allowed adjusted for team defense. They also use different fielding metrics. The two should never be compared directly.
What are the limitations of WAR? +
Defensive components are noisy over a single season, catchers are imperfectly measured because framing and game-calling are hard to capture, and short samples are unreliable. Differences of less than about one win are generally within the margin of error and should not be treated as meaningful.

Keep reading