Skip to main content
ProScoutAIQ

scouting fundamentals

The 20-80 Scouting Scale, Explained

How baseball's grading scale actually works, what each number means, and why the benchmark charts you find online are quietly out of date.

By Steve Denney

Every scouting report in professional baseball speaks the same numeric language. A scout in the Dominican Republic and a cross-checker in Southern California can put a number on a player’s arm and mean the same thing by it. That language is the 20-80 scale, and it is older, simpler, and more precise than most explanations make it sound.

Here is the short version: 50 is major league average. Every 10 points is one standard deviation from that average. A 60 arm is one standard deviation above big league average. A 30 arm is two below. Everything else follows from those two sentences.

Why the scale runs 20 to 80 and not 0 to 100

The range is not arbitrary, and this is the part most explanations skip.

In a normal distribution, three standard deviations in each direction captures 99.7% of the population. Set average at 50, make each standard deviation worth 10 points, and the useful range of human performance lands between 20 and 80. There is no 90 because a fourth standard deviation above major league average is, functionally, a player who does not exist.

That framing is worth holding onto, because it tells you how rare the top of the scale is meant to be. An 80 tool is not “excellent.” It is the best in the sport — a handful of players in a generation for any given tool. Scouts who hand out 70s and 80s freely are not being generous; they are being imprecise, and their reports become harder for a cross-checker to trust.

The scale is credited to Branch Rickey, who was building the first modern farm system and needed a way for reports from different scouts in different states to be compared against each other rather than against each scout’s private sense of good.

What each grade means

GradeMeaningRough frequency
80Top of the scaleBest in the sport
70Plus-plusPerennial All-Star quality in that tool
60PlusClearly above average, a carrying tool
55Above averageBetter than the median regular
50Major league averageThe median everyday big leaguer
45Below averagePlayable, usually needs support elsewhere
40Well below averageA weakness that shapes the roster fit
30PoorRarely survives at the major league level
20Bottom of the scaleEffectively unusable

Two things about this table trip people up.

The baseline is the major leagues, always. A 50 run tool means average for a big leaguer, not average for the high school field you are sitting at. A high school shortstop who is the fastest player in his county is very often a 45 runner on this scale. That feels harsh until you remember the scale exists to answer one question: what is this player relative to the population he is trying to join?

Most everyday major leaguers are 45s and 50s. A player with four 50s and a 55 is a solid regular, not a disappointment. If your reports are full of 60s, your scale has drifted, and the drift is invisible to you and obvious to whoever reads you next.

Some organizations use the full ladder in fives — 45, 50, 55 — and some restrict to tens. Neither is wrong. Being inconsistent within one organization is.

Present and future: the two-number grade

You will constantly see grades written as a pair: 40/55. That is not a range. It is present over future.

The first number is what the player can do today. The second is what the scout believes he will do at maturity, given his frame, his age, his swing, his delivery, and how quickly he has improved. A 17-year-old with a 40 present power grade and a 55 future power grade is a scout saying: he cannot drive the ball yet, and he will.

The gap between those two numbers is where scouting actually lives. Anyone can time a runner. The judgment — the thing that gets a scout promoted or fired — is the second number. It is also the number that takes years to be proven right or wrong about, which is why good organizations keep old reports and go back to read them.

For a 24-year-old in Double-A, present and future are usually close together. For a 17-year-old, they can be twenty points apart, and the report is mostly an argument for the gap.

The tools that get graded

Position players get five tools: hit, power, run, arm, field. Some organizations split hit into contact and approach, or field into range and hands. The five are the common core.

Pitchers get graded by pitch, plus command. Fastball, breaking ball or balls, changeup, each on the same 20-80 scale, plus a command grade and often a separate control grade. Command and control are not synonyms, and conflating them is one of the more common amateur mistakes: control is throwing strikes, command is throwing the ball where you intended within the strike zone.

Both get a makeup assessment, which is sometimes graded and sometimes written in prose, and which is the single most argued-about element of any report.

What the numbers attach to

Grades for the measurable tools anchor to observable benchmarks. This is where the scale stops being abstract.

Run times. For position players, the standard measurement is home to first. Right-handed and left-handed hitters get different scales, because a left-handed hitter starts roughly a step closer.

GradeRHHLHH
804.003.90
704.104.00
604.204.10
504.304.20
404.404.30

Amateur scouts also work in 60-yard dash times from showcases, which measure something related but not identical — a 60 time is straight-line speed with a running start on a track; home to first is baseball speed out of the box, including how well a hitter gets out of the batter’s box after a swing. Players who grade differently on the two are common and the difference is usually informative.

Pop time for catchers is the time from the ball hitting the mitt to it arriving at second base. The major league average on steal attempts of second is right around 2.00 seconds. Most big league catchers cluster between 1.90 and 2.00, and getting below 1.90 puts a catcher in genuinely rare company.

Notice how tight that band is. The entire major league population of catchers lives inside about a tenth of a second, which means a pop-time grade ladder has very little room in it and organizations legitimately disagree about where the lines fall. Do not trust a chart that claims false precision here. Record the actual time, note the conditions, and let the number speak.

The chart that is always out of date

Fastball velocity is the tool where published grade charts age fastest, and most of the ones circulating online are describing a game that no longer exists.

Not long ago an 88-91 mph fastball was considered major league average. In 2026, the average major league four-seam fastball sat at 94.7 mph — up from 94.5 the year before, and rising for the sixth consecutive season. Right-handed pitchers averaged 95.2.

Work through what that does to the scale. If 50 is major league average by definition, then a 91 mph fastball is no longer a 50. It is well below. Any chart that still says otherwise is not slightly stale; it is telling you a below-average pitch is average, which is the kind of error that puts a pitcher on a follow list he does not belong on.

Two practical consequences:

Velocity grades need re-anchoring, regularly. If your organization’s velocity chart has not moved in five years, it is quietly mis-grading every arm you see.

Velocity is not the fastball grade. A fastball earns its grade on the average velocity across an outing rather than the peak reading, and then that grade gets adjusted for life, angle, extension, and deception. A 93 with elite carry and a flat approach angle plays above a 96 that comes in straight and true. Radar guns measure one input to a grade, not the grade.

OFP: the number on top

Individual tool grades roll up into an Overall Future Potential grade, also on the 20-80 scale, describing the player as a whole rather than any single skill.

OFP is not an average of the tools. A player with one 70 and four 40s is usually more valuable than a player with five 50s, because a carrying tool creates a role and five average tools can create a bench player. Organizations weight the roll-up differently, and how a club computes OFP is a genuine piece of internal philosophy — it encodes what that front office believes wins games.

Rough translation, and it varies by club:

  • 70+ — franchise player, top of a draft
  • 60 — first-division regular, an everyday player on a good team
  • 50 — major league regular or a solid contributor
  • 45 — big league role player, second-division regular
  • 40 — up-and-down depth, organizational value

Where the scale gets misused

Grading against the wrong population. Grading a high school junior against other high school juniors produces a report full of 60s that means nothing to anyone above you. The population is always major league.

Letting the scale drift upward. Grades inflate over a season the way any subjective scale does, particularly after a stretch of watching weak competition. The correction is calibration against known quantities — periodically grading players whose major league outcomes are established, and checking whether your number matches what they turned out to be.

Treating a grade as a measurement. A run time is a measurement. A hit tool grade is a judgment with a number attached. Both belong in a report, and confusing which is which makes a report harder to argue with, not more convincing. The best reports keep the observation and the conclusion visibly separate, so a reader can accept your data and still disagree with your grade.

Grading without context. A 55 arm against varsity high school competition and a 55 arm in the Cape Cod League are different claims about a player, and a grade written without the level attached is a number nobody downstream can use.

The scale is a language, not an answer

The 20-80 scale does not evaluate anybody. It gives scouts a shared vocabulary precise enough that disagreement becomes productive — two scouts who both saw a 55 and a 45 on the same arm have something specific to argue about, and the argument is where the evaluation actually happens.

That is worth remembering in an era of a lot of automated measurement. A ball-tracking system will tell you the exit velocity to a decimal place. It will not tell you whether the swing that produced it will work against a breaking ball in two years. The number belongs in the report. So does the scout who watched it happen.

If you are building a system that stores these grades, the discipline that matters is the same one that makes the scale work: never let a number appear that a scout did not put there. A grade nobody assigned is worse than no grade, because it reads exactly like one that somebody did.


Next: What MLB Area Scouts Actually Do — the job behind the reports, from territory coverage to the draft room.

Get new posts in your inbox.

One email when a post goes live. Same list as the product waitlist.