Score Guide

Score types: standard, scaled, T, z and the rest

Assessment reports mix half a dozen score metrics, each with its own mean and scale. They all describe the same thing — where a score falls relative to a norm group — just in different units. Once you can convert between them, the whole report reads as one language.

One idea, many units

Almost every standardized score is a standard score in the general sense: it expresses how far a raw score sits from the average of a norm group, measured in standard deviations. The different metrics — standard scores, scaled scores, T-scores, z-scores, stanines, NCEs — are just different choices of where to put the mean and how wide to make each step. Learn the mean and standard deviation of each, and you can move between them freely.

The key that unlocks everything Every metric below is anchored to the same normal distribution. A score that's exactly average sits at the mean of every scale at once — Standard Score 100, Scaled 10, T 50, z 0, the 50th percentile. A score one standard deviation below average is SS 85, Scaled 7, T 40, z −1, the 16th percentile. Same location, different labels.

Conversion table

These rows all describe the same points on the normal curve, expressed in each metric. The percentile is the common translator — it's what every scale ultimately points to.

DescriptionSD from meanStandard (100/15)Scaled (10/3)T (50/10)z (0/1)Percentile
Well above average+21301670+2.098th
Above average+11151360+1.084th
High end of average+⅔1101257+0.775th
Average010010500.050th
Low end of average−⅔90843−0.725th
Below average−185740−1.016th
Well below average−270430−2.02nd

Percentiles are rounded to the values most commonly cited; different sources may round slightly differently.

Standard scores — mean 100, SD 15

The most common metric on cognitive and academic reports. Composite and index scores are almost always standard scores: a WISC-V Full Scale IQ, a WJ-V cluster, a BASC-3 composite. Because the mean is 100 and each standard deviation is 15, roughly 68% of people score between 85 and 115 — that is, within one standard deviation of the mean. That part is fixed by the math.

What isn't fixed is the label. Where "Average" begins and ends is a publisher's choice, not a statistical fact. Some descriptive systems call the whole 85–115 band "Average"; others reserve "Average" for a narrower 90–110 and split off "Low Average" and "High Average" at the edges. So the same standard score of 87 might be labeled "Average" by one test and "Low Average" by another — the number is comparable across tests, but the word attached to it is not. Keep the general rule of thumb in mind (around 85–115 is the broad average range), but always check the specific test's manual for how it defines the bands.

Number vs. label A standard score is comparable across tests; the descriptive term next to it may not be. When you translate scores into one common currency for a report or a parent conversation, standard scores are usually the clearest choice — just remember the range label comes from the publisher, not the score itself. (More on this in Descriptive Ranges →.)

Scaled scores — mean 10, SD 3

Used for subtest scores on the Wechsler batteries and many language and processing measures. The mean is 10 and each standard deviation is 3, so the average range runs from 7 to 13. A scaled score of 10 is exactly average; a 7 is one SD below, a 13 one SD above.

A common point of confusion: subtest scaled scores and composite standard scores appear side by side on the same report but live on completely different scales. A scaled score of 10 and a standard score of 100 mean the same thing — dead average — even though the numbers look nothing alike. Reading them as if they were on one scale (thinking a "10" is alarmingly low next to a "100") is a frequent mistake.

T-scores — mean 50, SD 10

The standard metric for behavior, social-emotional, and personality measures — BASC-3, Conners, the MMPI, and most rating scales. Mean 50, standard deviation 10, so the average range is 40 to 60.

One thing to watch with T-scores: on many clinical scales, higher is a greater concern, not a better result. A T-score of 70 on an anxiety or hyperactivity scale means the concern is well above typical — the opposite direction from a cognitive score, where higher is stronger. Always check what the scale is measuring before reading a high T-score as good or bad. (More on behavior & rating-scale scores →)

z-scores, stanines, and NCEs

z-scores — mean 0, SD 1

The z-score is the rawest expression of the same idea: it states, directly, how many standard deviations a score is from the mean. A z of 0 is average; +1 is one SD above; −1.5 is one and a half below. Every other metric is really a z-score dressed up with a friendlier mean and scale. z-scores are used mostly in research and whenever you need to compare across measures on different metrics, because they strip everything down to the common unit.

Stanines — mean 5, SD 2

A nine-point scale (the name comes from "standard nine"), with a mean of 5 and standard deviation of 2. Stanines compress the whole distribution into nine broad bands — 1 is lowest, 9 is highest, 5 is average. They're an older format still seen in some group-administered educational testing. The trade-off is deliberate: stanines are simple, but coarse. A single stanine covers a wide span of percentiles, so two students in the same stanine can differ meaningfully.

NCEs — mean 50, SD 21.06

Normal Curve Equivalents look almost like percentiles — they run 1 to 99 and meet percentiles at 1, 50, and 99 — but they're built on an equal-interval scale (mean 50, SD 21.06). That equal spacing is the whole point: unlike percentiles, NCEs can be averaged and compared arithmetically, which is why they're used in federal education program reporting. They are easy to mistake for percentiles at a glance; the values in between diverge, so it's worth confirming which one a report is actually using.

Percentiles — the common translator

A percentile rank states the percentage of the norm group a score met or exceeded: a score at the 37th percentile did as well as or better than 37% of same-age peers. Percentiles are the most intuitive metric for a lay audience, and they're the natural bridge between all the others — every scale above ultimately points to a percentile.

The one trap: percentiles aren't evenly spaced Percentiles bunch up in the middle and stretch out at the extremes. Near the average, a few points of standard score move the percentile a lot — SS 100 is the 50th percentile, SS 110 the 75th. Out at the edges, large standard-score differences barely move the percentile — SS 70 is the 2nd percentile and SS 55 is roughly the 0.1st, a 15-point gap that looks tiny in percentile terms. So the distance between two percentiles is not a fixed amount of ability; the same 10-percentile gap means something very different in the middle than at the tails.

This is also why percentiles and standard scores are worth reporting together: the standard score preserves the equal-interval spacing, and the percentile makes the standing intuitive. Each covers the other's weakness.

Putting it together

Once the metrics click into place, a line like "Scaled 7, SS 85, T 40, 16th percentile" stops being four numbers and becomes one fact stated four ways: one standard deviation below average. That fluency is what lets you read a mixed report smoothly, compare scores that arrive in different metrics, and translate any of them into the plain language a parent or teacher will actually follow.

Two companion guides build directly on this one: descriptive ranges (the "Average / Low Average" labels attached to these bands) and confidence intervals (the precision around any single score).

Plot any metric on one scale

Mix standard scores, scaled scores, T-scores, and percentiles on a single normal curve — the plotter converts and places them for you. Free, nothing stored.

Open the Normal Curve Plotter More score guides