IQ Percentile and Rarity Calculator

A modern IQ score is a standard score, not a ratio of ages. It says how far above or below the population mean you sit, measured in standard deviations, and the normal curve turns that distance into a percentile and a rarity. Enter a score and this calculator returns the percentile rank, the underlying z-score, how many people out of every N score at or above it, and the equivalent score on the other common deviation scale. Switch the direction and it works backwards, turning a percentile into the score that produces it — the calculation a report writer performs when a test manual gives only one of the two.

Calculator

This calculator runs in your browser. Enable JavaScript for live results — the inputs, formula and worked example below remain fully readable without it.

Inputs this calculator takes, with typical values
InputWhat to enterExample
Start fromChoose which number you already have; the other one is computed from it.An IQ score
IQ scoreThe full-scale or composite standard score printed on the report, not a subtest scaled score.115
Percentile rankThe percentage of the population scoring at or below the score you are looking for.90 %
Standard deviation of the scalePrinted in the test manual; almost every current clinical battery uses 15.15 — Wechsler, Stanford-Binet 5, WJ, KABC
Mean of the scaleStandardised IQ scales are centred on 100 by construction; change this only for an unusual instrument.100

It returns

  • Percentile rank — The share of the population scoring at or below this score.
  • IQ score
  • z-score (standard deviations from the mean)
  • Rarity at or above: 1 in N people
  • Rarity at or below: 1 in N people
  • Same rank on the other deviation scale — The score with the identical percentile when the scale's standard deviation is 15 instead of 16, or 16 instead of 15.

The formula

z=IQμσ,percentile=Φ(z)×100
IQ=μ+σΦ1(percentile100)

In plain text: z = (IQ − μ) / σ, percentile = Φ(z) × 100, rarity above = 1 / (1 − Φ(z))

  • IQThe reported standard score (points)
  • μMean of the scale, 100 on every standardised battery (points)
  • σStandard deviation of the scale, 15 on current clinical tests (points)
  • zDistance from the mean in standard deviations (SD)
  • ΦCumulative distribution function of the standard normal curve (—)

The deviation IQ is defined so that scores are normally distributed in the standardisation sample. All three quantities are therefore properties of the normal curve, not of the test items.

Updated Category Psychometrics & Standard Score Conversion Verified against published test cases Reading time 12 min

An IQ score is a position, not a quantity

The number on an intelligence test report is a deviation IQ. It is manufactured, not measured: the test publisher administers the battery to a large standardisation sample chosen to match the population on age, sex, region, ethnicity and parental education, then transforms the raw item totals so that the resulting scores have a mean of exactly 100 and a standard deviation of exactly 15. The number tells you where a person sits relative to that sample. It carries no units and it is not a count of anything.

This matters because it makes the percentile the more fundamental figure. Saying "115" is shorthand for "one standard deviation above the standardisation sample mean", which is shorthand for "above about 84% of the sample". The percentile survives a change of scale; the raw numeral does not. A score of 132 on the old Stanford-Binet Form L-M, which used a standard deviation of 16, describes the identical position as a 130 on a Wechsler scale. Anyone quoting an IQ without naming the scale has left out half the information.

The original ratio IQ, from Stern and the early Binet-Simon work, was genuinely a ratio: mental age divided by chronological age, times 100. Wechsler abandoned it in the 1930s because mental age stops growing in adulthood while chronological age does not, so the ratio drifts downward for every adult who ages. Every current battery — the Wechsler scales, Stanford-Binet 5, Woodcock-Johnson, KABC — uses the deviation form instead, with age-specific norms so that a 40-year-old is compared with 40-year-olds.

From score to percentile to rarity

Three steps take you from the reported score to everything else. First, standardise: subtract the mean and divide by the standard deviation, z = (IQ − 100) ÷ 15. The z-score is the distance from the mean expressed in standard deviations, and it is the only quantity that is comparable across tests.

Second, integrate the normal curve. The percentile is Φ(z) × 100, where Φ is the cumulative normal distribution — the area under the bell curve to the left of z. There is no closed-form expression for Φ; every table and every calculator, including this one, evaluates it numerically. The familiar landmarks are worth memorising: z = 1 gives 84.1%, z = 2 gives 97.7%, z = 3 gives 99.87%.

Third, invert the tail to get rarity. The proportion at or above the score is 1 − Φ(z), so the number of people you would expect to sample before finding one scoring that high is the reciprocal, 1 ÷ (1 − Φ(z)). This is the figure that makes the tails intuitive. A score of 115 is roughly 1 in 6. A score of 130 is roughly 1 in 44. A score of 145 is roughly 1 in 741. Each additional 15 points is not a fixed increment in rarity — it multiplies rarity by a factor that itself keeps growing.

Running the chain backwards uses the inverse normal function. Given a percentile, z = Φ⁻¹(p), and then IQ = 100 + 15z. Report writers need this when a manual prints percentile ranks but the referral question is phrased in standard scores, and gifted-programme coordinators need it when a cutoff is written as "the top 2%" rather than as a score.

Worked example: a full-scale score of 124 on a Wechsler scale

A report gives a Full Scale IQ of 124 with a 95% confidence interval of 119 to 129. The scale has a mean of 100 and a standard deviation of 15.

  1. Standardise. 124 − 100 = 24 points above the mean. Divide by the standard deviation: 24 ÷ 15 = z = 1.60.
  2. Find the area to the left. The standard normal table gives Φ(1.60) = 0.9452. The percentile rank is 0.9452 × 100 = 94.5. The score is at or above about 94.5% of the standardisation sample.
  3. Invert the upper tail. The proportion above is 1 − 0.9452 = 0.0548, so the rarity is 1 ÷ 0.0548 = 1 in 18.2 people.
  4. Translate to the other scale. The same z on a standard deviation of 16 gives 100 + 16 × 1.60 = 125.6. On an old Form L-M report the identical person would have been described with a slightly larger numeral.
  5. Carry the confidence interval through. The lower bound of 119 is z = 1.267, or the 89.7th percentile. The upper bound of 129 is z = 1.933, or the 97.3rd percentile. So the honest statement is "somewhere between roughly the 90th and the 97th percentile", not "the 94.5th percentile".

That last step is the one most often skipped, and it changes how the result should be described. A confidence interval eight percentile points wide is normal for a well-constructed full-scale score; the point estimate on its own overstates how precisely anyone has been located on the curve.

How to read the percentile and the rarity

Start with the confidence interval, not the point estimate. Every reputable report prints one, because a full-scale score carries a standard error of measurement of two to four points on most batteries. Two scores that differ by five points describe the same performance. If your report does not show an interval, treat the score as accurate to about plus or minus five points and convert both ends.

Read the percentile rather than the numeral when you are explaining the result to anyone. "Above about 95 of every 100 children of the same age" communicates the finding; "124" invites the false impression of a measured quantity on a ratio scale, as if 124 were twice as much of something as 62. It is not. There is no zero point on an IQ scale and no meaningful ratio between two scores.

Rarity is the right frame for selection decisions and the wrong frame for describing a person. Gifted programmes commonly set entry at the 95th, 97th or 98th percentile, which is a score of 125, 128 or 131 on a 15-point scale, and the rarity output tells an administrator how many children per hundred that cutoff will identify. But rarity in the tails is exactly where the normal model is least trustworthy: standardisation samples rarely contain enough people beyond z = 3 to verify that the curve keeps its shape, so a claim of "1 in 30,000" is a property of the mathematical model, not an observed frequency.

Composite scores also hide dispersion. A full-scale score of 100 built from a verbal index of 130 and a processing-speed index of 70 has almost nothing in common with a flat profile of 100 across the board, and most test manuals warn against interpreting the composite at all when index scores are that far apart. Convert each index separately here and compare the percentiles; the spread is usually the clinically interesting part. If you want to see the same standard-score machinery working on an admissions test, the SAT scaled score calculator shows how a fixed reporting scale is built from raw counts.

Score, percentile and rarity on both deviation scales

Every row describes the same position on the normal curve expressed three ways. Percentiles come from the standard normal distribution; rarity is the reciprocal of the upper tail.
zSD 15 scoreSD 16 scorePercentile1 in N at or aboveClassification (WAIS-IV)
+3.0014514899.87741Very superior
+2.67140142.799.62261Very superior
+2.0013013297.7244Very superior
+1.33120121.390.8811.0Superior
+1.0011511684.136.3High average
+0.67110110.774.754.0High average
0.0010010050.002.0Average
−0.679089.325.251.34Average
−1.00858415.871.19Low average
−1.338078.79.121.10Low average
−2.0070682.281.02Borderline
−3.0055520.131.001Extremely low

WAIS-5 and WISC-V renamed the bands: very superior became extremely high, superior became very high, and borderline became very low. The cut points are unchanged.

Errors that distort an IQ percentile

  • Ignoring the scale's standard deviation. A 132 on a 16-point scale and a 130 on a 15-point scale are the same rank. Comparing the numerals directly, which happens constantly with historic Form L-M scores, overstates the older result by about one point per ten points above the mean.
  • Quoting the point estimate without the confidence interval. Full-scale scores carry a standard error of two to four points, so a 95% interval spans about ten points. Convert both ends and describe the range.
  • Interpreting a composite over a scattered profile. When index scores differ by more than about 1.5 standard deviations, most manuals advise against interpreting the full-scale score at all. The percentile of a composite that averages a 130 and a 70 describes nobody.
  • Treating rarity in the far tails as an observed frequency. Beyond about z = 3 the normal model is extrapolating past the data in the standardisation sample. Ratios like one in a million are model output, not counts.
  • Comparing scores from different norm dates. Norms drift — the Flynn effect — so a test standardised decades ago yields higher scores than a freshly normed one for the same performance. Publishers restandardise for exactly this reason.
  • Using an online screening score as if it were a battery score. Unstandardised internet tests have no norm sample, no reliability estimate and no age norms, so there is no defensible z-score to convert.
  • Reading a percentile as a percentage correct. A percentile rank of 84 does not mean 84% of the items were answered correctly. It means 84% of the norm sample scored no higher.

Where deviation scores show up elsewhere

The same transformation runs through the whole of educational measurement. Subtest scaled scores on the Wechsler batteries use a mean of 10 and a standard deviation of 3, so a scaled score of 13 is z = 1 and sits at the 84th percentile, identical in position to a composite of 115. T-scores, common on behaviour rating scales, use a mean of 50 and a standard deviation of 10. Normal curve equivalents use 50 and 21.06, chosen so the scale is equal-interval and lines up with percentiles at 1, 50 and 99. Every one of these is the same z-score wearing different clothing.

Admissions testing works the same way with different constants. SAT sections are reported on a 200-800 scale, and converting a raw count to that scale is the job of the SAT score calculator; moving between the two admissions tests uses a rank-matching table rather than a formula, which the SAT to ACT conversion calculator applies. Curriculum-based measures used in schools are the exception worth knowing about: oral reading fluency norms, computed by the reading speed calculator, are reported directly as percentiles of words correct per minute because the underlying distribution is not normal enough to justify a standard score.

One caution to carry away. Everything on this page is arithmetic on a normal curve, and it is exact. What it cannot tell you is whether the score itself is a fair estimate of the person's ability, which depends on the test's reliability, its norm sample, the testing conditions, the examinee's language and schooling, and the clinician's judgement. Interpretation of a cognitive assessment belongs to a qualified professional. Use these numbers to understand a report, not to replace one.

Key terms

Deviation IQ
A standard score set so that the standardisation sample has a mean of 100 and a fixed standard deviation, usually 15. It expresses relative position, not a measured amount.
Percentile rank
The percentage of the norm sample scoring at or below a given score. Percentile ranks are ordinal: the gap between the 50th and 55th is far smaller in score points than the gap between the 94th and 99th.
Standard error of measurement
The typical difference between an observed score and the true score it estimates, used to build the confidence interval printed on a report.
Flynn effect
The long-run rise in raw test performance across generations, which forces publishers to restandardise. It makes scores from tests normed decades apart non-comparable.
Normal curve equivalent
A 1-to-99 scale with a mean of 50 and a standard deviation of 21.06, built so that equal differences represent equal amounts, unlike percentiles.

Frequently asked questions

What percentile is an IQ of 130?

The 97.7th, on a scale with a mean of 100 and a standard deviation of 15. That is z = 2 exactly, and the normal curve puts 97.725% of the population at or below it. In rarity terms it is about one person in 44. On a 16-point scale the same position is a score of 132, which is why older Stanford-Binet Form L-M reports show slightly larger numerals for the same rank.

How rare is an IQ of 145?

About one person in 740 on a 15-point scale. It is z = 3, the 99.87th percentile, so the upper tail holds 0.135% of the population and the reciprocal is roughly 741. Treat that figure as the output of the normal model rather than a measured frequency: standardisation samples contain too few people beyond three standard deviations to confirm the curve's shape there.

Why do some tests use a standard deviation of 16?

Historical choice. The Stanford-Binet Form L-M was constructed with a standard deviation of about 16, while Wechsler built his scales on 15, and both conventions persisted. Current editions of both families use 15. Always check the manual before comparing numerals, because the difference grows with distance from the mean: at the 99.9th percentile it is worth about three points.

What is the difference between percentile rank and percentage correct?

Percentile rank describes your position relative to other people; percentage correct describes your performance on the items. A percentile rank of 84 means 84% of the norm sample scored no higher than you, and says nothing about how many questions you answered. Two people with identical percentile ranks on different tests can have answered very different fractions of the items correctly.

Can I convert a subtest scaled score with this calculator?

Yes, if you change the mean and standard deviation. Wechsler subtest scaled scores use a mean of 10 and a standard deviation of 3, but this calculator's standard deviation menu offers 15, 16 and 24, so the direct route is to convert the scaled score to z by hand — (scaled − 10) ÷ 3 — and read the percentile from the reference table. A scaled score of 13 is z = 1, the 84th percentile.

What IQ corresponds to the top 2% for a gifted programme?

About 131 on a 15-point scale. The 98th percentile is z = 2.054, and 100 + 15 × 2.054 = 130.8. Programmes that use the 95th percentile are looking for 125, and the 97th percentile is 128. Switch this calculator to start from a percentile rank and it performs that inversion directly, which is the calculation a coordinator needs when the policy is written as a percentage rather than a score.

Why does my report show a range instead of a single score?

Because measurement is imperfect and reputable publishers say so. The confidence interval is built from the test's standard error of measurement, typically two to four points for a full-scale score, giving a 95% interval about ten points wide. The honest interpretation converts both ends of that interval to percentiles and describes the band, which for a score of 124 runs from roughly the 90th to the 97th percentile.

Does a higher IQ mean proportionally more ability?

No. The scale has no true zero and no meaningful ratios, so a 140 is not twice a 70 in any sense. The only defensible reading is positional: how far from the mean, and therefore what share of the population scores lower. This is why percentile rank and z-score are the quantities to reason with, and why differences of a few points should be treated as measurement noise rather than as real distinctions.

References

  • WAIS-IV Technical and Interpretive Manual — Pearson / NCS Pearson
  • Standards for Educational and Psychological Testing — American Educational Research Association, American Psychological Association and National Council on Measurement in Education
  • Essentials of Psychological Testing, 5th ed., Lee J. Cronbach — Harper & Row