The exam is shorter, taken on a screen, and the second half of each section changes difficulty based on how you did in the first. Raw right-answer counts stopped mapping onto scores the way they used to.
The SAT that students sit today is a different instrument from the one their older siblings sat. It is taken on a device rather than on paper, it runs closer to two hours than three, and — the part that actually changes how a score works — it adapts.
Each of the two sections is delivered in two modules. Everyone gets the same kind of first module. Performance on that first module then determines which version of the second module you receive: a harder set or an easier one. The exam is not adapting question by question — that is a different design — but stage by stage, which is why it is called multistage adaptive testing.
The practical effect is that two students can answer the same number of questions correctly and receive different scores, because they answered different questions. A correct answer on a hard item carries more information about ability than a correct answer on an easy one, and the scoring model uses that.
Two students, the same number of right answers, different scores — because the exam already knew the questions were not equally hard.
Keeping 400 to 1600 was not conservatism. Admissions offices, scholarship rules, state reporting and years of institutional data are all built on that scale, and moving it would have broken comparisons the whole system depends on. The work of keeping a new format on an old scale is done through equating — statistical procedures that place scores from different forms onto a common metric so that a given score means the same thing across administrations.
That is ordinary practice in large-scale testing and it is also where the technical scrutiny goes. Equating is what makes a score from one form comparable with a score from another, and it is the thing a testing programme has to document publicly for that comparability to be checkable.
The SAT is an achievement test. It is designed to measure what a student has learned in reading, writing and mathematics, against a curriculum. An IQ test is designed to estimate reasoning ability with as little dependence on taught content as its designers can manage. The two correlate — almost all cognitive measures correlate with each other to some degree — but they are built for different purposes, normed against different populations and validated against different outcomes.
This distinction gets blurred constantly, including by conversion tables that claim to turn an SAT score into an IQ. Those tables are not produced or endorsed by either the test maker or the publishers of intelligence scales, and the populations involved make the conversion unsound: SAT takers are a self-selected group of college applicants, not a representative sample of the general public, so a percentile in one group does not translate to a percentile in the other.
If what you actually want is a read on reasoning rather than on curriculum, an admissions exam is the wrong instrument, and converting one into the other is not something the numbers support. A general assessment reports a score on a named scale with the percentile against a stated reference sample and the confidence range around it — a different measurement with a different meaning.
Find your IQ score now! →For students the change is mostly logistical: a shorter exam, on a screen, where the second half responds to the first. For anyone reading scores, the useful correction is smaller and more specific — stop thinking in raw right-answer counts, because the exam stopped scoring that way.
Internationally in 2023 and in the United States from spring 2024. The format is multistage adaptive: each section is delivered in two modules, and the difficulty of the second depends on performance in the first.
Yes. The scale was kept deliberately so that scores remain comparable with earlier years for admissions, scholarships and institutional reporting. The comparability is maintained through equating procedures documented by the test maker.
Not soundly. The two are built for different purposes and normed on different populations — SAT takers are self-selected college applicants rather than a representative sample of the general public — so a percentile in one does not translate into a percentile in the other. Conversion tables circulating online are not endorsed by the test maker or by the publishers of intelligence scales.
Corrections: spotted an error? Email corrections@iqmetrics.org and we will update this story and note the change here.
On the scale most modern tests use, 120 sits around the 91st percentile. Change the scale and the same number moves. Add the measurement error every test carries and it stops being a point at all.
Aptitude tests are among the better-studied hiring tools, and for decades the headline validity figures were quoted with more confidence than the corrections behind them deserved. A 2022 reanalysis pulled those numbers down.
Two people can perform identically on two different online tests and walk away with scores twenty points apart. The scale, the comparison group and the margin of error are what separate an assessment from a quiz.
Our IIF-certified assessment reports your score with its scale, percentile and confidence range — and a breakdown of the cognitive domains behind it.
Start IQ Test →