An aptitude test result is rarely just the number of questions you answered correctly. The publisher converts that raw score into a comparison against a norm group, a reference sample of people who have already sat the same test, and it is the comparison that gets reported to the employer or the university. That is why the figure you are shown is usually a percentile.
Your percentile says where you sit relative to the norm group, not what proportion of the paper you got right. Scoring at the 70th percentile means you performed better than 70 per cent of the comparison group. It does not mean you answered 70 per cent of the questions correctly, and in practice those two numbers are often a long way apart.
Raw score, percentage and percentile are three different things
The raw score is the count of questions marked correct. On its own it tells you very little, because it depends entirely on how hard that particular test was and how much time you were given.
The percentage is the raw score expressed as a share of the questions available. It reads more naturally but it has the same weakness: 60 per cent on a punishing test and 60 per cent on a gentle one are not comparable results.
The percentile solves that by expressing your performance relative to other people. Because it is a rank rather than a mark, it stays meaningful across tests of different lengths and difficulties, which is exactly what an employer comparing a stack of candidates needs.
The consequence catches people out. It is entirely normal to answer only a minority of the questions on a hard, strictly timed test and still land in a high percentile, because everyone else found it hard too. Judging your own result by how many questions you got through is the most common way people misread a score report.
What the norm group actually is
A percentile is only as meaningful as the group behind it, so the norm group matters more than most candidates realise. A publisher might report your score against:
- A general population sample
- A graduate or university student sample
- A sample of people working in a particular occupation or industry
- The employer’s own applicants for that role, or its current post holders
The same performance can produce very different percentiles depending on which of those is used. Against a general population sample, a graduate level candidate will usually look strong. Against a norm group made up of other applicants to a competitive graduate scheme, the same raw score can sit squarely in the middle. Nothing about your ability has changed, only the comparison.
If you are sent a score report, look for the norm group. A good report names it. If it does not, treat the percentile as directional rather than precise.
Sten, stanine and standard scores
Percentiles are not the only scale in use, and a report can look alarming simply because you have not seen its scale before. You may also encounter:
- Sten scores, a scale of 1 to 10 where the middle of the distribution sits between 5 and 6
- Stanines, a scale of 1 to 9 with 5 in the middle
- Standard scores such as z-scores and T-scores, which express how far above or below the average you are in statistical units rather than in marks
These are all doing the same job as a percentile: placing you against a distribution rather than against a mark scheme. Personality questionnaires in particular tend to report on a sten scale, which is one reason a sten of 4 is not a failing result. It is a description of where you sit, not a grade.
How employers set the pass mark
Employers do not usually mark aptitude tests against a fixed pass mark in the way a school exam is marked. The threshold is set to serve the sift, and it moves.
The common approaches are:
- A benchmark percentile. Everyone at or above a set percentile against a chosen norm group goes through. This is the most typical arrangement.
- A volume driven cut. The employer takes candidates from the top down until it has as many as it can interview, so the effective threshold depends on that year’s field.
- Multiple hurdles. Each test has to be cleared in turn, and a weak result on one cannot be rescued by a strong result on another.
- A combined score. Results across several tests are pooled, so strength in one area can offset a weaker area.
Two practical points follow. First, the threshold is a property of the process rather than of you, so the same score can clear one employer’s sift and miss another’s. Second, where an employer uses multiple hurdles, your weakest test type is the one that decides the outcome, and that is where practice pays best.
Speed, accuracy and unanswered questions
Many aptitude tests are built so that few candidates answer every question. That is deliberate, because a test everyone finishes cannot separate the strongest candidates from each other. It also means your approach to unanswered questions matters.
Where there is no penalty for a wrong answer, and usually there is not, leaving a question blank is strictly worse than answering it, so a considered guess on the questions you cannot finish is worth making. Where a test is adaptive, meaning it chooses the next question based on how you answered the last one, the picture changes: rushing produces easier questions and caps the score you can reach, so accuracy carries more weight than volume. If you do not know which kind you are sitting, ask before test day.
What your result does not tell you
An aptitude test measures how you performed on a particular set of tasks, on one day, under time pressure. It is a useful signal, which is why employers and admissions teams use it, but it is a narrow one and it is rarely used alone. Interviews, work sample exercises and your record usually carry as much weight in the final decision.
A result also improves with familiarity. Much of what separates a first attempt from a fourth is knowing the question formats, the way the instructions are worded and where the time goes, and that part is entirely learnable. If a score report has just told you something you did not want to read, the useful response is to find out which test type pulled it down and work on that one specifically.
Practise the tests your result comes from
The quickest way to make sense of a score report is to sit the test types behind it and watch how your raw performance converts:
- Numerical reasoning tests
- Verbal reasoning tests
- Free aptitude tests across the main formats
- Aptitude test sample questions and answers with worked explanations