A situational judgement test hands you a workplace scenario and four things you could do about it, and asks you to pick. There is no arithmetic to check and no fact to recall, which is why candidates report finding SJTs harder to prepare for than any other stage of a selection process. The scoring is not arbitrary, though. It reads for five specific things, and once you know what they are the four options stop looking equally reasonable.
The Difference Between an SJT and Every Other Test in the Process
Aptitude items have a correct answer that exists independently of who is marking. Personality items have no correct answer at all — they are self-report, and the score is a description rather than a verdict. A situational judgement test sits between the two, and the position is genuinely awkward: the responses differ in effectiveness, but effectiveness is a judgement call rather than a fact.
What makes that workable is that the judgement calls are not close in most items. Presented with a colleague who has taken credit for your idea in a meeting, the options will typically include one that does nothing, one that escalates immediately and disproportionately, one that addresses it directly with the person, and one that is passive-aggressive. Reasonable people disagree about the ordering of the middle two. Almost nobody defends the fourth.
So the scoring works on the same principle as most human judgement: the extremes are easy and the middle is where the information is. Your overall band is largely determined by how consistently you avoid the clearly poor option, and your competency scores are determined by which of the two defensible options you tend to prefer.
Escalation: Knowing Whose Problem It Is
Escalation is the competency that measures whether you can tell the difference between a problem you should solve and a problem you should hand to someone with the authority to solve it. It is scored in both directions, which surprises people: escalating too readily is a failure, and so is refusing to escalate.
Under-escalation usually comes from wanting to be seen as capable. It produces the person who spends two days on something a two-minute conversation would have resolved, or who quietly absorbs a safety issue because raising it feels like complaining. Over-escalation usually comes from risk aversion, and produces the person whose manager becomes a routing layer for decisions they were hired to make.
The scenarios that separate the two are the ones where the fastest fix is available to you but is not yours to make. Choosing it is efficient and wrong, and recognising that is most of what the escalation score is reading.
Prioritisation: What You Drop
Prioritisation items are built around genuine conflict rather than a full inbox. Two things are due, both matter, and there is not time for both. The score is not reading whether you work hard — every response in a well-built item involves working hard — it is reading which one you drop and whether you tell anyone.
The response that quietly does both badly is the one most candidates recognise from their own working lives and the one that scores worst, because it converts a resource problem into a quality problem and hides it. The response that picks one, does it properly, and flags the other early scores best in almost every framing, because it leaves the second problem solvable by someone else.
This is the competency that overlaps most with plain reliability rather than judgement, which is why it is worth reading your prioritisation score alongside a result from the Work Ethics & Reliability test. Someone who prioritises well but does not follow through has a different problem from someone who follows through on everything and cannot choose.
Integrity: When Honesty Is Expensive
Integrity items are the ones where the honest response costs something — a deadline, a relationship, your own standing. Anyone will report an error that makes someone else look bad. The scenarios that discriminate are the ones where the error is yours, or where reporting it means a project that has been going well stops going well.
The scoring here is less about whether you would ever lie and more about your default under pressure. A response that delays disclosure until the situation is more convenient is not dishonest, exactly, but it scores lower than immediate disclosure because the delay is usually where the damage compounds. The item is asking whether you treat a cost as a reason to postpone.
It is worth being blunt about the limitation: a candidate who wants to score well on integrity items can score well on them. Self-presentation is available to everyone. What SJT integrity scores are genuinely useful for is the opposite signal — a low integrity score on a test somebody knew was being marked is unusual and worth a conversation.
Empathy and Boundaries: The Pair That Pulls Apart
These two are scored separately because they trade off against each other, and the trade-off is where most workplace friction actually lives. Empathy reads whether you recognise the state the other person is in before you respond to what they said. Boundaries reads whether you can decline, disagree or hold a line without either capitulating or escalating.
High empathy with low boundaries is the profile that absorbs everyone else’s overflow, is universally liked, and burns out. Low empathy with high boundaries is the profile that is technically correct in every exchange and leaves a trail of people who will not raise problems with them again. Neither is a character flaw; both are a pattern, and both are visible in the scores.
The angry-customer scenario is the standard test of the pair, because the effective response requires both at once: acknowledging the state the person is in without conceding a thing you cannot deliver. If you want a longer read on where you land, the Conflict Styles test looks at the same tension from the other direction.
Reading Your Result
Read the five competency scores before the overall band, for the same reason you read section scores before an overall aptitude band: the band is a summary for other people and the competencies are the information. Two candidates with the same band and opposite empathy-boundaries profiles will behave completely differently on their first difficult week.
Then look at the explanations rather than the scores. An SJT is one of the few assessments where the reasoning behind each scenario is more useful than the number, because the reasoning is transferable and the number is not. Understanding why the response you rejected was rated higher is the part that changes what you do on Monday.
Take the Situational Judgement Test — 25 real workplace scenarios, five competency scores, about 7 minutes — and read the competency breakdown rather than the band.