A pre-employment personality test reads how someone tends to work: how they plan, how they handle pressure, and how they get on with people. This page shows you a real candidate report, what the evidence actually supports, and what the law asks of you. Building an assessment and sending it costs nothing.
What does good look like for this role?
Fit scores every respondent against a target you set once. It takes about a minute, and you can change it later without re-inviting anyone.
Skip this and every dimension starts at the midpoint. You can set it any time.
Free · No signup · Live link in about 10 seconds
Not a screenshot. It is the live report, running the same scoring as the dashboard, for an illustrative Sales Development Rep. Click through the tabs. Hiring for that exact role? There is a dedicated sales hiring assessment, and a culture fit assessment for reading team complement rather than role fit.
Strong Alignment, Promising with Probes, Stretch Fit, Reconsider. The score behind each one is worked out in full at the foot of this page.
The shaded band is the range you set for that trait. The marker is the candidate. You see the gap rather than take our word for it.
Every watchout arrives with something to ask about it, so a low score turns into a conversation instead of a rejection.
A response-quality check flags rushing, straight-lining and inconsistency. It reads the answering, never the person.
Candidates keep their own copy of the report. The sample uses illustrative data, not a real applicant.
Partly. Enough to be worth running, nowhere near enough to decide on.
Conscientiousness predicts performance in almost every kind of job, which Barrick and Mount established in 1991[2]. The correlation is about .22. Critics of personality testing quote roughly the same figure, and they are right to. It is a real effect and a modest one.
Be careful with older league tables. Sackett and colleagues re-estimated the classic numbers in 2022[1] and several came down hard. Cognitive ability fell to about .31, behind a well-built structured interview at .42. The methods still work. The lesson is to stop crowning one of them and combine two or three that measure different things.
Tett and Christiansen[3] add the condition that makes any of this defensible: the trait has to connect to real work behaviour. That is why you set the target ranges first, and why you see them before a candidate answers anything.
Each one measures something different and predicts a different slice of the job. Most teams run two or three. Running two that measure the same thing buys you almost nothing.
| Type | What it measures | Best for | Watch out |
|---|---|---|---|
| Personality (Big Five) | How someone usually works: planning, pressure, collaboration, follow-through. | Role fit, interview questions, onboarding. | Normal-range traits only. Never a pass/fail gate on its own. |
| Cognitive ability | Reasoning and how quickly someone picks up unfamiliar work. | Complex roles with a steep learning curve. | The highest adverse-impact risk of the common methods. |
| Skills and job knowledge | Whether someone already has the specific skill the job needs. | Roles with a clear technical core: code, copy, bookkeeping. | Measures what someone can do now, not what they could learn. |
| Situational judgement | What a candidate chooses to do in realistic on-the-job dilemmas. | Customer-facing and people-heavy roles. | Only worth running if built from a real job analysis. |
| Integrity | Dependability and the pull toward rule-breaking at work. | Cash-handling, safety-critical and high-trust roles. | Can feel intrusive. One signal, never a verdict on character. |
| Work sample | Actual performance on a real slice of the job. | Almost any role you can cut down to a short, fair task. | Costs everyone time, and the scoring drifts unless you fix it. |
| Structured interview | Job-relevant behaviour, same questions and scoring for everyone. | Every hire you make. | Only works if you really score against the anchors. |
| Physical or role-specific | Physical capacities the job genuinely requires. | Safety-critical and physically demanding work. | The highest ADA exposure. Strict rules on timing. |
Popularity and defensibility are different things. Several well-known instruments are sold for hiring by people who should know better, including by their own resellers.
We sell one of the tools in this table, so read the first row with that in mind. What would change it: if the Big Five stopped predicting work behaviour in replicated research, that row should say no. The published evidence is linked at the foot of the page so you can check rather than trust us.
| Instrument | Question it answers | Select on it? | Why |
|---|---|---|---|
| Big Five (IPIP-NEO) | How does this person tend to work? | Yes, role-mapped | The strongest evidence base for workplace behaviour. Tie each trait to a job behaviour and pair it with an interview. |
| Structured interview | How has this person handled the work before? | Yes | The cheapest accuracy upgrade most teams can make, and the strongest single method in the 2022 re-analysis. |
| Work sample | Can this person do the job? | Yes | The closest thing to watching someone work. Standardise the scoring before you start. |
| Cognitive ability | How fast will this person pick the work up? | Yes, carefully | Predictive, but it carries the highest adverse-impact risk. Document why the job needs it. |
| Situational judgement | What would this person do in a tricky moment? | Only if built for the job | Off-the-shelf versions measure very little. One written from a real job analysis measures a lot. |
| DISC | How does this person prefer to communicate? | No | Publishers position DISC as a communication and development language, not a selection instrument. |
| MBTI / 16 types | How does this person prefer to think? | No | The Myers-Briggs Company states plainly that the MBTI should not be used for hiring decisions. |
| CliftonStrengths | What language helps this person develop? | No | Built for coaching people you have already hired. It was never designed to choose between candidates. |
| Enneagram | What self-reflection fits this person? | No | No accepted validation for selection. It belongs in personal development, not in a hiring file. |
"No" here means not defensible as a basis for choosing between candidates. Several of these are perfectly good tools for developing people you have already hired.
Type the job title, or paste the description you already wrote. You get a live link straight away, with no account.
You see the target range for every trait before a single candidate answers anything. Change any of them. Your judgement wins.
Each candidate returns a fit band, the gaps worth probing, and questions to ask about them. Then you interview properly.
Ten minutes on a phone. No account to create, no password to forget, plain workplace language, and a progress bar that tells the truth. None of the clinical framing that makes people feel examined rather than met.
They also keep their own copy of the results. A test that gives something back feels like a fair trade rather than data extraction, and that goodwill survives whether or not you hire them. You can take the same questionnaire yourself before you send it to anyone, which is the fastest way to judge whether you would resent receiving it.
A number should sharpen the next conversation, not end it. The moment a score rejects someone on its own, you have stopped hiring and started gating.
Good teams are built from differences people understand and support. They are not one person copied seven times.
Tune the criteria once you know who applied and you will always find what you hoped to find. Lock the profile first.
A careful operations role and a high-ambiguity sales role should not share a template. Being specific is the entire point.
If your interviewers cannot explain the score, they should not be using it. People need to be able to argue with the data honestly.
Four things matter, and none of them are complicated. Most of the trouble comes from skipping the boring parts, not from the test itself.
The EEOC counts a personality test as a selection procedure[6], so it has to relate to the work. Write down why each trait matters for that role before you send anything, and keep it.
The ADA restricts medical examinations before a job offer, and a test built to reveal a condition counts as one. This is a normal-range workplace questionnaire and asks nothing clinical. Keep it that way.
This is what adverse impact means: one protected group rejected at a much higher rate than another. It does not make a test illegal, but it does put the burden on you to show the method is job-related. Same test, same scoring, same process for everyone, and keep the records.
New York City treats some hiring software as an automated employment decision tool, or AEDT, and may require a bias audit and candidate notices[7]. The EU restricts decisions made purely by machine. Both get simpler when a person makes the actual decision, which is how this is meant to be used anyway.
This is general information, not legal advice. If you hire at volume, in regulated roles, or across several countries, have someone qualified read your process.
Candidates answer the IPIP-NEO, a public-domain Big Five inventory[4]. Answers become percentile scores against the norm sample, at both domain and facet level.
For each trait you care about, you set a target range. A candidate inside that range scores full marks for it. Outside it, the score falls in proportion to the distance, softened by how wide you made the range plus a fixed 20-point margin. Scoring above a range costs half as much as scoring below it, because being more conscientious than the job needs is rarely the problem.
The headline number is the average of those trait scores, weighted by how much you said each one counts. Above 75 reads as Strong Alignment, above 55 as Promising with Probes, above 35 as Stretch Fit, and below that as Reconsider.
Two limits worth stating plainly. Self-report can be shaded when a job is at stake, so the report includes a response-quality check and we treat the result as a source of questions. And norms are population-level: they describe how an answer compares with a large sample, not what any individual is capable of.
If you want to see it rather than read about it, the builder is at the top of the page.