TestGorilla alternatives: what they cost, and what they predict
By Michael Hodge, BSc Psychology (University of Wollongong)Last reviewed 16 August 2026About a 20-minute read
A TestGorilla alternative is any pre-employment assessment tool you would use instead of TestGorilla to
screen candidates before interviewing them. The twelve here differ less in features than in the two
things that decide your bill: what they charge per candidate, and whether they measure skill,
ability or behaviour.
What are you assessing?
Role you're hiring for
What does good look like for this role?
Fit scores every respondent against a target you set once. It takes about a minute, and you can change it later without re-inviting anyone.
Skip this and every dimension starts at the midpoint. You can set it any time.
Free · No signup · Live link in about 10 seconds
Every price below is dated and sourced4 of the 12 publish no price at allNothing on this page is gated
TestGorilla costs $1,704 a year, and most pages telling you otherwise are wrong
We read the pricing page on 16 August 2026 and then checked it against every other page ranking for
this question. Of the eight that quote a number, seven disagree with the live page.
Start with the plans as published. The prices below are the US dollar view.[1] That qualifier matters
more than it should, and the next table explains why.
Read from testgorilla.com/pricing on 16 August 2026
Plan
Price
Credits
Seats
Watch for
Free
$0
10 a month, no rollover
1
5 preset tests only
Core
$142/mo, billed $1,704 annually
Starts at 400 a year
2
No ATS, no API
Plus
From $400/mo, from $4,800 annually
Tiers quoted by sales
Unlimited
ATS, API, custom tests
Enterprise
Contact sales
Unlimited
Unlimited
No published figure
Three things the price alone does not tell you
It is an annual contract, shown as a monthly rate. TestGorilla's own billing FAQ
states that Core and Plus are annual upfront plans with no monthly payment option, and that there is
no free trial.[1] The page still leads with "$142/mo". Nearly every billing complaint in the review
sites traces back to that gap.
Credits expire. One credit covers one candidate taking an assessment of up to five
skills tests, which is generous per candidate. But unused credits do not roll over at the end of the
contract period on any plan, free or paid.[2] If your hiring is uneven, you are buying capacity you
will lose.
Integration is a tier away. ATS integrations and the API are on Plus, which starts
at $4,800 a year.[1] Core gives you two seats that can spend credits. Hiring managers can view for
free, but a third recruiter cannot run anything.
The same plan costs a different amount, and includes a different bundle
The pricing page carries a currency switcher, and it does not simply convert. The credit allowance
changes too.
Currency
Core monthly
Core annual
Credits included
USD
$142
$1,704
400
EUR
€84
€1,008
250
GBP
£71
£852
250
AUD
$118
$1,416
250
INR
₹6,250
₹75,000
250
Two numbers we could not reconcile
The US pricing page says Core "starts at 400 annual credits". TestGorilla's own help centre says
"your Core plan includes 250 credits". Both were live on 16 August 2026. We have used the pricing
page figure throughout and are flagging the conflict rather than quietly picking one, because the
difference changes the cost per candidate by 60%. Ask them which applies to your contract before
you sign.
What everyone else says it costs
This is worth showing, because if you have spent the morning reading comparison pages you have been
given at least six different prices. Prices do move, and pages go stale. Here is what each said on
the day we checked.
We will be wrong eventually too. TestGorilla has changed pricing model four times in about three
years, most recently overhauling credits in January 2026 and retiring the last legacy plans in May.
The date on this section is the only claim we can really stand behind, which is why it is there.
Cost per candidate
The number that decides this is cost per candidate, and nobody prints it
Every page in this category argues about annual fees. No buyer thinks in annual fees. The arithmetic
is one division and it changes the answer completely.
At full utilisation, $1,704 across 400 credits is $4.26 a candidate, which is a good price. Almost
nobody runs at full utilisation. A company hiring two roles a year with thirty candidates each uses
sixty credits and loses the other 340.
A 400-credit bundle against the hiring an occasional employer actually does. The sticker rate and
the real rate differ by a factor of six.
Move the sliders to your own numbers. The comparison underneath uses published list prices only.
What one year of screening actually costsPublished rates, USD, checked 16 August 2026
Everyone who starts an assessment, not everyone who applies.
TestGorilla bundles up to five into one credit. Equip charges per test.
60
TestGorilla credits needed
$1,704
Their annual bill
$28.40
Per candidate screened
TestGorilla$1,704 · $28.40
TestDome$1,000 · $16.67
Equip$180 · $3.00
SeeMyPersonality$948 · $15.80
TestGorilla. Core covers this. 340 of the 400 credits expire unused, because credits do not roll over.
The others. Smallest pack that covers it: 100 candidates for $1,000. Packs never expire. $1 per candidate per test, so 3 tests each. Nothing expires. 5 scored respondents a month fits the $948 annual tier, assuming the year is spread evenly. Hire in bursts and you would need the next tier up, where TestGorilla's annual bundle would not care. One personality profile, not a test library.
These are list prices for a like-for-like year of screening, not quotes. They buy different things: TestGorilla and TestDome sell a library of skills tests, Equip sells per-test screening, and ours is one personality profile read four ways. A cheaper row is not a better fit, and the section below on what a skills test cannot see is the part that should decide it.
Where the calculator gives up, and why that is the point
Push the volume past the bundled credits and it stops returning a figure. That is not a gap in our
model. TestGorilla does not publish a per-credit top-up rate anywhere, and Plus credit tiers are
quoted by sales. Above roughly 400 candidates a year, the cost of the product cannot be determined
from the information the company publishes. Four of the twelve alternatives below have the same
problem.
Why teams leave
Five reasons teams leave, in the order they actually come up
Ordered by how often they recur across G2, Capterra and Trustpilot, read on 16 August 2026. The first
surprise is that price is not the top one.
4.5
G2, about 1,450 reviews
Employer-weighted
4.1
Capterra, 266 reviews
Support rated 3.9
3.8
Trustpilot, about 1,711 reviews
Candidate-weighted
That spread is the most useful thing on any review site here. G2 and Capterra are where employers
go. Trustpilot is where the people who were made to sit the test go. A page quoting only the 4.5 is
telling you half the story, and it is the half that does not have to take the assessment.
1
The plan reads monthly and bills annually
TestGorilla advertises "$142/mo" and its own FAQ says there is no monthly payment option. Reviewers describe finding out after the fact. There is also no free trial, which its FAQ states plainly.
2
You buy credits by the year and lose what you do not spend
Credits expire at the end of the contract period and do not roll over on any plan. An employer hiring in bursts pays for a full bundle and uses a fraction of it.
3
The tests feel generic for the role
This is the single most common complaint on G2, tallied at 23 mentions, ahead of price. The recurring phrasing is that questions are abstract and do not reflect the skills the job needs.
4
The feature you assumed was included is on Plus
ATS integrations and the API are Plus only, which moves the floor from $1,704 to $4,800. Core also caps you at two seats that can spend credits.
5
Getting money back is harder than spending it
Refund and cancellation difficulty runs through G2, Capterra and Trustpilot, and TestGorilla has replied to 3% of its negative Trustpilot reviews. These are user accounts, not findings we can verify.
Said plainly
The billing complaints in point five are reported experiences from named review platforms, with
dates. We have not verified any individual account and we are not presenting them as findings. What
we can verify is the contract structure that produces them, and TestGorilla documents that itself.
When to stay
When TestGorilla is the right answer
Google shows a question box asking who TestGorilla's pricing suits. Thirteen of the fifteen pages
ranking for these terms do not answer it, because they are all selling something else. Here is the
honest version.
Stay if you hire steadily and at volume. Between roughly 250 and 400 candidates a
year, on two recruiter seats, without an ATS integration, Core works out near $5 a candidate for a
library of 381 tests.[3] Very little in this market beats that, and switching would cost you more in
disruption than you would save.
Stay if you need breadth more than depth. Their own category counts come to 162
role-specific tests, 83 programming, 59 software, 37 language, 17 cognitive and 13 situational
judgment.[3] If you hire across many different roles, that breadth is genuinely hard to replace.
Stay if proctoring matters and budget is tight. This is the thing competitor pages
least like to mention: TestGorilla's anti-cheat features are available on every plan, including the
free one. Most rivals put proctoring behind a paid tier.
Stay if you are already integrated and mid-contract. The cost of moving assessment
history, retraining a panel and rebuilding scorecards is real, and it rarely appears in anyone's
comparison table.
Credit where it is due on one more point. TestGorilla publishes per-test reliability figures with
sample sizes, which is more than most skills-first vendors do at all. Their abstract reasoning test
reports a Cronbach's alpha of .89 on 617 people.[3] The limits of that disclosure are covered in the
fairness section, but the disclosure itself is above average for this category.
The twelve
Twelve alternatives, grouped by how they charge you
Not ranked. Ranking these against each other would mean pretending they do the same job, and the
pricing model is the thing that most often decides whether a tool fits.
We make one of these, and it is in the table in its place rather than at the top. On published
prices at low volume, Equip and TestDome undercut us and nearly everyone else.[18] If cost per
candidate is your binding constraint and you only need skills screening, start there rather than
with us.
Every price read from the vendor's own site on 16 August 2026
Tool
Measures
How it charges
Published price
Suits
Equip
Skills
Per candidate, per test
$1 per candidate per test
Occasional hiring, tiny budgets. Nothing expires.
TestDome
Skills
Per candidate, prepaid
$20 down to $7 a candidate
Clear unit pricing. Packs never expire.
Alva Labs
Cognitive + personality
Base fee plus per job slot
€370 an extra role. Base not published
European teams wanting published validity.
Bryq
Personality (16PF)
Flat, banded by headcount
$69/mo billed annually
Small teams wanting traits plus ability.
Adaface
Skills
Annual, bundled credits
From $180 a year
Low-volume technical screening.
Canditech
Skills
Annual, tiered by volume
$150/mo billed annually
Job-simulation style tasks.
HackerRank
Coding only
Tier plus candidate attempts
$99/mo, or $948 a year
Engineering hiring at volume.
Codility
Coding only
Subscription plus invites
From $1,200 a year
Engineering, real IDE conditions.
SeeMyPersonality
Personality
Monthly, by scored respondents
$99/mo, or $948 a year
One profile read as fit, team, culture and growth. Ours.
Criteria Corp
Cognitive
Quote, scaled by headcount
Not published
Deep aptitude bench, heavy compliance.
Vervoe
Skills
Quote, enterprise only
Not published
AI-graded work samples at scale.
iMocha
Skills intelligence
Quote
Not published
Mapping skills of staff you already have.
Predictive Index
Personality
Flat annual, by headcount
About $10,000 a year
Whole-company behavioural programme.
Two corrections worth carrying into your shortlist
Toggl Hire is gone. Toggl retired it, its own page says it is no longer accepting
signups, and its app subdomains no longer resolve. Several pages ranking for this term still list it
as a live option.[22]Bryq is $69 a month, not $168. The $168 figure circulating in
comparison tables is a "save $168" badge on their pricing page that has been parsed as a price.[20]
Go deeper: what "not published" tells you
Four of the twelve publish no starting price: Vervoe, Criteria Corp and iMocha route to a sales
form, and Alva Labs publishes its per-role overage of €370 but not its base fee.[21] iMocha's
pricing URL redirects to a demo booking.
Treat that as information rather than an obstacle. Quote-only pricing usually means the price
varies by headcount, which means it will rise as you grow whether or not your hiring does. It also
means any number you read for those vendors on a comparison page came from somewhere other than
the vendor, and in this category those numbers are wrong more often than they are right.
What skills tests miss
A skills test tells you if they can. It does not tell you if they will.
This is the real fork in the decision, and almost no comparison page frames it, because most of them
are written by skills-testing vendors.
TestGorilla's library is 381 tests by its own category counts. Six of them are personality and
culture. That ratio is not a criticism, it is a description of what the product is for: it is a
skills library with a small behavioural corner attached, and it is sold to people whose problem is
"can this person do the work".
If your problem is that people you hire can clearly do the work and then leave, clash or stall, a
bigger skills library will not touch it. That is a different measurement.
What the behavioural side honestly adds, and what it does not
We should be careful here, because we sell a personality instrument and the evidence does not
support overclaiming. Once you already have a structured interview, biodata and an integrity check
in place, a general conscientiousness measure contributes about 3% of the predictable variance in
job performance.[6] Anyone telling you a personality test is the main event is selling.
Where it earns its place is narrower and more useful. Personality measures predict counterproductive
behaviour, which a skills test cannot see at all: agreeableness is the strongest predictor of
interpersonal deviance and conscientiousness of organisational deviance.[9] And how you ask matters.
Conscientiousness measured with work-framed items reaches .25 with essentially no variation across
studies, against .19 for generic wording.[4][8]
There is also a live disagreement worth knowing about. In 2007 six former journal editors argued
that self-report personality tests should be abandoned for selection outright, and were answered in
the same issue.[12] That argument has not been settled. We think the honest position is that a
personality measure is a useful second input and a poor first one.
Applicants do inflate. Across 33 studies they score about half a standard deviation higher than
non-applicants on conscientiousness and emotional stability.[10] The usual vendor answer is that this
leaves overall validity intact,[11] which is broadly true and slightly beside the point: you do not hire
on a correlation, you hire from the top of a ranked list, and a shift concentrated among the people
who chose to inflate reshuffles exactly that part of it. Work-framed items and a structured
interview do more about this than any faking-detection scale.
What predicts performance
The validity numbers on vendor science pages were withdrawn in 2022
If a vendor tells you work samples predict performance at .54, they are quoting a 1998 figure that the
field corrected downward four years ago. You can check this in five minutes on any of their sites.
The 1998 estimates everyone quotes came from Schmidt and Hunter, who applied a range-restriction
correction uniformly across studies.[7] Most of those studies did not warrant it. In 2022 Sackett,
Zhang, Berry and Lievens redid the analysis and published a corrected matrix.[5][4]
Operational validity for predicting overall job performance, before and after the 2022 correction.
Higher is better. The work sample, which is what a skills test is, fell furthest.
Two things follow. Structured interviews, not tests, are now the strongest single method. And the
largest downward revision of all landed on work samples, which is the category most of this page's
alternatives sell. Codility, to their credit, cite the corrected .33 on their own validity page
rather than the flattering older number.[19]
Method
2022
1998
Score gap
Grade
What to remember
Structured interview
.42
.51
0.23
A
Now the strongest single method
Job knowledge test
.40
.48
0.54
A
Narrow, and only for known content
Biodata, empirically keyed
.38
.35
0.33
A
Higher floor than interviews
Work sample or skills test
.33
.54
0.67
A
The largest downward revision
Cognitive ability
.31
.51
0.79
A
Largest group differences
Integrity test
.31
.41
0.10
A
Good fairness, candidates dislike it
Conscientiousness, work framed
.25
n/a
−0.07
A
Beats generic wording
Conscientiousness, generic
.19
.31
−0.07
A
Adds 3% once other methods are in
Unstructured interview
.19
.38
0.32
B
Can make predictions worse
Grade A means a meta-analysis or a replicated finding. Grade B means one strong study. Score gap is
the standardised Black and White difference on the method itself, which the next section is about.
Do not read this as "assessment does not work". Sensible combinations of several methods still
average about .47, barely down from .51 under the old estimates.[4] The correction hits single-method
marketing claims, not the practice.
One more finding worth the detour, because it is probably the cheapest improvement available to most
buyers. In a study where people predicted students' grades, adding an unstructured interview made
their predictions worse than using prior grades alone. In one condition the interviewee
answered at random. Interviewers did not notice, and formed confident impressions anyway.[15] If you
run unstructured interviews after your assessment, that is where your money is going.
Numbers we deliberately did not use
You will meet these on nearly every page in this category, and we could not stand any of them up.
"A bad hire costs 30% of first-year salary, according to the US Department of Labor"
has no locatable primary source; every citation points at another blog.
"Recruiters spend six seconds on a CV" comes from a 2012 marketing study of about
thirty recruiters, revised upward by its own authors in 2018.
"The MBTI is used by 89 of the Fortune 100" is the publisher's own marketing claim,
and that publisher's ethics guidance says not to use it for hiring.
"75% of CVs are rejected by an ATS before a human sees them" traces to a sales pitch
by a company that closed in 2013. We would rather have a shorter page than use them.
The fairness numbers
Skills tests are not the fair alternative to aptitude tests
The category's central marketing claim is that testing skills removes the bias in CVs and IQ tests. On
the standard measure of adverse impact, work samples sit second worst of the eight methods here.
Each method plotted by how well it predicts performance and how large a group difference it
produces. Up is fairer, right is more predictive. Structured interviews occupy the corner everyone
says they want.
Work samples produce a Black and White score gap of 0.67, against 0.79 for cognitive ability and
0.23 for a structured interview.[4] A skills test is roughly three times the gap of the interview it is
often sold as an improvement on. Integrity tests, at 0.10, are the fairest thing on the chart and the
method candidates like least.
This is not an argument against skills testing. It is an argument for knowing the number before your
general counsel does. The trade-off is also smaller than it used to look: reaching the four-fifths
ratio with a sensible weighting of methods costs about .09 of validity,[6] not the collapse the older
literature implied.
The sentence that should decide your shortlist
The Uniform Guidelines on Employee Selection Procedures rule out, in terms,
"all forms of promotional literature" as evidence that a selection procedure is valid.[17] A vendor
badge reading "scientifically validated" or "EEOC compliant" carries no legal weight, and the EEOC
certifies nothing. If a test contributes to rejecting applicants, the burden of showing it is job
related sits with you.
So the question to ask a vendor is narrow and answerable: do you publish criterion validity
and adverse impact for the specific tests I will use? Reliability alone does not answer it.
A test can be highly consistent and still predict nothing.
Applied to TestGorilla, the fair reading is mixed. They publish per-test reliability with sample
sizes, which most skills-first rivals do not. On the two tests we checked, criterion validity and
the race and ethnicity analysis both read "pending", and there is no technical manual or published
bias audit. Alva Labs and Bryq publish more on this than most, and it is reasonable to ask the
others why they publish less.
What candidates go through
Somebody has to sit the thing, and no comparison page asks how that goes
Every page on this SERP is written for the buyer. The people who actually experience the product are
on Trustpilot and Reddit, and they are the reason those scores are lower.
Straight from TestGorilla's own candidate help centre, on 16 August 2026:[2]
Most assessments take about 60 to 90 minutes, unpaid.
You can take an assessment once.
One question per screen, and you cannot go back after submitting, even if you skipped it.
Full screen is required. Leaving hides the question, the timer keeps running, and the employer sees it.
Whether you ever see your score is the employer's decision.
None of that is unusual for the category, and the anti-cheat rationale for most of it is real. It is
still an hour and a half of somebody's evening, and half of the complaints in this category are
about being asked for that before a human has spoken to them.
Shorter is not the fix everyone assumes
Every vendor in this market competes on assessment length. The research does not support it.
Assessment length did not predict candidates quitting; most drop-off happens in the first twenty
minutes whatever the total; and giving a longer, more honest estimate upfront was associated with
fewer people quitting, not more.[16] Some of the drop-off you would remove by shortening is weaker
candidates removing themselves, which is the test doing its job.
What candidates do respond to is whether the task looks like the work. Across 86 samples and nearly
49,000 people, interviews and work samples are rated most favourably, cognitive tests lower, and
personality inventories lower still.[13] That ordering held across 17 countries.[14] Face validity, the
sense that this is obviously relevant, is the strongest driver of whether people think a process
was fair.
The practical read: a relevant 25-minute assessment beats an abstract 8-minute one on both
completion and goodwill. Which is awkward for a spec war fought on minutes.
Choosing
Start from your situation, not from somebody's shortlist
Four entry states cover most people who search this. Find yours and the shortlist mostly writes
itself.
You hire a few times a year
Buy by the candidate, not by the year
Two roles and sixty candidates on a 400-credit annual bundle works out at $28.40 a candidate, and 340 credits expire. A prepaid or per-candidate tool prices the same year in the low hundreds. This is the clearest case for leaving, and it has nothing to do with product quality.
You hire steadily, one or two recruiters
The incumbent is probably fine
At 300 to 400 candidates a year on two seats with no ATS integration, Core lands near $5 a candidate, which almost nothing beats for a library that size. Switching costs you more than you save.
You need it wired into your ATS
Price the Plus tier before you compare anything
ATS integrations and the API sit on Plus, from $4,800 a year. Several tools include integrations on their entry plan. Compare like for like or the shortlist is meaningless.
You keep hiring people who can do the job but do not last
You have a fit problem, not a skills problem
A skills test answers whether someone can do the work. It says nothing about how they handle pressure, feedback or a team. That is a different instrument, and the honest version of it is a structured interview first, with a personality measure filling the gap it leaves.
Ten questions to ask before you sign anything
Works for any vendor on this page, including us.
Is the advertised monthly price a real monthly option?TestGorilla shows $142/mo and bills $1,704 upfront. Ask for the contract term in writing.
Do unused credits or invites expire?Expiry turns a low sticker price into a high per-candidate one the moment your hiring is uneven.
What does a year cost at MY volume, not at full use?Divide the annual fee by the candidates you honestly expect. That number is the one to compare.
Which features move me to the next tier?ATS integration, the API and extra seats are the three that most often triple the bill.
How many people can spend credits, not just log in?View-only seats are often unlimited while the seats that matter are capped at two.
Can I see the questions before I buy?Several vendors will not show you the items, which makes the score hard to defend later.
Do they publish criterion validity, or only reliability?A high alpha means the test is consistent. It does not mean it predicts the job.
Do they publish adverse impact by race, age and gender?If those cells read "pending", the legal exposure is yours to carry, not theirs.
How long is it for the candidate, and do they get anything back?Sixty to ninety unpaid minutes with no feedback is a real cost to your employer brand.
What happens to my data and results if I leave?Contract lock-in is the most cited complaint in this category and the least documented.
Take it into the demo call. The answers to questions seven and eight are the ones vendors are least
prepared for.
Questions
TestGorilla alternatives and pricing, answered
Is TestGorilla legit?
Yes. It is an established Dutch company holding SOC 2 Type II and GDPR compliance, with roughly 1,450 reviews on G2 at 4.5 out of 5 and 266 on Capterra at 4.1. Its Trustpilot score is lower at 3.8 across about 1,711 reviews, mostly because Trustpilot attracts job candidates rather than employers. The complaints that recur are about billing terms and test relevance, not legitimacy.
How much does TestGorilla cost?
Checked on 16 August 2026, the Core plan is $142 a month billed as $1,704 annually and includes a bundle starting at 400 credits in US dollars. Plus starts at $400 a month, or $4,800 a year. There is a free plan with 10 credits a month. Prices and credit bundles differ by currency: the euro and sterling versions include 250 credits rather than 400.
Does TestGorilla have a monthly plan?
No. The price is displayed per month but its own billing FAQ states that Core and Plus are annual upfront plans with no monthly payment option. This is the most common source of complaints about the product.
Is TestGorilla free?
There is a free plan, and it is a real free plan rather than a trial. It gives you 10 credits a month that do not roll over, five preset tests instead of the full library, and one seat. It excludes ATS integrations, the API, analytics and custom branding. TestGorilla states separately that it does not offer a free trial of the paid plans.
Do TestGorilla credits roll over?
No. Credits expire at the end of the contract period on both free and paid plans. One credit covers a candidate taking an assessment of up to five skills tests. Credits are only deducted when a candidate actually starts, so people who never show up cost you nothing.
What is the cheapest alternative to TestGorilla?
On published list prices, Equip at $1 per candidate per test and TestDome from $7 to $20 per candidate are the cheapest credible options, and both charge only for candidates you actually assess. Neither includes a personality instrument. Cheapest and best fit are different questions, and the answer depends on whether you need skill, ability or behaviour.
Which TestGorilla alternatives publish their prices?
Of the twelve here, eight publish a real starting price: Equip, TestDome, Adaface, Canditech, Bryq, HackerRank, Codility and ours. Vervoe, Criteria Corp and iMocha publish nothing and route to a sales form, and Alva Labs publishes its per-role overage but not its base fee. Treat any number you read elsewhere for those four with caution.
Do pre-employment assessments actually predict job performance?
Some do, and by less than vendors claim. A 2022 re-analysis by Sackett and colleagues corrected the widely quoted 1998 figures downward: structured interviews now sit at .42, work samples and skills tests at .33, cognitive ability at .31 and general conscientiousness at .19. Well-built combinations of methods still reach about .47, so assessment works. Single-method claims above .5 are quoting numbers that were withdrawn.
Are personality tests legal for hiring?
In the United States, yes, provided the test is job related and consistent with business necessity, and does not produce unjustified adverse impact. The Uniform Guidelines are explicit that a vendor’s own marketing cannot serve as evidence of validity, so the burden of showing a test is job related sits with the employer, not the supplier. Medical or disability-related inquiries are separately restricted.
Can candidates cheat on pre-employment assessments?
On unproctored skills tests, yes, which is why anti-cheat tooling exists. On self-report personality measures the more useful question is inflation rather than cheating: applicants score about half a standard deviation higher on conscientiousness and emotional stability than non-applicants. That shifts who reaches your shortlist even where it leaves the overall correlation with performance broadly intact.
How long should an assessment take?
Shorter is not automatically kinder. Published research found assessment length did not predict candidates dropping out, that most quitting happens in the first twenty minutes whatever the total length, and that giving a longer, more honest time estimate upfront was associated with fewer people quitting. What candidates respond to is whether the task looks like the job.
What should I ask a vendor before signing?
Whether the contract is annual and whether the advertised monthly price is a real monthly option. Whether unused credits expire. Which features sit behind the next tier, especially ATS integration and extra seats. What the per-candidate cost is at your real volume rather than at full utilisation. And whether they publish criterion validity and adverse impact figures for the specific tests you will use, not just reliability.
Every price on this page was read from the vendor's own website on 16 August 2026, in a browser, not
taken from a comparison page. Where a vendor publishes nothing, this page says "not published"
rather than repeating a third-party figure, because in this category those figures are wrong more
often than they are right. TestGorilla's US pricing page, its billing FAQ, its credits help article
and its candidate help centre were all read the same day.
How the evidence was graded
Validity and adverse impact figures come from the 2022 and 2023 Sackett and 2024 Berry papers rather
than from the 1998 estimates still in circulation. Grade A means a meta-analysis or replicated
finding, B a single strong study, C a famous but contested one. Nothing graded C appears as fact on
this page, and the numbers we rejected are listed in section seven rather than quietly dropped.
Limits, stated plainly
Pricing in this category changes often, and TestGorilla has changed model four times in about three
years, so treat 16 August 2026 as the shelf life of section one rather than a permanent claim. The
cost calculator uses list prices and cannot account for negotiated discounts, which at the top of
this market are normal. We have not run every one of these tools, and the review-site complaints are
other people's reported experiences, not our findings. Validity figures are averages across many
studies with wide spreads, so they describe methods in general and not the specific product you buy.
Reviewed by Michael HodgeContent last reviewed 16 August 2026
Disclosure: SeeMyPersonality is our own product and appears in the comparison table above. Nothing
on this page is for sale, there is no signup, and the assessment builder at the top is free to use
without an account.
If you only take one thing from this page, make it the division: your annual fee over the candidates you
honestly expect. Build a free assessment at the top if you want to see what the
behavioural half looks like before you decide.