Talent Assessment Tools: 12 Platforms Compared
Talent assessment tools compared: 12 platforms, cost per hire at 3, 10, and 30 hires a year, what the research says, and the rules tests fall under.
Talent Assessment Tools: 12 Platforms Compared
Why the phrase points at two unrelated products, what each type of test actually measures, what a licence costs per hire when you hire three people a year, what the research now says about which methods predict performance, and the legal standard every test has to meet
A note before the list, because it will save some readers a click. The phrase talent assessment describes two products that have nothing to do with each other. One is bought by employers to evaluate candidates, and the results go to a hiring team. The other is taken by individuals to explore what career might suit them, and the results go to the person taking it. Search results mix them. Everything here is the first kind.
The second thing worth saying early is that this category is priced on the assumption you hire constantly. Most platforms sell annual subscriptions that cost the same whether you test thirty candidates or three hundred, which is excellent value at volume and poor value for an employer filling a handful of roles a year. Almost no comparison does that arithmetic, and it is the calculation that decides whether any of this is worth buying.
This page covers 12 platforms across four groups: broad skills tools priced for smaller employers, technical specialists, behavioural and psychometric instruments, and enterprise volume platforms. It states which vendors publish rates, models cost per hire at 3, 10, and 30 hires a year, sets out what the validity research now says about which methods actually predict performance, and covers the legal standard every test has to meet.
Two different things share this name
This distinction is worth thirty seconds because searching the phrase returns both and they are not substitutes in either direction.
| Employer-side talent assessment | Career or self-assessment | |
|---|---|---|
| Who buys it | A company evaluating applicants or staff | An individual exploring career options |
| Who sees the result | The hiring team | The person who took it |
| What it is for | Comparing candidates against defined criteria | Personal insight and direction |
| Validation evidence | Expected, and legally relevant | Not required |
| Typical price | Annual subscription or per assessment | Free or a one-off fee |
| Also called | Pre-employment testing, candidate assessment software | Career aptitude test, career quiz |
Everything below is the first column. If you arrived looking for the second, the useful adjustment is to search career aptitude test instead, which separates the two result sets cleanly.
What these tools actually measure
Assessment platforms bundle several distinct instrument types under one subscription, and knowing which is which matters because they differ sharply in what they predict and how defensible they are.
| Instrument type | What it measures | Where it fits best | Main caution |
|---|---|---|---|
| Skills and job knowledge | Whether someone can already do specific work | Roles with concrete, testable competencies | Tests knowledge that can be learned, so it favours the experienced |
| Work samples and simulations | Performance on a realistic task from the actual job | Almost any role where a representative task exists | Takes candidate time; keep it short and paid where substantial |
| Cognitive ability | Reasoning and speed of learning | Roles where people must learn quickly | Historically the highest adverse impact risk of the four |
| Personality and behavioural | Self-reported working style and preferences | Team fit discussions and development, not screening | Self-report is fakeable; weakest basis for a reject decision |
The fourth row is the one small employers most often buy first and should probably buy last. A personality questionnaire is cheap, interesting to read, and the least defensible instrument on the list to reject someone on.
Do you need one at your hiring volume
This question goes before the product list because for a lot of employers the honest answer is not yet, and the reason is arithmetic rather than opinion.
| Signal | Points toward buying | Points toward not buying |
|---|---|---|
| Hiring volume | Enough hires a year that per-hire licence cost is small | A handful of hires; the subscription costs more per hire than it saves |
| Role repetition | The same role filled repeatedly, so one test gets reused | Every hire is a different job needing a different instrument |
| Applicant volume | More applicants than anyone can screen properly | You can interview everyone who is plausibly qualified |
| Role type | Concrete testable skills, especially technical | Judgement-heavy roles a test cannot represent |
| Who screens | Several people, needing comparable scores | One person deciding, with no comparability problem |
| What is broken | Interviews produce inconsistent, unfalsifiable impressions | Interviews work; the problem is elsewhere in the process |
The right column describes most small employers, and there is a genuinely good alternative for them that costs nothing. Asking every candidate the same questions in the same order and scoring against the same criteria is a structured interview, and on the current evidence it is one of the strongest predictors available. Adding a short realistic task gets most of the remaining benefit.
12 assessment platforms at a glance
The table below covers all four groups, with the facts that decide most of the evaluation: what each actually tests, and whether anyone will tell you the price.
| Platform | Test focus | Published pricing | Free tier | Entry cost | Billing | Best for |
|---|---|---|---|---|---|---|
| TestGorilla | Broad skills | Free, then about $1,650 a year | Annual only | Multi-role hiring across job families | ||
| Vervoe | Work simulations | Reported from about $19 | Monthly available | Judging people on realistic tasks | ||
| eSkill | Broad skills | Reported from about $415 | Subscription or per use | Customised test batteries | ||
| Mercer Mettl | Broad plus proctoring | Reported usage-based | Usage-based | Remote testing needing invigilation | ||
| HackerRank | Technical only | About $1,990 a year | Annual tiers | Engineering hiring at volume | ||
| Codility | Technical only | Reported from about $1,200 a year | Pay per test option | Engineering teams wanting per-test billing | ||
| Criteria Corp | Cognitive and personality | Quote only | Annual contract | Small and mid-sized employers wanting validated tests | ||
| Predictive Index | Behavioural | Reported from about $5,000 a year | Annual licence | Companies building a behavioural framework | ||
| Wonderlic | Cognitive and personality | Quote only | Quote | Employers wanting a long-established test | ||
| SHL | Full psychometric | Quote only | Enterprise contract | Large employers with validation requirements | ||
| HireVue | Video and game-based | Quote only | Enterprise contract | High-volume structured video screening | ||
| Harver | Volume screening | Reported from about $5,000 monthly | Enterprise contract | Retail, hospitality, and call centre volume |
How we evaluated these platforms
Vendor pages here describe near-identical libraries and claim similar outcomes, and several cite time-to-hire improvements that are marketing rather than independent findings. We applied four tests instead, identically to all twelve.
Priced for smaller employers
These four are the realistic shortlist for a company that is not hiring at volume. Three publish rates and one has the only meaningful free tier in the comparison.
TestGorilla has the broadest library at the accessible end of this category and the only free plan substantial enough to evaluate the approach properly. For a company hiring a customer support person, a bookkeeper, and a developer in the same year, one subscription covers all three, which the specialist tools cannot do. Ratings are strong on the major review platforms.
The pricing model is where small buyers report friction, consistently and specifically. Every paid plan is an annual commitment with no monthly option, the credit system means unused capacity expires rather than carrying forward, and reviewers repeatedly describe candidate limits and rules around repurchasing assessments as expensive surprises. Integration with an applicant tracking system and custom test creation sit on the higher tier, roughly three times the Core price.
Vervoe builds around simulations rather than questionnaires, which puts it closest to the instrument type the evidence treats most favourably. A candidate completes a task resembling the actual job and is ranked on the output, which is both more predictive and considerably easier to explain to a rejected applicant than a personality score. It also has the lowest reported entry point in this comparison.
Building good simulations takes work, and the template library reduces but does not remove that. AI grading of open-ended responses is a genuine convenience and also a place where a human should review before a rejection, particularly given how selection procedures are regulated. Reported pricing varies enough between sources that the tier you actually land on needs confirming directly.
eSkill differentiates on assembly rather than breadth: you build a battery from individual subject areas to match a specific role rather than picking the nearest template. For a business with unusual role combinations, an administrator who also needs spreadsheet and industry knowledge for instance, that granularity produces a more relevant test than a generic library.
Assembly is also effort, and someone has to decide what belongs in the battery and why, which is exactly the judgement that determines whether the test is job-related. Published pricing is thinner than the reported entry figure suggests, so budgeting needs a conversation. It is less known than the market leaders, so independent review evidence is correspondingly thinner.
Mettl, part of a large consultancy group, is strongest on proctoring: identity checks, monitored sessions, and integrity signals designed for testing at distance where nobody can watch the room. For an employer whose test result carries real weight in a decision, that layer is the difference between a score you can rely on and one you cannot.
Usage-based quoting means the cost depends on volume you may not be able to forecast, and the entry figures circulating are third-party reports rather than a rate card. Proctoring also imposes a heavier candidate experience, including camera requirements, which some applicants decline. For a handful of hires a year the invigilation apparatus is disproportionate to the risk.
Technical and coding specialists
These two do one thing and do it better than the general libraries. If you are not hiring engineers they are irrelevant, and if you are they are usually the right answer.
HackerRank is the most recognised name in technical screening, and candidate familiarity is itself an advantage: engineers know the format, which removes a source of noise from the result. Plagiarism detection matters more than it used to now that candidates have capable assistance available, and the live interview environment covers the pairing stage as well as the screen.
The Starter tier caps attempts and gives one seat, so a hiring manager and a technical reviewer both needing access pushes you up. It is useless outside engineering, which means it is a second subscription for any company that also hires non-technical staff. And a well-designed take-home task relevant to your actual codebase is free and arguably more informative for a small team.
The pay-as-you-go option is the reason Codility belongs on a page written for smaller employers: a team hiring two engineers a year can test candidates for a few tens of dollars rather than committing to a subscription that costs the same whether used or not. Its reporting also describes how a candidate approached a problem, not only whether the tests passed, which is more useful in a debrief.
Task quality varies across the library and picking the right difficulty for the role takes judgement, which is true of every coding platform and worth budgeting time for. Brand recognition among candidates trails the market leader slightly. As with all technical tools, it addresses one slice of hiring and nothing else.
Behavioural and psychometric instruments
These three sell validated psychometric instruments rather than skills libraries. All are quote-led, and all carry more legal weight than a skills test because of what they measure.
Criteria occupies a genuinely useful position: psychometric rigour comparable to the enterprise houses, sold to companies that are not enterprises. The validation documentation and adverse impact reporting matter more than they sound, because they are what an employer relies on if a selection decision is ever challenged, and most skills-library vendors provide considerably less of it.
Nothing is published, so evaluation starts with a sales conversation, and reported annual figures span an order of magnitude depending on assessment volume. The instruments are also less flexible than an assembled skills battery: you are buying validated tests as designed rather than composing your own, which is the point but constrains unusual roles.
Predictive Index is bought as a framework rather than a test. The behavioural profile is used in hiring but also in team composition, management conversations, and development, and the licence includes training that turns it into shared vocabulary across a company. Organisations that adopt it properly tend to use it far beyond recruitment, which is where the value sits.
It is priced accordingly and reported entry figures put it several times above the skills libraries, which is a lot for an employer that simply wants to screen applicants. Behavioural instruments are also self-report and the weakest basis on this page for a rejection decision, so using it as a screening gate rather than a discussion input is both less effective and more exposed.
Wonderlic has one of the longest track records in employment testing, and the depth of normative data behind it is a genuine advantage: comparisons are against a large, well-documented population rather than a vendor sample. For an employer whose concern is defensibility, that history is worth something.
Nothing is published, which is standard in psychometrics and still a friction for a small buyer. The heavy cognitive component also carries the highest adverse impact risk of the instrument families on this page, which is not a reason to avoid it but is a reason to run the numbers and document the job-relatedness before using it as a gate.
Enterprise volume platforms
These three dominate category rankings and none will sell to a small employer on sensible terms. Included so you can recognise them and move on.
SHL is the reference house in occupational psychometrics and has been for decades. Its advantage is depth of evidence: normative data across many populations and job families, published validation research, and situational judgement and simulation content built to professional standards rather than assembled from a question bank. Where a selection process genuinely has to withstand challenge, this is the level of documentation that supports it.
None of that is accessible at small scale. Nothing is published, contracts are enterprise-shaped, and the product assumes somebody on the buyer side who understands psychometrics well enough to specify what is needed. A company hiring four people a year is buying rigour for a risk it does not carry at that volume.
HireVue solves a specific problem well: applying the same structured questions to a very large applicant pool without consuming interviewer time linearly. Because every candidate answers identical questions, the comparison is genuinely like-for-like, which is the same property that makes structured interviews strong predictors. At programme scale that consistency is difficult to achieve any other way.
It is scoped and priced for that scale and nothing below it. Automated scoring of video responses also sits squarely in the area jurisdictions have begun regulating, so an employer using it needs to understand its bias audit position and candidate notice obligations rather than assuming the vendor has resolved them. For a small employer, a scheduled video call achieves the same consistency for nothing.
Harver is built around throughput rather than precision on any individual hire. For an employer filling the same hourly role hundreds of times, small improvements in screening efficiency compound into large operational savings, and the assessment content is designed for candidates completing it on a phone without support.
The reported entry figure is the highest on this page, and the economics only work at volumes a small employer will never reach. The whole design assumption is role repetition at scale, so a company with varied roles and occasional hiring gets none of the benefit and all of the cost.
The pattern across this group is consistent: each is credible, each assumes a dedicated talent acquisition function, and each requires a sales process before you learn whether it is affordable. For an employer where the person evaluating this also runs payroll, that is a real cost and a reasonable basis for excluding the group early.
What it costs per hire at three volumes
This is the calculation that decides the purchase and almost nobody publishes it. The table models annual cost against hiring volume rather than quoting monthly rates.
| Option | Rate basis | 3 hires a year | 10 hires a year | 30 hires a year | Notes |
|---|---|---|---|---|---|
| Structured interview process | Your own time | $0 | $0 | $0 | No licence; costs preparation time instead |
| Free tier only | TestGorilla free plan | $0 | Not viable | Not viable | One seat and a handful of tests |
| Vervoe entry | Reported from about $19 monthly | $228 | $228 | $228 | Pay-as-you-go tier reported at the low end |
| Codility pay per test | Reported about $3 per test | About $30 | About $100 | About $300 | Assumes several candidates tested per hire |
| eSkill | Reported from about $415 | $415 | $415 | $415 | Reported entry package, technical roles aside |
| Criteria Corp | Reported $10 to $25 per assessment | $1,200 plus | $1,200 plus | $1,200 plus | Reported annual contracts from about $1,200 |
| TestGorilla Core | About $1,650 a year | $1,650 | $1,650 | $1,650 | Credit-based; unused credits do not carry over |
| HackerRank Starter | About $1,990 a year | $1,990 | $1,990 | $1,990 | One user and a capped number of attempts |
| Predictive Index | Reported from about $5,000 a year | $5,000 | $5,000 | $5,000 | Behavioural licence priced by headcount band |
| TestGorilla Plus | From about $4,800 a year | $4,800 | $4,800 | $4,800 | Where ATS integration and custom tests sit |
| SHL, HireVue, Harver | Quote only | Not published | Not published | Not published | Enterprise contracts scoped well above this range |
Read the subscription rows across and nothing changes, which is the entire point: an annual licence near $1,650 is about $550 per hire at three hires a year and about $55 at thirty. The product did not get better; you simply used it more. That means the question is not whether an assessment platform is worth $1,650, it is whether it is worth $550 per hire against a structured interview that costs preparation time. At thirty hires the same question answers itself the other way.
What the research now says about what predicts performance
Vendor content in this category still cites validity figures from a 1998 summary as though they were settled. They were substantially revised, and the revision changes what a small employer should spend money on.
The practical consequence for an employer with a small budget is direct. If the single strongest instrument is a structured interview, and a structured interview requires a question set, a scoring rubric, and the discipline to use both, then the first investment is process rather than software. Work samples and job knowledge tests also perform well and are cheap to construct from your own work. Personality questionnaires, which are the easiest thing to buy, sit lowest on this list as a basis for a reject decision.
A test is a selection procedure, with everything that follows
This is the part vendor content reduces to a compliance badge, and it deserves more space because the obligation sits with the employer rather than the vendor.
Any test used to make an employment decision is a selection procedure under federal anti-discrimination law, and the long-standing guidance on employment tests and selection procedures sets the framework. The core principle is adverse impact: if a test screens out members of a protected group at a substantially different rate than others, the employer must be able to show the test is job-related for the position and consistent with business necessity. Buying a validated instrument helps that argument and does not complete it, because validity evidence has to relate to the job you are actually hiring for.
What happens after the test
Assessment platforms end at a score. Two things follow that they do not handle, and both are where the value of a good hiring decision is either realised or lost.
The first is the decision record. If a test contributed to rejecting someone, the reasoning and the result are exactly what an employer needs to be able to produce later, and they usually live in the assessment platform rather than in the employment file. The second is what happens to the person you did hire: a well-assessed candidate who arrives to an improvised first week is a hiring win converted into a retention risk, and no assessment score protects against that.
How to choose a talent assessment tool
Five questions settle this, and the first decides whether to buy anything at all.
A closing note on sequencing. Before buying, run one role with a written question set, the same questions for every candidate, and a short task drawn from the actual work. Score it, and see whether the decisions felt better. If they did, you have found the cheapest improvement available and can decide whether software adds to it.
Frequently Asked Questions
What are talent assessment tools?
Platforms employers use to evaluate candidates or employees beyond a resume using structured tests. They cover four families: skills and job knowledge tests, cognitive ability tests, personality and behavioural questionnaires, and work samples or simulations. The category is also sold as talent assessment software, a talent assessment platform, pre-employment testing, and candidate assessment software, and the terms describe the same products.
What is the difference between talent assessment and a career test?
They point in opposite directions. Employer-side assessment is bought by a company to evaluate applicants, and results go to the hiring team. Career assessments are taken by an individual exploring what work might suit them, and results go to that person. Searching the phrase surfaces both, but the products are entirely separate: a career test has no employer dashboard, no candidate comparison, and no validation evidence intended to defend a hiring decision.
How much do talent assessment tools cost?
Published entry points run from roughly $19 a month at the lightest end to $5,000 or more a year for behavioural licences, and half the well-known platforms publish nothing. TestGorilla lists a free plan and a Core tier reported between about $1,620 and $1,700 a year, with Plus from around $4,800. HackerRank starts near $1,990. Criteria Corp, SHL, HireVue, and Wonderlic are quote-only. Most are annual subscriptions that do not scale down with hiring volume.
Do small businesses need talent assessment software?
At low hiring volume, usually not. An annual subscription near $1,650 across three hires is roughly $550 per hire, and across thirty it is $55. These platforms are built for the second case. For an employer making a handful of hires, a structured interview asking every candidate the same questions plus a short relevant work sample captures most of the benefit at no licence cost. The threshold is volume and role repetition rather than company size.
Is there free talent assessment software?
There are free tiers, though they are starting points. TestGorilla publishes a free plan with a small number of library tests, resume scoring, and one seat, which is enough to see how the platform works and not enough to run hiring across parallel roles. Beyond vendor tiers the genuinely free options are methodological: a structured interview guide and a work sample task cost nothing and compare well with paid instruments on current evidence.
Which assessment methods actually predict job performance?
The evidence base changed. A 2022 re-analysis by Sackett and colleagues found that long-standing validity estimates had been inflated by an inappropriate statistical correction, and that structured interviews rather than cognitive ability tests emerged as the predictor with the highest mean validity. The cognitive ability estimate was revised substantially downward. The revision is debated and rebuttals have been published, so treat it as live disagreement.
Are pre-employment tests legal?
Yes, and they are a regulated selection procedure. Federal anti-discrimination law applies to any test used in employment decisions, and guidance sets expectations around job-relatedness and adverse impact: if a test screens out a protected group at a substantially different rate, the employer must show it is job-related and consistent with business necessity. Reasonable accommodation obligations apply to the test itself, and some jurisdictions add bias audit and notice duties for automated tools. Confirm with employment counsel.
What is the difference between a talent assessment platform and an applicant tracking system?
They do different jobs at adjacent points. An applicant tracking system collects applications, manages the pipeline, and moves candidates to an offer. An assessment platform measures candidates against defined criteria and returns a score. Most assessment vendors integrate with common tracking systems so results appear against the candidate record, and several tracking systems include light screening questions that are not the same as a validated test. Neither replaces the other.