The Ultimate Guide to Hiring Assessments: Types, Costs, and How to Choose

Hiring Assessment Guide
Search

Try SmoothHiring for free for 14 days

See SmoothHiring in action. Know our features and get insights on how our friendly software helps you with successful hiring. Learn how data and predictive analytics help in hiring the right candidate.

The predictive hiring platform for finding the best employees

Know how predictive hiring helps you find the right people for the right job and increase employee productivity.

Author: SmoothHiring Team

Not finding the best employees?

Schedule a demo to learn how SmoothHiring will help you find the best fit for the job using predictive analytics.

Most hiring teams land in one of two places with assessments: they skip them entirely and hope resumes and interviews are enough, or they open a vendor comparison page, see 40 tools with different pricing models and vague feature lists, and close the tab. Neither gets you a better hire.

This guide covers the three questions that actually matter. What types of hiring assessments exist and what does each one measure. What they cost, from free tiers to enterprise contracts. And how to pick the right one for your hiring volume, your roles, and your budget, without needing a psychology degree to read the vendor’s spec sheet.

It’s written for HR managers, recruiters, and small to mid-size business owners who are building or rebuilding a hiring process, not for candidates preparing to take a test.

A hiring assessment is any standardized tool used to measure a candidate’s ability, traits, or likely job performance before making an offer. That covers a wide range: a 10-minute typing test, a 45-minute personality inventory, a coding challenge, a scenario-based judgment test. What they share is standardization. Every candidate gets the same questions, scored the same way, which is what separates an assessment from an unstructured “gut feeling” interview.

Adoption has grown quickly. The global employee screening market, which includes background checks alongside pre-employment testing, is projected to reach $5.46 billion by 2026. Most of that growth is defensive: a bad hire is expensive, and assessments are one of the few hiring tools with published, decades-old research behind their accuracy.

Every hiring assessment on the market is a variation on seven core types. Some vendors bundle several into one platform; others specialize in a single type.

Cognitive tests measure reasoning, problem-solving, and learning speed rather than a specific job skill. They’re built on decades of industrial-organizational psychology research and are among the most heavily studied assessment types that exist. General mental ability, the academic term for this category, carries a validity coefficient around 0.51 in Schmidt and Hunter’s classic 1998 meta-analysis, meaning it’s one of the strongest single predictors of job performance ever measured. A 2021 reanalysis by Sackett and colleagues applied stricter statistical corrections and put the figure closer to 0.31, still meaningful, though the exact number remains debated among researchers. Either way, cognitive tests work better for roles with real complexity and judgment calls than for highly routine work.

Personality assessments measure stable traits, like the Big Five model’s openness, conscientiousness, extraversion, agreeableness, and emotional stability, that hold steady across situations and over time. Conscientiousness in particular has a long, consistent research record as a performance predictor across nearly every occupation studied. This is a large enough topic that it deserves its own explanation; see our full breakdown of behavioral, psychometric, and personality assessments.

Behavioral Assessments, DISC and the Predictive Index being the two most common, measure how someone tends to act and communicate in specific situations rather than fixed internal traits. They overlap heavily with personality assessments in practice, to the point that the same tool often gets marketed under both labels depending on the vendor. Our behavioral vs. psychometric vs. personality guide covers this overlap in detail.

Skills tests measure whether a candidate can actually do a specific, job-relevant task: writing a SQL query, formatting an Excel model, drafting a customer email, operating a piece of equipment. Job knowledge tests are a close cousin, testing what a candidate knows rather than what they can produce live. In Schmidt and Hunter’s meta-analysis, job knowledge tests carry a validity coefficient of 0.48, among the higher scores of any single method, and they’re generally the easiest type for a hiring manager to build without outside help, since the questions come directly from the job itself.

A situational judgment test, usually shortened to SJT, presents a realistic workplace scenario and asks the candidate to choose or rank how they’d respond. SJTs are built to capture soft skills, judgment, teamwork, adaptability, that don’t show up cleanly in a resume. Research published in the Journal of Applied Psychology by Lievens and Sackett found that SJTs add real incremental validity on top of cognitive and job knowledge tests, meaning they measure something those tests miss. The tradeoff: a well-built SJT requires scenarios specific to your workplace, since a generic scenario bank tends to measure agreement with generic corporate norms rather than fit for your team.

Integrity tests measure honesty, reliability, and how a candidate is likely to handle ethical gray areas on the job, things like handling cash, following safety procedures, or reporting a colleague’s mistake. Some are overt, asking directly about attitudes toward theft or rule-breaking; others are indirect and personality-adjacent. In Schmidt and Hunter’s data, integrity tests carry a standalone validity of 0.41, but combining an integrity test with a cognitive test produced the single highest combined validity of any pairing studied, 0.65. One clarification worth making explicitly: paper-and-pencil integrity tests are not the same as polygraph exams. The Employee Polygraph Protection Act of 1988 bars most private employers from using lie-detector tests in hiring; it does not restrict written integrity assessments.

Work samples ask a candidate to complete a task that mirrors real job content: a mock sales call, a short coding exercise, a customer service roleplay. They’re the highest single predictor of job performance in Schmidt and Hunter’s meta-analysis at 0.54, edging out even cognitive tests, largely because they measure the actual behavior in question rather than a proxy for it. The limitation is practical rather than statistical: work samples only work for candidates who already have some familiarity with the task, which makes them a poor fit for entry-level or career-change hiring.

The table below summarizes Schmidt and Hunter’s (1998) foundational validity coefficients, still the most-cited benchmark in the field, alongside how much each method adds when paired with a cognitive ability test.

Assessment TypeValidity AloneCombined with Cognitive Test
Work samples0.540.63
Structured interviews0.510.63
Cognitive ability tests0.51N/A
Job knowledge tests0.480.58
Integrity tests0.410.65
Unstructured interviews0.38N/A
Years of job experience0.18N/A

A 2021 reanalysis by Sackett and colleagues revised several of these figures downward under stricter statistical corrections (cognitive ability closer to 0.31, for instance) while keeping the same overall ranking. The practical lesson holds under either analysis: combining two methods consistently outperforms relying on one, and the most popular hiring method of all, the unstructured interview, is one of the weakest predictors on this list.

Assessment pricing falls into four general models. Which one makes sense depends almost entirely on how often you’re hiring.

Free and freemium tools. Several platforms offer a limited free tier, typically a handful of tests per month with basic reporting. This works for occasional, low-volume hiring but usually caps out fast once you’re filling more than one or two roles at a time.

Pay-per-candidate. You buy test credits and pay per candidate tested, generally somewhere in the $19 to $40 range per test depending on the vendor and test type, with no ongoing subscription. This model suits businesses that hire in bursts rather than continuously, since you’re only paying when you’re actually screening someone.

Monthly or annual subscriptions. Most mid-market platforms price by seat count and candidate volume, with tiers commonly starting around $75 to $100 a month for small teams and scaling into the low thousands per month for larger hiring volumes. This is the most common model for companies hiring regularly across multiple roles.

Enterprise contracts. Large organizations with high hiring volume or custom validation needs often move to negotiated annual contracts, which can run $40,000 or more a year depending on scale, integrations, and support level. This tier usually includes dedicated validation studies and adverse-impact analysis as part of the package.

The cost of skipping assessments entirely. This is the number that puts the above in perspective. The U.S. Department of Labor’s long-standing estimate puts the cost of a bad hire at a minimum of 30% of that employee’s first-year salary. Separately, CareerBuilder’s employer surveys put the average financial hit at roughly $17,000 for an entry- to mid-level bad hire, climbing to $240,000 or more for a senior-level one. That’s a different number from the average cost to fill a role in the first place, which SHRM puts at roughly $4,700 to $4,800. In other words: hiring already costs money whether or not you assess candidates. Assessment spend is a small, fixed cost measured against a bad-hire cost that’s large and unpredictable.

Start with hiring volume, not brand reputation. A company filling one or two roles a quarter rarely needs the same tool as one running twenty req’s a month. Low-volume hiring usually fits a free tier or pay-per-candidate model; frequent hiring usually earns back a subscription’s cost quickly through better time-to-fill and lower turnover.

Technical and specialist roles call for skills or job knowledge tests over generic personality inventories; a coding test tells you more about a developer candidate than a Big Five score does. Customer-facing and team-heavy roles benefit from behavioral or situational judgment tests, since communication style and real-time judgment matter more than raw cognitive horsepower. Leadership and management roles tend to benefit from combining a personality assessment with a situational judgment test, since both people skills and judgment under ambiguity matter. Roles involving cash handling, safety compliance, or sensitive data are the clearest case for an integrity assessment.

1. Do you have published validity data, ideally from a study specific to a role similar to mine, or only general marketing claims?

2. What adverse-impact data do you have across race, gender, and age groups for this specific test?

3. Does the assessment integrate with the ATS I already use, or will scores live in a separate system?

4. What’s the average completion time, and has that been tested against real candidate drop-off rates?

    Watch for these red flags. A vendor that can’t produce any validity data beyond testimonials. A single generic test marketed as a fit for every role and every industry. Any product built around a polygraph or “lie detector,” which is illegal for most private employers to use in hiring under federal law. And pricing that’s opaque until you’re on a sales call, which usually means the per-candidate cost is higher than competitors who post pricing openly.

    Using one generic test for every role. A cognitive test built for a management role and reused for a warehouse position is measuring the wrong thing for at least one of those jobs. Validity comes from matching the test to the actual demands of the role, not from picking one assessment and standardizing on it company-wide.

    Skipping the validity and adverse-impact review. It’s tempting to pick a test because a competitor uses it or because the demo looked polished. A test with no published validity data and no adverse-impact analysis is a liability wearing the costume of a solution, especially once a rejected candidate asks why.

    Stacking too many assessments in one process. Each additional test after the second one adds candidate drop-off without a matching gain in predictive accuracy. A 90-minute battery of five tests filters out strong candidates who simply don’t have that much free time, not just weak ones.

    Treating the assessment score as the whole decision. The strongest predictive validity in the research comes from combining methods, not from any single score. A cognitive test paired with a structured interview outperforms either one alone. An assessment score should narrow the field and inform the interview, not replace it.

    SmoothHiring’s pre-employment assessment library covers skills tests, cognitive and aptitude tests, situational judgment tests, and integrity checks inside the same ATS used for the rest of the hiring pipeline, so assessment scores sit next to resumes and interview notes rather than in a separate tool. Plans and assessment access are laid out on the pricing page for teams sizing up which tier fits their hiring volume.

    Frequently Asked Questions

    A personality test is one specific type of hiring assessment. “Hiring Assessment” is the broader category that also includes cognitive tests, skills tests, situational judgment tests, integrity tests, and work samples. A given hiring process might use one type or several in combination.

    It depends on the model. Free tiers exist for light use, pay-per-candidate tools run roughly $19 to $40 per test, subscription platforms start around $75 to $100 a month for small teams, and enterprise contracts can run $40,000 or more a year for large-scale hiring with custom validation.

    No. They’re optional, though widely used. What’s legally required, under the EEOC’s Uniform Guidelines on Employee Selection Procedures, is that any test used in hiring be job-related and applied consistently, especially if it produces a disproportionate pass or fail rate between protected groups. Talk with an employment attorney before rolling out a new assessment at scale.

    Yes. Pay-per-candidate pricing and free tiers exist specifically for lower-volume hiring, and the per-test cost is small compared with the cost of a bad hire, which the U.S. Department of Labor estimates at a minimum of 30% of that employee’s first-year salary.

    Most individual assessments run 10 to 45 minutes depending on type. Skills and cognitive tests tend toward the shorter end; personality and situational judgment tests run a bit longer. Stacking more than two or three assessments in one hiring process tends to increase candidate drop-off.

    A skills test measures whether a candidate can do a specific, job-relevant task right now, like writing a query or formatting a spreadsheet. A cognitive ability test measures general reasoning and problem-solving that predicts how quickly someone will learn a new task, regardless of whether they’ve done it before.

    The research supports it indirectly rather than by measuring turnover directly: assessments with strong validity data predict job performance more accurately than resumes or unstructured interviews, and better-performing hires are less likely to be a poor fit that leads to early turnover. Results vary by how well the specific assessment was validated for the specific role.

    No. Paper-and-pencil integrity tests are legal for most private employers to use in hiring. Polygraph, or lie-detector, tests are restricted for most private employers under the Employee Polygraph Protection Act of 1988. A vendor offering polygraph-based screening for a standard hiring process is a legal red flag.

    Most well-designed processes use one or two, rarely three. Each additional assessment adds candidate drop-off risk without a proportional increase in predictive accuracy, especially once you’ve already combined two complementary methods, like a skills test plus a situational judgment test.

    Yes, if the assessment isn’t job-related or produces adverse impact without a validity defense. The risk isn’t in using assessments, it’s in using an unvalidated one inconsistently. Vendors with published validity and adverse-impact data, applied the same way to every candidate for a role, are the lower-risk path.

    Hiring assessments aren’t one product, they’re seven overlapping categories with very different price points and very different jobs to do. Cognitive tests and work samples carry the strongest individual research record. Skills and job knowledge tests are the easiest to build in-house. Personality, behavioral, situational judgment, and integrity tools each add something the others miss, and combining two well-chosen methods consistently beats relying on any single one.

    The right starting point isn’t the assessment with the best marketing page. It’s an honest look at your hiring volume, your riskiest roles, and your budget, matched against a vendor who can actually show you validity data instead of just a demo.

    Let us provide you with a detailed tour

    Tell us about your problems, and we will present you with the most intriguing choices?

    Hiring Assessment Guide

    Get Started Today

    Let us profile your top performers and put together comprehensive WHY data that you can use immediately to hire. Speak with one of our representatives to learn how to save time and money while making dramatically better people decisions:

    Get Started Today

    Let us profile your top performers and put together comprehensive WHY data that you can use immediately to hire. Speak with one of our representatives to learn how to save time and money while making dramatically better people decisions: