Reading Time
8 min

Skills Testing Tools ROI

Skills testing tools deliver provable ROI when they show three things at once. Predictive validity for the role you are filling, a measurable effect on time to hire and drop-off, and results from a customer that looks like you. Miss one and your business case is a guess, not a calculation.

That sounds like a low bar. It is not. Work through the assessment market asking every vendor for the number plus the sample it was calculated on, and you end up with a short list. This page shows you how to build that list yourself, and where we sit on it.

What ROI on a skills test actually means

ROI on a skills test is not the same as feeling good about quality. It is the gap between what a better selection decision earns you and what the measuring costs. That gap sits in four places.

Recruiter time. Hours not spent scanning CVs and chasing no-shows. This is the easiest to measure, which is why it is often the only thing measured, which is why most business cases come out too small.

Funnel drop-off. Candidates who quit mid-application. Every drop-off you prevent is a sourcing euro you do not spend twice.

Quality of hire. First-year performance and the odds that someone stays. This is the biggest item and the hardest, because you need performance data to see it.

Mishires. The hire who leaves within six months. One prevented mishire often pays for an annual licence, provided you know what that mishire actually costs instead of guessing.

Build your case on recruiter time alone and you buy a tool that saves time and says nothing about who you hire. That is efficiency, not ROI. The distinction is not academic, because it decides which tool you should want.

Five criteria to judge a tool on

  1. Predictive validity for your role. Ask for the correlation between test score and job performance, and for the sample it was calculated on. A tool with 2,500 tests and no validity reporting measures a lot and predicts nothing.
  2. Completion rate. The share of candidates who finish the test. A test with strong predictive power that drives half your candidates away costs you more than it returns.
  3. Adverse impact. The difference in selection ratio between groups. Ask for the analysis, not for the promise that the test is objective.
  4. Time to decision. Not how long the test takes, but how long it takes to get from application to go or no go, including the manual work the tool leaves you with.
  5. Reusability of the data. Can you put the scores next to performance data a year later? Without that step you can never establish ROI, only assume it.

What counts as evidence and what does not

What the research says about predictive power

The most recent large recalculation of selection methods is Sackett and colleagues, 2022. They corrected a systematic overestimate in the older Schmidt and Hunter figures from 1998. The resulting order.

Selection methodPredictive validity
Structured interview.42
Job knowledge test.40
Empirically keyed biodata.38
Work sample test.33
Cognitive ability test.31
Interests, measured as fit with the job.24

Two things stand out. The structured interview comes first, not the cognitive ability test. And interests measured as fit with a specific job do more than twice as well as interests as a general category, which sat at .10 in the older figures.

What that means for your choice. A tool that sells you a cognitive ability test alone buys you .31. A tool that pairs that test with a structured interview guide on the same competencies sits higher, because you are stacking two predictors instead of relying on one. That is the cheapest return in the whole hiring chain and the one most often skipped.

Four red flags in vendor claims

  • A validity figure with no sample and no role. A bare .58 means nothing if you do not know who it was measured on and against which performance measure.
  • Percentages with no denominator. "70% more accurate" is not a number, it is a sentence. More accurate than what, measured how?
  • Customer logos with no metric. A logo wall tells you someone bought it, not that it worked.
  • Test count as a quality argument. 400 tests, 2,500 tests, the count says nothing about whether one of them predicts your role.

These four apply to us too. Further down you can see which of our own numbers survive them and which do not.

Vendor overview, what each one proves in public

Below is only what was publicly available on 26 August 2026. "Not found" does not mean it does not exist, it means you have to ask for it before you count on it.

ToolWhat it measuresPublic evidenceWhere the ROI sits
Selection LabCognitive ability, personality, language, hard and soft skills, plus screening and scheduling in one flowCustomer cases with a metric per customer. No per-test validity coefficients in publicManual work removed and lower drop-off at volume
TestGorillaBroad library you assemble your own test set fromFact sheets per test with reliability and validity information, plus an explanation of their own validation studies. No coefficients or sample sizes on the science page itselfSpeed and volume, provided you pick the right tests
IxlyDutch psychometric publisher, ability and personalityTest manuals online plus COTAN ratings per testRoles where a bad hire is expensive
Aon Assessment, formerly cut-eInternational test batteries for volume and cross-border hiringNo public validity report found. Manuals go through the vendorInternational volume hiring on one norm
Codility and HackerRankWork samples for code and technical tasksCodility points to a technical manual and states plainly that it is making no claim about outcome tracking today. HackerRank publishes whitepapers per assessment typeInside engineering, where the task sits close to the job

One observation worth the detour. The only vendor in that table with an independent, publicly checkable review is Ixly, and the review is not spotless. COTAN, the Dutch national test review committee, rated Ixly's ACT general intelligence test good on theoretical foundation, test material and manual, adequate on reliability and construct validity, and inadequate on norms and criterion validity. Ixly wrote about it themselves, noting that criterion validity is very hard to establish in the field and that few Dutch intelligence tests reach an adequate rating there.

That is not a weakness of Ixly. That is what a real review looks like. Put it next to a vendor that prints "scientifically validated" on the homepage and publishes nothing, and you can see who has let themselves be checked.

How to calculate the ROI yourself

Fill in your own numbers. The formula.

Annual return = (mishires prevented x cost per mishire) + (recruiter hours saved x hourly cost) + (extra placements x margin per placement)

ROI = (annual return minus annual licence cost minus implementation cost) divided by (annual licence cost plus implementation cost)

The item that decides the outcome is almost always the mishire. It is also where the uncertainty sits, because hardly anyone knows that figure for their own organisation. So there is no industry benchmark below. Put your own figure in and watch what it does to the answer.

What you fill inWhere you get it
Hires per yearYour ATS
Leavers within six monthsYour HR system
Cost per mishireSourcing cost plus onboarding time plus lost output plus the hiring manager's hours
Applicants per yearYour ATS
Minutes saved per applicantMeasure this in the pilot, do not take it from a demo
Recruiter hourly costYour own payroll

Without the cost per mishire you cannot calculate the ROI of any tool. Establishing that number first is worth more than sitting through three demos. It takes an afternoon and it makes every vendor claim after it checkable.

What the EU AI Act does to your business case

Skills testing tools that use AI to evaluate candidates fall under Annex III of the EU AI Act and count as high risk. The Digital Omnibus on AI entered into force on 27 July 2026, after publication in the Official Journal on 24 July 2026. That moved the deadline for stand-alone Annex III systems from 2 August 2026 to 2 December 2027.

What does apply as of 2 August 2026 are the transparency obligations in Article 50. Candidates have to know they are dealing with an AI system, and synthetic output has to be marked in machine-readable form, with a transition period to 2 December 2026 for systems already on the market.

Two consequences for your business case. You have longer than you thought to get the documentation in order, and it has become a purchasing criterion. A vendor who cannot show you risk management and technical documentation now is passing that cost to you.

What we can and cannot prove

We do not publish per-test validity coefficients. Anyone who wants to compare at that level has to ask us for it, the same as with every other vendor in the table above. What we do have is results per customer. One of them measures quality of hire. The rest measure process.

The number that is about quality. At Dentons the quality-of-hire ratio rose by 13.5%. Of the candidates assessed, 32% were hired and 77% of those hires performed at or above average, measured with performance data after they started. Candidates scoring higher on morality, self-control and enthusiasm performed better.

The numbers that are about process. Carrefour saves 6 hours per recruiter per week. DPD saw the cost of the hiring process come out 15% lower. CoBuilders improved its placement ratio by 8%. Welten saves 6,700 minutes a month. FrieslandCampina hired 40 trainees within a month, out of 23,000 applications across 18 countries. At IG&H, 603 candidates were assessed, with 94% positive candidate feedback.

That second list says something about time and cost and nothing about who got hired. Which is the exact distinction this page is about, and it applies to us as well. If you want to know whether a tool makes your hires better, there is only one route. Keep the test scores, put them next to performance data a year later, and draw your conclusions then.

Frequently asked questions

Which skills testing tools actually deliver ROI in recruitment?

The ones that can show you three things. A validity figure with the sample attached, a completion rate from live use, and customer numbers with the metric attached. Which tool wins depends on your volume and on what a mishire costs you. For engineering roles that is a coding platform, for volume hiring an automated selection platform, for expensive key roles a psychometric publisher.

How do I calculate the ROI of an assessment tool?

Set the return from prevented mishires, saved recruiter hours and extra placements against licence and implementation costs. The item that decides the outcome is almost always the mishire. If you do not know that cost, you are not calculating, you are estimating.

Do skills tests predict job performance better than interviews?

Better than a CV, but not better than a structured interview. In the Sackett and colleagues recalculation from 2022 the structured interview reaches .42, a job knowledge test .40 and a cognitive ability test .31. The strongest combination is a test plus a structured interview on the same competencies.

How long before a skills testing tool pays for itself?

That depends on your hiring volume, not on the tool. Time savings show up immediately. Effect on quality of hire takes six to twelve months, because you need performance data from the new hires. Agree upfront which data you put side by side and when.

Does the EU AI Act apply to skills testing tools?

AI systems that evaluate candidates fall under Annex III and count as high risk. The Digital Omnibus on AI moved that deadline to 2 December 2027. The Article 50 transparency obligations have applied since 2 August 2026, so candidates already have to know an AI system is involved.

Comparing tools rather than building the business case? Start with the comparison of skill test tools, see how to improve your quality of hire, or read the Dentons case.

FAQ

Can game-based assessments promote diversity in the hiring process?

Yes, game-based assessments can support diversity by focusing on skills and behaviors rather than traditional criteria like résumés, which may contain unconscious biases. This gives candidates from diverse backgrounds a fairer chance to demonstrate their potential.

What is a game-based assessment?

A game-based assessment is a method that uses game mechanics to evaluate a candidate’s skills, competencies, and personality traits. While playing these games, candidates are assessed on aspects like problem-solving, cognitive ability, and behavior under pressure in an interactive way.

What are the advantages of game-based assessments?

Game-based assessments offer a more engaging and interactive experience for candidates, which can lead to a more positive perception of the hiring process—especially among certain groups. For employers, they provide deeper insights into both cognitive and behavioral traits, which traditional tests may miss. They also reduce the chance of socially desirable answers, as candidates tend to respond more authentically in a game environment.

How reliable are game-based assessments compared to traditional tests?

When well-designed, game-based assessments can be just as reliable—or even more reliable—than traditional tests. They assess a wide range of behaviors and cognitive abilities in a dynamic setting. However, the quality of these assessments varies greatly, so careful evaluation is essential.

How does a game-based assessment work?

Candidates participate in interactive games designed to measure specific skills and behaviors. Evaluation goes beyond just the final score—it also considers how the candidate makes decisions, handles challenges, and responds to different scenarios. These insights reveal underlying thought processes and behavioral patterns.

Are game-based assessments scientifically validated?

The main drawback is that many game-based assessments are relatively new and have not yet been extensively researched by independent academics. Providers often cite their own research, which is rarely externally validated. Without independent studies, the reliability of these assessments remains uncertain—something to keep in mind when selecting one.

How can game based assessments contribute to a better candidate experience

This can vary significantly by audience. The playful, interactive nature of game-based assessments can lower stress levels for some candidates compared to traditional tests. However, research shows that certain groups, especially those over 35, may find them more stressful. Men also tend to rate the experience more positively than women.

Can you practice game-based assessment?

You can familiarize yourself with the style of games used, but it’s difficult to "practice" for them in a traditional sense. These assessments are designed to measure natural reactions and authentic behavior, so repeated practice typically has less effect on performance than with traditional tests.

Will game-based assessments replace traditional tests in the future?

It’s likely that game-based assessments will become more common in hiring processes, but they probably won’t fully replace traditional tests. Both approaches have value and can complement each other depending on the role and the company’s needs.

How are the results of a game-based assessment analyzed?

Results are analyzed based on predefined criteria such as problem-solving ability, reaction time, and behavior under pressure. Advanced algorithms collect and interpret this data to provide a reliable, objective evaluation of a candidate’s strengths.

What kind of skills do game-based assessments measure?

They assess a wide range of abilities, including problem-solving, adaptability, decision-making under pressure, teamwork, and emotional intelligence. Depending on the design, they may also evaluate cognitive skills like memory, attention, and pattern recognition.

How long does a game-based assessment take?

Typically, these assessments last between 15 and 60 minutes, depending on the game’s complexity and the number of skills being tested. They’re usually shorter and more engaging than traditional assessments, making for a smoother candidate experience.

Are game-based assessments suitable for all roles?

They are especially effective for roles that require flexibility, creativity, problem-solving, and strong interpersonal skills. For highly technical or specialized roles, additional assessments may be needed to measure specific knowledge.

What’s the difference between a game-based and a gamified assessment?

A gamified assessment adds game-like elements (such as points or rewards) to a traditional test to increase engagement. A game-based assessment, on the other hand, is a standalone game designed specifically to evaluate certain competencies. The game itself is the primary evaluation tool, not just an enhancement.

FAQ

How can I improve my company’s retention rate?

The retention rate can be improved by investing in employee development and satisfaction. This includes offering training, career opportunities, and recognition for their contributions. A culture of open communication and attention to work-life balance can also contribute to higher retention. Additionally, offering competitive compensation and involving employees in decision-making can strengthen loyalty.

What are the benefits of growth opportunities for employee retention?

Growth opportunities can promote employee retention by giving staff a sense of direction and motivation. When they have the chance to learn and develop professionally within the company, they feel valued, which increases their loyalty. This can prevent them from leaving to seek better opportunities elsewhere. kunnen het behoud van personeel bevorderen door medewerkers een gevoel van richting en motivatie te geven. Wanneer zij de kans krijgen om te leren en zich professioneel te ontwikkelen binnen het bedrijf, voelen zij zich gewaardeerd, wat hun loyaliteit vergroot. Dit kan voorkomen dat ze vertrekken om elders betere kansen te zoeken.

What are the key factors that influence employee retention?

Key factors that influence employee retention include salary and benefits, opportunities for professional development, work-life balance, company culture, and the relationship with supervisors. Employees tend to stay longer when they feel valued, challenged, and supported in their work environment.

Why is employee retention so important for organizations?

Employee retention is important because it helps reduce recruitment and training costs for new employees, and it contributes to retaining knowledge and experience within the organization. High retention also ensures continuity within teams, leading to a more stable company culture, higher customer satisfaction, and improved business outcomes.

Which recruitment strategies help improve retention?

Recruitment strategies that can improve retention include identifying candidates who align with the company culture, using assessments to evaluate soft skills, and providing transparency about role expectations during the hiring process. Employees who feel connected to the organization and have clarity about their role are more likely to stay longer.

How can a good onboarding process contribute to higher retention?

An effective onboarding process can contribute to higher retention by helping new employees quickly adapt to their role, the company culture, and expectations. By providing support and clear information from the start, their engagement is increased, and the likelihood of them leaving early due to feelings of being overwhelmed or lacking guidance is reduced.

What is the role of company culture in retaining employees?

Company culture plays a crucial role in employee retention. When employees feel heard, valued, and connected to the values and norms of the company, they are more likely to stay. A positive culture that fosters collaboration, respect, and personal growth can significantly enhance employee motivation and satisfaction.

How can leadership and management style influence retention?

Leadership and management style have a significant impact on retention. Leaders who inspire, support, and coach their team can increase employee engagement and satisfaction. Offering autonomy and trust can lead to higher loyalty, while inefficient or negative management styles can contribute to dissatisfaction and increased employee turnover.

What is the importance of recognition and rewards for employee retention?

Recognition and rewards play an important role in employee retention by showing staff that their work is valued. This can increase their motivation and loyalty. In addition to financial rewards, compliments, promotions, and other forms of recognition can also contribute to satisfaction and retaining employees.

What role does work-life balance play in improving retention?

A balanced work-life balance plays an important role in increasing retention. By reducing stress and improving job satisfaction, employees are more likely to stay with the company. Initiatives such as flexible working hours, remote work options, and respect for personal time can contribute to this balance.

What does increasing retention mean within a company?

Increasing retention within a company means implementing strategies to keep employees with the organization for longer. This can be achieved by improving job satisfaction, offering growth opportunities, and fostering a positive and supportive company culture.

How do I measure the success of my retention strategy?

The success of a retention strategy can be measured by tracking retention rates and turnover rates, and by gaining insights from exit interviews. Additionally, employee satisfaction surveys and feedback from performance evaluations can provide valuable information about the effectiveness of the strategies applied.

What are the costs of a low retention rate?

A low retention rate can bring significant costs, such as increased expenses for recruiting and training new employees. Furthermore, the loss of experienced staff can lead to lower productivity, reduced knowledge transfer, and a negative impact on company culture.

How can I increase employee engagement?

To increase employee engagement, involve them in decision-making processes, regularly ask for their feedback, and recognize their contributions. Offering development opportunities and maintaining transparent communication can also contribute to greater engagement.

How can technology help improve employee retention?

Technology can be a tool for improving employee retention by facilitating communication, feedback, and development. By using online platforms for training, recognition, and evaluation, companies can create a more engaged and satisfied workforce.

FAQ

How long does it take to complete the tool?

Less than 10 minutes. You’ll answer 30 guided questions and get a summary of what to look for in your next assessment platform.

Can this checklist help me compare assessment providers?

Yes. By clarifying what matters most to your team, it makes comparing providers' features, pricing, and strengths much easier and more strategic.

How can I use this checklist if I’m not doing a formal RFI?

It’s equally valuable for internal evaluations, exploring new tools, or improving your current hiring process even if you’re not issuing an RFI or RFQ.

What should I look for in a modern assessment tool?

Prioritize platforms with user-friendly design, mobile compatibility, strong analytics, ATS integrations, and inclusive features like neurodiversity support.

What types of assessments should I consider in 2025?

Leading tools combine cognitive testing, situational judgment tests (SJTs), behavior assessments, and predictive AI to evaluate candidates more holistically.

Who should use an assessment checklist?

HR professionals, hiring managers, and procurement teams evaluating pre-selection solutions, especially those comparing AI-powered or compliance-driven assessment platforms.

How does this checklist help with RFIs and RFQs for assessments?

The checklist helps you define your exact requirements so you can confidently draft or respond to Requests for Information (RFI) or Requests for Quotation (RFQ) for assessment tools.

What is an assessment tool in hiring?

An assessment tool evaluates candidates’ skills, behaviors, and fit during the recruitment process. It helps improve hiring decisions and streamline pre-selection.

Game-based assessment packs

← Our Blog

Skills Testing Tools ROI

Five criteria, a vendor overview built on public evidence only, and a model you fill in with your own numbers.
Read time: Approx
8 min

Skills testing tools deliver provable ROI when they show three things at once. Predictive validity for the role you are filling, a measurable effect on time to hire and drop-off, and results from a customer that looks like you. Miss one and your business case is a guess, not a calculation.

That sounds like a low bar. It is not. Work through the assessment market asking every vendor for the number plus the sample it was calculated on, and you end up with a short list. This page shows you how to build that list yourself, and where we sit on it.

What ROI on a skills test actually means

ROI on a skills test is not the same as feeling good about quality. It is the gap between what a better selection decision earns you and what the measuring costs. That gap sits in four places.

Recruiter time. Hours not spent scanning CVs and chasing no-shows. This is the easiest to measure, which is why it is often the only thing measured, which is why most business cases come out too small.

Funnel drop-off. Candidates who quit mid-application. Every drop-off you prevent is a sourcing euro you do not spend twice.

Quality of hire. First-year performance and the odds that someone stays. This is the biggest item and the hardest, because you need performance data to see it.

Mishires. The hire who leaves within six months. One prevented mishire often pays for an annual licence, provided you know what that mishire actually costs instead of guessing.

Build your case on recruiter time alone and you buy a tool that saves time and says nothing about who you hire. That is efficiency, not ROI. The distinction is not academic, because it decides which tool you should want.

Five criteria to judge a tool on

  1. Predictive validity for your role. Ask for the correlation between test score and job performance, and for the sample it was calculated on. A tool with 2,500 tests and no validity reporting measures a lot and predicts nothing.
  2. Completion rate. The share of candidates who finish the test. A test with strong predictive power that drives half your candidates away costs you more than it returns.
  3. Adverse impact. The difference in selection ratio between groups. Ask for the analysis, not for the promise that the test is objective.
  4. Time to decision. Not how long the test takes, but how long it takes to get from application to go or no go, including the manual work the tool leaves you with.
  5. Reusability of the data. Can you put the scores next to performance data a year later? Without that step you can never establish ROI, only assume it.

What counts as evidence and what does not

What the research says about predictive power

The most recent large recalculation of selection methods is Sackett and colleagues, 2022. They corrected a systematic overestimate in the older Schmidt and Hunter figures from 1998. The resulting order.

Selection methodPredictive validity
Structured interview.42
Job knowledge test.40
Empirically keyed biodata.38
Work sample test.33
Cognitive ability test.31
Interests, measured as fit with the job.24

Two things stand out. The structured interview comes first, not the cognitive ability test. And interests measured as fit with a specific job do more than twice as well as interests as a general category, which sat at .10 in the older figures.

What that means for your choice. A tool that sells you a cognitive ability test alone buys you .31. A tool that pairs that test with a structured interview guide on the same competencies sits higher, because you are stacking two predictors instead of relying on one. That is the cheapest return in the whole hiring chain and the one most often skipped.

Four red flags in vendor claims

  • A validity figure with no sample and no role. A bare .58 means nothing if you do not know who it was measured on and against which performance measure.
  • Percentages with no denominator. "70% more accurate" is not a number, it is a sentence. More accurate than what, measured how?
  • Customer logos with no metric. A logo wall tells you someone bought it, not that it worked.
  • Test count as a quality argument. 400 tests, 2,500 tests, the count says nothing about whether one of them predicts your role.

These four apply to us too. Further down you can see which of our own numbers survive them and which do not.

Vendor overview, what each one proves in public

Below is only what was publicly available on 26 August 2026. "Not found" does not mean it does not exist, it means you have to ask for it before you count on it.

ToolWhat it measuresPublic evidenceWhere the ROI sits
Selection LabCognitive ability, personality, language, hard and soft skills, plus screening and scheduling in one flowCustomer cases with a metric per customer. No per-test validity coefficients in publicManual work removed and lower drop-off at volume
TestGorillaBroad library you assemble your own test set fromFact sheets per test with reliability and validity information, plus an explanation of their own validation studies. No coefficients or sample sizes on the science page itselfSpeed and volume, provided you pick the right tests
IxlyDutch psychometric publisher, ability and personalityTest manuals online plus COTAN ratings per testRoles where a bad hire is expensive
Aon Assessment, formerly cut-eInternational test batteries for volume and cross-border hiringNo public validity report found. Manuals go through the vendorInternational volume hiring on one norm
Codility and HackerRankWork samples for code and technical tasksCodility points to a technical manual and states plainly that it is making no claim about outcome tracking today. HackerRank publishes whitepapers per assessment typeInside engineering, where the task sits close to the job

One observation worth the detour. The only vendor in that table with an independent, publicly checkable review is Ixly, and the review is not spotless. COTAN, the Dutch national test review committee, rated Ixly's ACT general intelligence test good on theoretical foundation, test material and manual, adequate on reliability and construct validity, and inadequate on norms and criterion validity. Ixly wrote about it themselves, noting that criterion validity is very hard to establish in the field and that few Dutch intelligence tests reach an adequate rating there.

That is not a weakness of Ixly. That is what a real review looks like. Put it next to a vendor that prints "scientifically validated" on the homepage and publishes nothing, and you can see who has let themselves be checked.

How to calculate the ROI yourself

Fill in your own numbers. The formula.

Annual return = (mishires prevented x cost per mishire) + (recruiter hours saved x hourly cost) + (extra placements x margin per placement)

ROI = (annual return minus annual licence cost minus implementation cost) divided by (annual licence cost plus implementation cost)

The item that decides the outcome is almost always the mishire. It is also where the uncertainty sits, because hardly anyone knows that figure for their own organisation. So there is no industry benchmark below. Put your own figure in and watch what it does to the answer.

What you fill inWhere you get it
Hires per yearYour ATS
Leavers within six monthsYour HR system
Cost per mishireSourcing cost plus onboarding time plus lost output plus the hiring manager's hours
Applicants per yearYour ATS
Minutes saved per applicantMeasure this in the pilot, do not take it from a demo
Recruiter hourly costYour own payroll

Without the cost per mishire you cannot calculate the ROI of any tool. Establishing that number first is worth more than sitting through three demos. It takes an afternoon and it makes every vendor claim after it checkable.

What the EU AI Act does to your business case

Skills testing tools that use AI to evaluate candidates fall under Annex III of the EU AI Act and count as high risk. The Digital Omnibus on AI entered into force on 27 July 2026, after publication in the Official Journal on 24 July 2026. That moved the deadline for stand-alone Annex III systems from 2 August 2026 to 2 December 2027.

What does apply as of 2 August 2026 are the transparency obligations in Article 50. Candidates have to know they are dealing with an AI system, and synthetic output has to be marked in machine-readable form, with a transition period to 2 December 2026 for systems already on the market.

Two consequences for your business case. You have longer than you thought to get the documentation in order, and it has become a purchasing criterion. A vendor who cannot show you risk management and technical documentation now is passing that cost to you.

What we can and cannot prove

We do not publish per-test validity coefficients. Anyone who wants to compare at that level has to ask us for it, the same as with every other vendor in the table above. What we do have is results per customer. One of them measures quality of hire. The rest measure process.

The number that is about quality. At Dentons the quality-of-hire ratio rose by 13.5%. Of the candidates assessed, 32% were hired and 77% of those hires performed at or above average, measured with performance data after they started. Candidates scoring higher on morality, self-control and enthusiasm performed better.

The numbers that are about process. Carrefour saves 6 hours per recruiter per week. DPD saw the cost of the hiring process come out 15% lower. CoBuilders improved its placement ratio by 8%. Welten saves 6,700 minutes a month. FrieslandCampina hired 40 trainees within a month, out of 23,000 applications across 18 countries. At IG&H, 603 candidates were assessed, with 94% positive candidate feedback.

That second list says something about time and cost and nothing about who got hired. Which is the exact distinction this page is about, and it applies to us as well. If you want to know whether a tool makes your hires better, there is only one route. Keep the test scores, put them next to performance data a year later, and draw your conclusions then.

Frequently asked questions

Which skills testing tools actually deliver ROI in recruitment?

The ones that can show you three things. A validity figure with the sample attached, a completion rate from live use, and customer numbers with the metric attached. Which tool wins depends on your volume and on what a mishire costs you. For engineering roles that is a coding platform, for volume hiring an automated selection platform, for expensive key roles a psychometric publisher.

How do I calculate the ROI of an assessment tool?

Set the return from prevented mishires, saved recruiter hours and extra placements against licence and implementation costs. The item that decides the outcome is almost always the mishire. If you do not know that cost, you are not calculating, you are estimating.

Do skills tests predict job performance better than interviews?

Better than a CV, but not better than a structured interview. In the Sackett and colleagues recalculation from 2022 the structured interview reaches .42, a job knowledge test .40 and a cognitive ability test .31. The strongest combination is a test plus a structured interview on the same competencies.

How long before a skills testing tool pays for itself?

That depends on your hiring volume, not on the tool. Time savings show up immediately. Effect on quality of hire takes six to twelve months, because you need performance data from the new hires. Agree upfront which data you put side by side and when.

Does the EU AI Act apply to skills testing tools?

AI systems that evaluate candidates fall under Annex III and count as high risk. The Digital Omnibus on AI moved that deadline to 2 December 2027. The Article 50 transparency obligations have applied since 2 August 2026, so candidates already have to know an AI system is involved.

Comparing tools rather than building the business case? Start with the comparison of skill test tools, see how to improve your quality of hire, or read the Dentons case.