Home AI Foundations for HR Teams A Practical Checklist for Vetting Any AI Feature Your HR Vendor Just Shipped

A Practical Checklist for Vetting Any AI Feature Your HR Vendor Just Shipped

Your ATS pushed an update with a new 'AI assistant' toggle — here's how to evaluate it before you turn it on for real candidates.

By Theo Brandt, a former talent-acquisition lead who now trains HR teams on AI · Published 9 July 2026 · 8 min read · Reviewed against our editorial standards

ADVERTISEMENT

The pattern is familiar by now. You log into Greenhouse, iCIMS, Workday, SmartRecruiters, or Paradox on a Tuesday and there's a new panel: AI-generated interview questions, an "assistant" that drafts candidate rejections, a summarizer that condenses a resume into three bullets, a score you didn't ask for. It shipped in a release note. Nobody asked whether you wanted it, and the toggle may already be on.

These features are often genuinely useful. But "it came from a vendor we already trust" is not the same as "it's safe to point at live candidates." Here's the checklist I walk HR teams through before they flip one of these on for real.

1. What does it actually do — and does it decide anything?

Separate features into two buckets. Assistive features help a human do their own work faster: draft this outreach email, summarize this call, suggest interview questions. Decisional features influence who advances: score this candidate, rank this pool, recommend a shortlist.

The bar is completely different. An assistive drafting tool that writes a mediocre email costs you a rewrite. A decisional tool that quietly filters people carries legal and ethical weight and may fall under laws like NYC's Local Law 144 or the EU AI Act's high-risk category. Get precise about which bucket you're in, because vendors love to describe decisional features in assistive language. "Helps you prioritize" often means "ranks candidates."

2. What's the model, and where does the data go?

Ask three concrete questions and get answers in writing:

ADVERTISEMENT

3. Ask specifically about bias and testing

For anything decisional, ask what the vendor has done to test for disparate impact, and ask for documentation. Watch how they respond. A serious vendor has an answer — a bias audit, fairness testing, a methodology. A vendor that gets defensive, waves at "our AI is unbiased," or points only to the fact that the model "can't see race" is telling you something. Every model can learn proxies. "It doesn't see the protected attribute" is not a bias defense; it's a misunderstanding.

4. Can you see the reasoning, or just the output?

A score with no explanation is hard to defend and hard to trust. Prefer features that show their work: which parts of a resume drove a summary, why a candidate was flagged, what the recommendation is based on. If a candidate or a regulator asks "why was I screened out," "the software said so" is not an answer you want to give. Explainability isn't a nice-to-have here; it's what makes the feature auditable.

ADVERTISEMENT

5. Who is accountable when it's wrong — and can you turn it off?

Two practical questions:

6. Test it on data where you already know the answer

This is the step teams skip, and it's the most valuable one. Before you trust a new feature, run it against a set of past cases where you know the outcome.

An afternoon of this tells you more than any vendor deck. You're not looking for perfection — you're calibrating how much to trust it and where it fails.

ADVERTISEMENT

7. Document the decision

Once you've vetted it, write down what you found, who signed off, and on what date. When you turned it on, what testing you did, what the vendor told you about bias and data. This record costs ten minutes and is exactly what you'll want if a candidate complaint or a regulator's question lands eighteen months later. It also forces the decision to be a decision, rather than a toggle someone left on by default.

The one-line version

Before you point a new AI feature at real candidates, know what it decides, where the data goes, how it was tested for bias, whether you can see its reasoning, whether you can turn it off, and how it performs on cases you already understand — then write down that you checked. New features are guilty until tested, even from vendors you like.

This is general guidance for HR practitioners, not legal advice. Whether a specific feature triggers legal obligations depends on your jurisdiction and facts — involve qualified counsel and your privacy team before deploying decisional AI in hiring.

vendor-vettingprocurementchecklistgovernance

Put this into practice

Paste any text to estimate how many tokens it uses, and see what that text would cost to send to each major model.

Open the Token Estimator →

A note on shelf life. AI products change fast. This guide deliberately focuses on the parts that stay true — how to judge a tool, what the trade-offs are — rather than ranking products that will have changed by the time you read it. Prices and feature claims should always be checked against the provider before you rely on them.

ADVERTISEMENT