An AI ATS is an applicant tracking system with machine-learning features built into the hiring workflow: parsing resumes into structured fields, ranking applicants against a role, summarising interviews, and drafting job posts or outreach. The tracking layer is the product. The AI is a layer on top of it, and its quality varies far more than the demos suggest.
Free 1-user plan · No credit card · Talk to a real recruiter
Strip the marketing and you find five distinct things vendors call AI, and they are not equally hard to build. Parsing turns a resume into structured fields. Semantic search finds people in your database by meaning rather than exact keywords. Ranking scores an applicant pool against a role. Generation drafts job posts, screening questions, rejection notes and outreach. Summarisation condenses an interview recording or a long profile into a paragraph. A standard applicant tracking system already handles the storage, stages, notifications and reporting underneath all five, and those are described in the ATS capability guide. Ask which of the five a vendor ships today and which sit on a roadmap slide. Rules-based automation gets relabelled as AI constantly: a filter that removes anyone without a required licence is a rule, not a model, and it is often the more useful of the two.
Sort every claim into two piles. Checkable: hand the vendor an awkward resume — a scanned PDF, a two-column design template, a career gap, a college the model has probably never encountered — and watch the parse happen in front of you. Ask them to rank twenty of your own applications live. Ask why the top candidate is top, and notice whether the answer comes from the product or from the salesperson. Ask for a generated job post for a role in your industry, then read it properly. Unverifiable: accuracy percentages, model names, corpus size, and anything phrased as trained on millions of profiles. You cannot test any of those from your side of the screen, so give them no weight in your scoring sheet. Push every claim into the first pile or drop it. A vendor who will not run your data during a demo has answered you already.
Want this priced against your own hiring volume?
Free forever for 1 user · no credit card
Backtest it. Pick two or three requisitions you closed in the last year where you remember the outcome and the applicant pool was large enough to sort. Load those applicants into the trial account with the outcome stripped out. Run the ranking. Then look for the people you actually interviewed and the person you hired. If your eventual hire lands in the bottom half, the model is not reading that role the way you do. If they surface near the top, that is real signal rather than a testimonial. Do it for a role you filled easily and one you struggled with, because the difficult one tells you far more. Then have a recruiter order the same pool blind and compare the two lists. Disagreement is not automatic failure, but each disagreement is a conversation worth having before money changes hands, not after.
A thin layer is a generic model wired to a text box, sold as though it were trained for hiring. Six questions surface it quickly, and none of them require you to understand machine learning. The pattern to watch for is a vendor who answers each one with a benefit rather than a mechanism. Ask what happens on day one, before you have given the system a single hiring decision to learn from — a product that needs months of your data to be useful should say so plainly. Ask whether a score changes when the same resume is uploaded twice. Ask who can see a score, and whether a hiring manager sees the same number the recruiter does. Ranking that nobody can explain, switch off, or override is not a feature you are buying. It is a liability you are inheriting.
Because a hiring decision has to be defensible to a candidate, a manager and possibly a regulator, and none of them accept the model said so. Explainability is operational before it is ethical: a recruiter who cannot see why someone ranked low will either trust the order blindly or ignore it entirely, and both waste the feature. Look for reasons attached to each score, an audit trail of who overrode what, the ability to hide name, gender markers and institution during a first pass, and a recorded human decision at every rejection point. Ask what the vendor can demonstrate rather than what they can certify. Rules on automated decision-making and candidate data differ by country and by state, and they change — confirm your obligations with your own legal or compliance advisor rather than relying on a vendor's summary of them.
A trial is an experiment, so decide what would make you say no before it starts. Run one live requisition end to end and one closed requisition as a backtest. Put at least three people in it: whoever screens, whoever interviews, and whoever will answer for the decision later. Measure four things — hours to a first shortlist, how often recruiters override the ranking, how often the override was right on review, and whether any candidate asked a question the team could not answer. Import a realistic volume rather than a tidy sample, including duplicates and half-finished applications, because that is what production looks like. Wire scoring into your pipeline and your recruitment reporting instead of running it in a side tool nobody opens. At day thirty, compare those four numbers against how the same roles ran before you started, and write the decision down.
It struggles wherever there is nothing to generalise from. A role with eleven applicants does not need ranking, it needs sourcing. Confidential and senior searches happen in conversations the system never sees, so a model has no basis to score them. Referrals arrive pre-qualified and get sorted alongside cold applications as though they were the same thing. Dirty data quietly poisons everything: duplicate profiles, five spellings of one company, resumes attached to the wrong person. The most common failure is upstream of the software entirely — a job description nobody agreed on produces a ranking nobody trusts, and the model gets blamed for a disagreement between two humans. Fix the definition first, then automate the sorting, and keep the screening logic visible enough that a recruiter can argue with it when it is wrong. A model cannot settle an argument two people have not had yet.
| Claim you will hear | What it usually means | How to test it live | What a failure looks like |
|---|---|---|---|
| Our AI parses any resume | A parser tuned on the formats it has seen most | Upload a scanned PDF and a two-column template | Fields land in the wrong place or come back empty |
| It surfaces your best candidates first | A similarity score between resume text and the job description | Rank a closed role's pool and find the person you hired | Your actual hire sits in the bottom half |
| The model learns from your decisions | Feedback is stored, but may not change scoring | Ask what measurably changes after fifty overrides | Nobody can describe the mechanism |
| Bias-free screening | Some fields were excluded from the model input | Ask which fields it reads and which you can hide | You get a certificate instead of a field list |
| AI writes your job posts | A prompt template over a general-purpose model | Generate a post for a niche role in your industry | Generic copy your hiring manager rewrites entirely |
| Trained on millions of profiles | A statement about the vendor, not about your roles | Ask how that improves output on the roles you hire | No answer that survives one follow-up question |
Pitch N Hire is an applicant tracking system. Post roles, screen applicants, run structured interviews, and make offers from a single pipeline — free for 1 user.
Free for 1 user · No credit card · Talk to a real hiring expert
Book a working session with our team, or start on the free-forever single-user plan and run the backtest yourself.
Prefer to talk? Book a demo · Talk to sales · View pricing
Free 1-user plan · No credit card · Talk to a real hiring expert
See your true cost-per-hire and how much Pitch N Hire could save you — our free Recruitment ROI Calculator gives you the numbers in under a minute. No signup required.
Open the free ROI calculatorPrefer a tailored walkthrough on your real roles? Drop your work email:
★ Free 1-user plan · No spam · Talk to a real hiring expert