By task · Candidate evaluation
Candidate evaluation software and candidate assessment scorecards with the evidence behind every rating
Candidate evaluation software is where most hiring processes quietly fall apart. Five interviewers, five different ideas of what good looks like, and a debrief where the loudest opinion wins. HireAgent fixes the first half of that before anyone books an interview: it turns your role brief into weighted criteria, reads every applicant and sourced candidate against them, and returns a ranked shortlist in which every rating links to the resume or profile line that earned it.
Source candidates · match-scored shortlist · personalized outreach
Ranked shortlist
Each candidate card shows the score, a criterion-by-criterion breakdown, what is missing and a ready outreach draft. Your team still runs the interviews and makes every hiring decision; the agent makes sure everyone walks into the debrief looking at the same evidence.
If you only need to put an existing applicant pool in order, candidate ranking software is the narrower purchase. If you are comparing screening vendors on price, the candidate screening software comparison covers twelve of them.
The short answer
Candidate evaluation software scores every candidate for a role against the same weighted criteria, records the evidence behind each rating and lines candidates up side by side so the hiring team decides on facts rather than memory. It comes as ATS scorecards, skills tests, interview notetakers and AI agents that evaluate resumes. Insist on two things: weights you set per role, and a rating you can trace to evidence. HireAgent does both from $299 a month.
A role in a ranked shortlist out
The agent sources you hire
Why it works
What you get with candidate evaluation software
One scorecard per role
Must-haves and nice-to-haves come from your brief and carry the weights you set, so a candidate who ticks eight minor boxes cannot outrank one who has the two that matter.
Ratings you can trace
Every criterion on the card links to the resume or profile line behind it, and a criterion with no evidence says so instead of guessing.
A record of the decision
The criteria version, the scores and who advanced or rejected each person are stored, which is what an auditor or a candidate challenge asks for.
What it handles
A role in, a match-scored shortlist out
Describe the role and HireAgent sources candidates, screens them against your criteria, and returns a ranked shortlist with a match score and the evidence behind it, then drafts personalized outreach and schedules interviews. The agent does the legwork, you make the hire.
- Turns a role brief into weighted, must-have and nice-to-have criteria
- Scores applicants, past candidates and sourced profiles on one scale
- Links each rating to the evidence that earned it
- Flags missing must-haves instead of averaging them away
- Lines candidates up side by side for the debrief
- Keeps an audit trail of criteria, scores and decisions
evidence · Direct experience with the role requirements.
evidence · Strong overlap; one stack tool is adjacent.
evidence · Slightly junior for the scope as described.
evidence · In timezone and open to a move now.
Start here
What does candidate evaluation software actually do?
Candidate evaluation software does four jobs: it fixes the criteria for a role before anyone reads a resume, it gathers evidence against those criteria from resumes, tests and interviews, it converts that evidence into ratings on one scale, and it keeps a record of who decided what and why. Tools differ mainly in which of the four they cover.
The first job is the one teams skip. A hiring manager says "strong communicator, senior, knows our stack", the recruiter writes a job ad, and three interviewers each carry a slightly different version of the role into their conversations. Evaluation software forces the version into writing: five to eight criteria, each marked as a must-have or a nice-to-have, each with a weight. Once that exists, every later rating has something to point at.
The second and third jobs are where the categories split. An applicant tracking system collects ratings from interviewers on a scorecard form. A skills testing tool gathers evidence by setting tasks. An interview notetaker records the conversation and summarizes it against the questions. An AI evaluation agent reads the resume or profile itself and rates each criterion with the line that supports it. HireAgent sits in that last group, which is why it is most useful at the top of the funnel, when there are 200 applicants and nobody has spoken to any of them.
The fourth job, the record, is now a legal question as much as an operational one. New York City, Illinois and California all put duties on employers whose tools substantially drive who advances, and a debrief note that says "felt like a better fit" does not survive a challenge. A stored scorecard with criteria, ratings and evidence does.
Categories
Four kinds of candidate evaluation software and who each one suits
| Kind | Evidence it evaluates | Best for | Where it falls short |
|---|---|---|---|
| ATS scorecards (Greenhouse, Ashby, Manatal) | Interviewer ratings typed into a form after each interview | Teams that already run structured interviews and need the ratings in one place | Nothing is evaluated until a human has interviewed, so the top of the funnel is still read by hand |
| Skills tests (TestGorilla and similar) | Test scores, work samples, recorded answers | Roles where a task predicts performance: support, sales development, data entry, junior engineering | Candidates drop out of long tests, and a test says nothing about experience or career trajectory |
| Interview notetakers (Metaview and similar) | Transcripts and summaries of interviews | Teams running many interviews who want consistent notes and less writing after each call | Only evaluates people who have already reached an interview |
| AI evaluation agents (HireAgent) | Resumes, profiles and applications read against the weighted criteria | Teams with too many applicants or too few, who need a ranked, explained shortlist before interviews start | Cannot judge what is not written down, so thin profiles are flagged for a human rather than scored low |
Most teams end up with two of these, not four. The common pairing is an AI evaluation layer at the top of the funnel and ATS scorecards for the interview loop, with a skills test added only for roles where a task genuinely predicts the job. Buying a test for every role is the most frequent way evaluation budgets get wasted: completion rates fall on senior roles, and the strongest passive candidates simply do not take a 45-minute assessment for a job they have not decided they want.
Method
How do you compare candidates objectively?
To compare candidates objectively, agree on weighted criteria before you see anyone, rate every candidate on the same scale against written evidence, and total the weighted scores only after each criterion is rated on its own. Reading the whole resume and then picking a number is the habit that lets a strong first impression decide everything.
Here is a worked example for a customer success manager role with five criteria. Weights add to 100.
| Criterion | Weight | Candidate A | Candidate B | Candidate C |
|---|---|---|---|---|
| Owned a B2B book of 40+ accounts (must-have) | 30 | 5 | 4 | 2 |
| Renewal or expansion targets met | 25 | 3 | 5 | 4 |
| SaaS onboarding experience | 20 | 4 | 3 | 5 |
| Salesforce or HubSpot daily use | 15 | 5 | 4 | 5 |
| Mentored junior CSMs | 10 | 1 | 4 | 3 |
| Weighted score out of 5 | 3.90 | 4.05 | 3.65 |
Candidate C has the best onboarding and tooling ratings and the most impressive resume to skim, yet ranks last, because the must-have carries a 2. Candidate A looks strongest on the must-have and loses to B on renewals. That is the conversation a debrief should have, and it only happens when the ratings are written down criterion by criterion before anyone argues.
Two rules make the method hold. A missing must-have caps the total instead of being averaged away; in HireAgent a candidate without a must-have cannot rank above a candidate who has all of them. And every rating needs a sentence of evidence next to it. A 4 with "managed 52 mid-market accounts at a payroll SaaS, 2021 to 2024" is a rating. A 4 with nothing next to it is a feeling.
Scorecards
What should a candidate evaluation scorecard include?
A candidate evaluation scorecard should include five to eight job-related criteria, a must-have or nice-to-have flag and a weight for each, a rating scale with a written definition of every point, a space for the evidence behind each rating, and an overall recommendation recorded separately from the criterion ratings.
The rating definitions are what most scorecards leave out, and they are what makes two interviewers comparable. "3 = meets the bar" means nothing on its own. "3 = has owned a book of accounts of this size for at least a year, with named outcomes" means the same thing to everyone in the loop.
Keep the criteria about the work. Years of experience is a weak criterion on its own; what the person did in those years is a strong one. Avoid criteria that act as proxies for protected characteristics, such as graduation year, a specific university list or "culture fit" with no definition, because those are exactly what an adverse-impact review looks for.
Finally, separate the overall recommendation from the ratings and collect it last. If interviewers give a hire or no-hire first and fill in the criteria afterwards, the criteria get bent to match the verdict. HireAgent applies the same discipline to its own output: it rates each criterion on evidence, totals them, and leaves the recommendation and every advance or reject to a person on your team.
US rules
Is AI candidate evaluation legal for US employers?
Yes. US employers can use AI to evaluate candidates, but where a tool substantially drives who advances, several jurisdictions require notice, bias audits and record keeping, and the rules follow where the candidate is located.
New York City Local Law 144 requires an independent bias audit within the past year, a published summary and candidate notice at least 10 business days before an automated employment decision tool is used for a New York City role. Illinois requires notice when AI is used in employment decisions and bars using zip code as a proxy for a protected class. California regulations on automated-decision systems under FEHA, in effect since 1 October 2025, prohibit discriminatory use and require four years of record retention. Title VII adverse-impact analysis applies to any selection procedure, automated or not.
There is also an open question about consumer reporting. A proposed class action filed in California on 20 January 2026 alleges that one large talent-intelligence vendor compiled data on candidates and delivered match scores to employers in a way that makes it a consumer reporting agency under the Fair Credit Reporting Act. The case is unresolved and the allegations are only allegations, but the practical lesson for buyers is clear: ask every vendor what data it evaluates beyond what the candidate submitted, and whether candidates can see and dispute it.
HireAgent evaluates the application, the resume and the profile the candidate or your sourcing brief points to, stores the criteria version, the rating and the evidence for every score, and leaves every decision to a human. That is the record each of these rules asks for.
Cost
How much does candidate evaluation software cost?
Candidate evaluation software costs from about $15 per user a month for ATS scorecards with AI scoring, to $215 a month and up for a skills-testing account, to $299 a month for an AI agent that evaluates and ranks candidates, to $100,000 a year and more for enterprise talent-intelligence platforms. The figures below are read from each vendor's own pricing page or marketplace listing.
The enterprise end deserves a number because almost nobody publishes one. Eightfold lists a Starter Edition on the SAP Store at USD 25,000 per quarter with a one-year minimum, priced by the customer's employee count, which puts the entry point at $100,000 a year before anything is configured. Our Eightfold AI pricing breakdown covers what that figure includes and what it leaves out.
HireAgent publishes every plan. Solo is $299 a month for one seat and up to three open roles, or $1,794 a year billed annually. Growth is $799 a month, or $4,794 a year, for three seats, up to ten roles, autosend outreach and ATS integration with Greenhouse, Lever and Ashby. Scale is $1,999 a month for ten seats and SSO. Every plan includes sourcing, evaluation, ranking, outreach and scheduling.
The comparison most buyers actually face is time, not another tool. If a recruiter spends 6 minutes on each of 250 applicants for one role, that is 25 hours of reading before the first phone screen. Cutting it to the top 40 with evidence already attached gives back most of a working week per role. The AI recruiting software pricing page sets every vendor we have priced side by side.
Who buys it
Candidate evaluation software for in-house teams, agencies and hiring managers
In-house talent teams buy evaluation software to make a debrief defensible. A mid-size US company running 30 open roles has dozens of interviewers who were never trained on the same bar, and the evaluation layer is how the recruiter keeps the loop honest. For high applicant volume, the high volume hiring software page covers the extra controls.
Staffing and search firms buy it for the submittal. A client asks why these three candidates and not the other forty, and a criterion-by-criterion card with evidence answers that in one page. The AI recruiter for staffing agencies page walks through that workflow.
Hiring managers without a recruiter buy it to avoid reading 300 resumes at night. A founder or department head can write the brief, set the must-haves and get a ranked, explained shortlist the next morning; the AI recruiter for hiring managers page shows how that runs.
Side by side
Candidate evaluation software compared on published US prices
Read from each vendor's own pricing page or marketplace listing. Per-user, per-account and per-company prices are not directly comparable, so the second column says what the price buys.
| Tool | What it evaluates | Published US price | Best fit |
|---|---|---|---|
| Manatal | Applicants in its ATS, with AI recommendations and scoring on every plan | $15 a user a month annual, $19 monthly (Professional) | Small agencies that want an ATS with basic AI scoring |
| Ashby | Interview scorecards plus AI-assisted application review, metered by AI credits | $300 to $900 a month by company size, up to 100 employees | In-house teams replacing their ATS |
| Greenhouse | Structured interview kits and scorecards inside the ATS | Quoted, no price on its pricing page | Mid-market and enterprise teams with formal interview loops |
| TestGorilla | Skills tests, video questions and resume scoring | Assessments from $215 a month per account, billed annually; Plus from $520 | Roles where a work sample predicts performance |
| Metaview | Interview notes, and applications through its Application Review agent | Notetaker Pro $60 a user a month; Application Review $150 to $25,000 a month by volume | Teams running many interviews |
| Eightfold | Talent intelligence across hiring and internal mobility | Starter Edition USD 25,000 per quarter on the SAP Store, one-year minimum | Enterprises with thousands of employees |
| HireAgent | Applicants, ATS records and sourced profiles against weighted criteria, with evidence per rating | $299 a month (Solo), $799 (Growth), $1,999 (Scale) | Teams that need a ranked, explained shortlist before interviews |
Free tiers and trial plans are not listed. Prices change; check each vendor before you sign.
Why HireAgent
One agent that sources, screens and ranks candidates
Not a job-board blast and not a resume pile. HireAgent sources candidates, screens them against your criteria, and returns a match-scored shortlist with the evidence behind each fit, then drafts outreach and books interviews. The agent does the legwork, you make the hire.
Criteria-based screening
Every candidate is screened against the same structured criteria you set, with a match score on a red to amber to green scale, so screening stays consistent and fair.
Evidence behind every match
Each match score links to the experience that earned it, the role, the skill, the timeline, so the fit is auditable and your shortlist is defensible.
A ranked shortlist
Match scores roll up into a ranked list, so the strongest candidates are already at the top and your team reviews the best fits first.
Good questions
Questions about candidate evaluation software
Explore more
More ways to recruit with HireAgent
Hiring and interviewing software
One agent that sources, screens, ranks and books, instead of four tools and a gap.
Learn moreAI talent acquisition
Run talent acquisition end to end with one accountable AI agent.
Learn moreRecruitment automation tools
Replace a stack of automation tools with one recruiter agent.
Learn moreStop digging through resumes. Put recruiting on autopilot.
Describe the role and HireAgent sources candidates, screens them against your criteria, and returns a match-scored shortlist, then drafts personalized outreach and schedules interviews. The agent does the legwork, you make the hire.
Engineering, data, sales, support & product · consent and AI disclosure · the agent sources, you hire