# AI Recruiting Evaluation Scorecard

> Score an AI recruiting system across eight evidence, control, risk, and outcome dimensions before running or expanding a real hiring pilot.

- Canonical HTML: https://openjobs.genedai.me/evaluation-scorecard
- Last substantive review: 2026-08-07
- Maximum score: 24
- Scope: Evaluation aid; not proof of legal compliance, fairness, or business value.

## Scoring scale

- **0: No evidence**
- **1: Claim or scripted demo only**
- **2: Pilot evidence with gaps**
- **3: Repeatable, auditable evidence**

Use evidence from a real workflow rather than a scripted demo.

## Dimensions

### 1. Outcome definition

The team agrees what "qualified, interested, and worth interviewing" means before reviewing results.

### 2. Brief fidelity

Must-haves, preferences, trade-offs, geography, and evidence of seniority are explicit and approved.

### 3. Search provenance

Sources, coverage, refresh limits, exclusions, and likely blind spots are documented for the role.

### 4. Match evidence

Every recommendation can be traced to role-relevant evidence, and reviewers can inspect rejects and borderline cases.

### 5. External action control

Sender identity, channel, message, timing, follow-up, suppression, and approval boundaries are visible and enforceable.

### 6. Candidate experience and accessibility

People can understand the process, request an accommodation, correct relevant data, opt out, and reach a human.

### 7. Handoff quality

The hiring team receives evidence, current interest, unresolved questions, and a clear next step. A raw queue does not meet this standard.

### 8. Audit and fallback

Inputs, edits, approvals, actions, and outcomes remain distinguishable; automation can pause without losing the workflow.

## Interpretation

- **0 to 8, stop:** Evidence or controls are too weak to proceed.
- **9 to 16, narrow pilot only:** Limit scope, keep the workflow reversible, and close identified evidence gaps.
- **17 to 24, ready for real-role validation:** The system has enough evidence to be tested on a real role; this is not a compliance or effectiveness conclusion.

## Deep review guides

- [Vendor evaluation checklist](https://openjobs.genedai.me/vendor-checklist)
- [Pilot design and metrics](https://openjobs.genedai.me/pilot-design)
- [Sourcing and ranking evaluation](https://openjobs.genedai.me/sourcing-evaluation)
- [Candidate screening evaluation](https://openjobs.genedai.me/screening-evaluation)
- [Agent reliability evaluation](https://openjobs.genedai.me/agent-reliability)
- [Evidence and scoring methodology](https://openjobs.genedai.me/methodology)

## Next step

Ask the vendor to demonstrate every scored dimension using one role, including weak matches, corrections, approval points, and the final handoff. Read the [field guide](https://openjobs.genedai.me/) and verify claims against the [primary-source ledger](https://openjobs.genedai.me/sources).

The interactive HTML scorecard runs entirely in the browser and does not store or submit answers.
