Stage purpose
Observe the workflow with real users and tightly bounded production conditions while preserving independent review and a safe fallback.
Use-case boundary
AI can draft requisitions, answer candidate questions, schedule, and organize applicant evidence. When it ranks, filters, recommends, or materially influences a selection decision, job relatedness, accessibility, notice, audit, and human accountability become central.
Write the specific population, users, systems, source records, proposed AI contribution, human decision, allowed action, and business consequence. State what remains outside the stage. A bounded plan prevents a successful test of one narrow task from becoming an unsupported approval for a broader operating process.
Entry condition
The controlled test met its thresholds, material risks have owners, affected people are trained, and the organization can stop or reverse the workflow.
Do not waive the entry gate because a tool is already licensed or a provider offers a short implementation window. Existing access can reduce procurement time, but it does not resolve purpose, authority, evidence, ownership, privacy, security, operating fit, or measurement.
Work to complete
- limit users, population, duration, and actions
- monitor quality, adoption, exceptions, and rework
- sample retained evidence
- review incidents and complaints
- measure full operating cost and outcome
Limited operating trial scenario
For recruiting and candidate support, select one decision with a known outcome and one unresolved case that represents the edge of the intended scope. Document the people, source systems, records, timing, current work, consequences, and existing controls. Run only the actions allowed at the limited operating trial stage, and keep any generated or recommended output outside a broader production decision until the exit gate is met.
The stage owner should be able to explain why this population is representative, which groups or situations are excluded, how a user challenges an output, where a difficult exception goes, and what evidence will support the next decision. If those answers are not yet available, the correct result may be to narrow the stage rather than accelerate it.
Test design
Use representative records and preserve the denominator. Include a normal path, missing information, contradictory evidence, an unusual case, an authorized override, and a changed source, policy, model, or integration. Capture input, version, output, reviewer action, time, error, rework, exception, and downstream consequence for every test case.
Decision questions
- Does the tool materially influence who advances?
- What validated job criteria support the output?
- How can a candidate request accommodation or challenge an error?
Evidence requirements
- traceable inputs
- reviewable outputs
- human decision record
- measured outcome and failure evidence
Risk and incident controls
- disparate impact
- inaccessible assessment
- automation bias
Name the person who can stop the stage, the event that requires immediate pause, the fallback process, how affected records will be corrected, who must be notified, and what evidence is needed before work can resume. The plan should also address participant feedback and challenge when outputs affect people, customers, partners, investors, or regulated activity.
Measures
| Dimension | Measure | Decision use |
|---|---|---|
| Quality | Correct, incomplete, unsupported, conflicting, and materially wrong outputs | Determine whether review is practical and error is acceptable |
| Work | Cycle time, touch time, rework, exceptions, and support burden | Test the complete operating case rather than generation speed |
| Outcome | Role-specific business result against the baseline and comparison group | Separate activity from value |
| Risk | Incidents, near misses, complaints, overrides, and affected populations | Test whether controls and escalation work |
| Adoption | Correct use, avoidance, workarounds, challenge, and confidence calibration | Understand whether the operating model is usable |
Authority and policy checkpoint
AI Principles for Worker Well-Being
Evaluate engagement, rights, data, job quality, and accountability.
The authority record does not certify a product, provider, program, or organization and does not determine buyer-specific applicability.
New York City Local Law 144 AEDT rules
Track scope, bias audit, publication, and notice responsibilities.
The authority record does not certify a product, provider, program, or organization and does not determine buyer-specific applicability.
Official sources for the stage review
AI Principles for Worker Well-Being — U.S. Department of Labor. The authority record does not certify a product, provider, program, or organization and does not determine buyer-specific applicability.
New York City Local Law 144 AEDT rules — NYC Department of Consumer and Worker Protection. The authority record does not certify a product, provider, program, or organization and does not determine buyer-specific applicability.
Exit condition
A dated decision states whether to stop, redesign, extend the trial, or prepare for scale, with evidence and dissent preserved.
The exit record should state what was observed, which claims were supported or rejected, which limitations remain, whether the population was representative, who approved the decision, and what evidence could reverse it. Silence or project momentum is not approval.
The publication supports research and executive decision preparation. It does not provide legal, financial, accounting, employment, clinical, cybersecurity, investment, procurement, or implementation advice.