Free · Fillable · Use it yourself

Test AI on one real task before investing in a bigger system.

Use this 14-day test plan to define the job, what information AI may use, what a person still owns, how the result gets checked and when the test should stop.

Two people and an AI assistant build a working plan with visible ownership and safeguards.

Complete this with the person who owns the result and someone who performs the work. If those are the same person, answer from both perspectives.

1. Responsibility

  1. Work we are trying to improve: Name one recurring responsibility, not “use AI.”
  2. Current friction: What repeats, slows down, gets missed or requires avoidable rework?
  3. Human decision: Which judgment, promise or relationship remains accountable to a person?

2. Boundaries

  1. AI-supported portion: What small gathering, organizing, comparing, routing or drafting task may receive assistance?
  2. Permitted inputs: What information is approved for this system?
  3. Prohibited inputs: What private, regulated, sensitive or unverified information must stay out?
  4. Source of truth: Which approved records, policies or examples should the result follow?

3. Review

  1. Expected output: What should the system return, in what format, and for whom?
  2. Named reviewer: Who checks the result before anyone relies on it?
  3. Quality standard: What must be accurate, supported, complete, on-voice and appropriate?
  4. Escalation: When must the system stop, mark uncertainty or hand the work to a person?

4. 14-day test

  1. Baseline: How does the work happen now, and where does the friction appear?
  2. Small measure: Compare repeated steps, preparation time, missed contacts, preventable revision or another observable condition.
  3. Stopping condition: What failure, risk or extra supervision means the test should pause?
  4. Final decision: Keep, revise, stop or investigate a different problem?
The test is not ready to begin until a named person has the time, standard and authority to review the result. Record that person, the review standard and the escalation path before the test starts.

Reusable instruction frame

You are supporting [bounded task] for [person or team]. Use only [approved sources]. Do not use or infer [prohibited information]. Produce [specific output]. Mark uncertainty and missing information instead of guessing. Stop and request human review when [escalation condition]. The final decision and approval belong to [named role].

After the test

Put what works into a workflow people can repeat.

A successful experiment still needs a clear trigger, approved information, human review and a place in the regular work.