Three steps, six stages. At the end of each stage you see the evidence and decide whether to spend more money, add more people, or give the system more autonomy.
| Step | Stage | Question answered | What must be true to move on |
|---|---|---|---|
| Scope | 1 Select | Which work matters most, and is it ready? | Named owner, measurable outcome, usable data |
| Scope | 2 Diagnose | How does the work really happen, and where is the bottleneck? | Baseline agreed with the business owner |
| Build | 3 Design | What is the lowest level of automation that solves it? | Design approved by the owner |
| Build | 4 Prove | Does it meet the pass mark set in advance, and beat the non-AI option? | Pass mark met, zero unacceptable errors |
| Run | 5 Pilot | Does it work for real users under enforceable controls? | Used independently, no unresolved critical incidents |
| Run | 6 Decide | Did the value arrive, and what happens next? | Scale, iterate, or stop, with reasons |
A project can stop at any stage. Stopping early with evidence is a good outcome. It saves the investment that would have followed.
Every step is tested against four questions before any AI is considered. Most stop at the first or second. Across our four published builds, 59% of steps run on rules.
You agree the pass mark before testing starts. It performs, or it stops.
Business, AI operations, engineering, and risk each own a part. Any of them can pause it.
| When | What lands on your desk |
|---|---|
| End of week 2 | Ranked shortlist, map of the work, agreed baseline |
| Design | Signed-off design, with each step marked rules, AI-assisted, or human |
| Prove | Test results against your pass mark |
| Pilot | Your team running it, with training and monitoring in place |
| Decide | Results against the baseline and a decision memo |
Five tools take you from how people say the work happens to how it actually happens, and then to a step-by-step call on rules, AI, or a person.
Half an hour with Andrew. If AI is not the answer, you will hear that too.