AI agent for director of operations
Process Pilot Tracker Agent
Decide every pilot on evidence: scale, change or stop.
What it does
A team trials a new picking method for two months, everyone has an opinion and no one checks the numbers against the starting point. When a pilot is approved, the agent sets the baseline from the metrics the pilot should change, such as cycle time, error rate, cost or throughput, and records the plan and the end date. Each week it collects the pilot's results and compares them with the baseline and with a control group where there is one. It checks whether the sample is large enough and whether side effects appear, for example quality dropping while speed improves. If the sample is too small to judge, it keeps the pilot running and checks again next week. At the end, it recommends scale, change or stop. The leader approves. Edge case: a seasonal spike that distorts results.
How it works
Follow the arrows from top to bottom. The orange dashed arrow is the loop: when a check fails, the agent goes back and tries again.
Read the steps as a list
- A pilot is approved
- Collect baseline data for the main and side metrics
- Record the target, sample size needed and end date
- Each week, pull the pilot results
- Compare with the baseline and the control group
- Is the sample large enough to judge?If not: keep the pilot running and check again next week. Back to step 4.
- Check side effects and seasonal distortions
- Is the main metric at or above the target without a harmful side effect?If not: identify the likely cause and propose a change to the pilot. Back to step 4.
- Write the recommendation: scale, change or stop
- Leader approves the decisionThe agent waits here for your OK.
- Pilot closed and result logged
How it decides
A pilot is judged only when the sample reaches the minimum size. Scale when the main metric improves by the target without a harmful side effect, otherwise change or stop.
- Judge a pilot only after at least 30 cycles or 4 weeks
- Scale when the main metric improves by 10 percent or more
- Stop when a side effect worsens by more than 5 percent
- Adjust for seasonal weeks
Make it yours
Every agent is a starting point. You choose these settings for your own situation.
- Minimum sample (default 30 cycles)
- Improvement target
- Side-effect limits
- Review day
- Metrics used
What keeps you in control
It always asks you first
- Leader approves scale, change or stop
- Leader approves extending a pilot
Hard limits
- Never end or expand a pilot itself
- Report side effects even when the main result is good
It stops when
- Done: decision approved and logged
- Stop: data collection fails for 2 weeks, tell the leader
Set it up
We guide you through the set-up, step by step
Members get the full set-up guide for this agent. No technical skills needed: you copy, paste and upload.
- One set of instructions to paste into your AI, with the clicks for ChatGPT, Claude, Microsoft 365 Copilot, Gemini and Grok
- The agent then walks you through connecting your own data, one source at a time
- A downloadable copy with the flow chart, the rules and the full guide