Goal
You give your agent a trigger, run it on this week’s real work, define your checkpoint — and hand it to one colleague.
Why this matters
An agent that only ran on test input is a prototype. This week it becomes part of your workflow: real input, a defined gate, and the smallest possible test of whether it works beyond your own setup.
The drill
- Give it a trigger. Decide how it runs: on a schedule, so your Monday report is drafted before you arrive, or on demand from where you work. Match the cadence to the task you scouted in week 2. (Scheduling and invocation options depend on the agent type and your licence — confirm what’s available in your tenant.)
- Run it on this week’s real work. Not a test input. The actual project, the actual data. Read the output the way the recipient would: where does it hold at game speed, where does it wobble?
- Build the checkpoint. Define exactly where you review before anything leaves your hands. The loop is simple and non-negotiable: the agent produces, you check, then it ships. For anything going to a customer, a board, or outside your team, you are the gate — write down where that gate is.
- Hand it to one person. Share the agent with a single colleague and have them run it on their work. If it only works for you, it’s a personal tool. If it works for someone else, it’s a capability.
Extra round
Note where your colleague’s result differed from yours. Those differences are your next instruction refinements — and your first evidence of whether the agent travels.
Micro-win
Your agent did a real piece of this week’s work, and one other person used it.
Remember
The honest posture is supervised autopilot: on narrow, repeated, rule-based work, a well-instructed agent is genuinely reliable — and it still needs a human gate for anything consequential, because the cost of the occasional miss reaching a customer is too high. Widen the gate as trust in a specific agent grows — real minutes, watched closely, more responsibility as it’s earned.