An agent only its builder understands is a personal project running in production. A named owner, a written operating guide and a periodic test set are the minimum bar before calling it finished.
This closes the cluster this series has spent five pieces building: what an agent actually is, why the method has to come first, and where the absolute limits sit. This piece answers the practical question that follows once an agent is genuinely built and working: what stops it quietly breaking, or quietly drifting into something nobody understands, the day the person who built it moves on. The honest answer is three artefacts, none of them technical, all of them cheap compared to the cost of finding out the hard way that nobody else can touch the thing.
An agent that only its builder understands is not actually finished. It is a personal tool wearing the appearance of organisational infrastructure, and the gap between those two things only becomes visible at the worst possible moment, usually right after the builder has already left.
The person who built the agent and the person who owns it going forward do not need to be the same person, and for most organisations this size, keeping them the same is actually the risk. Ownership here means something specific: one named person, not a team, who is accountable for the agent still doing what it is supposed to do, who gets told first when something looks wrong, and who has the authority to switch it off. If the builder leaves and nobody was ever named as owner, the agent does not fail cleanly. It keeps running, unattended, until something breaks badly enough that someone finally asks who is meant to be watching it.
The operating guide is the artefact that actually transfers the knowledge the builder is carrying around in their head: what the agent does, what it does not do, what normal output looks like, what a wrong output looks like, and what to do when one shows up. This is not a technical specification. It is closer to the method this series argued for building before the agent existed in the first place, now updated to describe what actually got built rather than what was intended. An owner who inherits an agent with no operating guide is not really an owner. They are a person with responsibility for something they cannot actually evaluate.
The test set is a small, fixed collection of real inputs with known-correct outputs, run periodically against the agent to confirm it still behaves the way it did when it was accepted. This matters because the way an agent quietly stops working is rarely a dramatic failure. It is a slow drift, a platform update, a small change in the documents it reads, a shift in the task itself, that nobody notices until the output has been subtly wrong for weeks. A test set turns that invisible drift into something a person can actually check on a schedule, rather than something that only surfaces once a beneficiary or a funder points it out.
A working agent with no named owner, no operating guide, and no test set is not really an organisational asset. It is a personal project that happens to be running in production, and the organisation will not find out the difference until the person who built it is no longer there to explain it.
Treat a named owner, a written operating guide, and a periodic test set as the minimum bar for calling an agent finished, not optional extras added once time allows, since without all three the agent only survives for as long as its builder stays.
No external statistic cited; this article presents an internal practical framework rather than third-party evidence.
A short, absolute list of tasks no agent should be configured to do: safeguarding referrals, eligibility decisions, payment authorisation and public statements on live matters, where the human is the safeguard.
Custom Agents & Tools · 3 minAnalysis · 30 September 2026If nobody can write down in plain steps how a task is done today, buying an agent for it is premature. The written method is the real readiness test, and it surfaces hidden disagreements early.
Custom Agents & Tools · 4 minAnalysis · 30 September 2026A tool, a skill and an agent cost very different amounts to build, run and govern. One question, whether the task must decide what happens next, tells you which one to choose.
Custom Agents & Tools · 4 minIf this is the question on your desk, a thirty-minute call tells you whether the service fits, or that you do not need us yet.