AI Agent Architecture Cheat Sheet
Every moving part drawn out, including the point where a person signs off.
Twenty-four questions your agent has to answer before it is allowed to act without a person checking each run. Six of them are hard stops. Score it honestly, get a verdict, and leave with the three things to fix first.
Most agents do not fail at the demo. They fail in month three, when the person who watched every run has moved on, the prompt has been edited four times, and nobody can say who approved the version that is now sending the emails. The gate makes the hand-over from supervised pilot to unattended operation an explicit act by a named owner — with the conditions written down before the decision, not reconstructed after the incident.
Bring the person who owns the outcome and the person who built the agent. Half the questions are about intent, half about mechanics, and the gaps live where those two have not talked.
Next three moves
Paste the scorecard into your enquiry and we start from your numbers, not from a blank page.
A go-live gate is a fixed set of conditions an AI agent must meet before it is allowed to act without a person checking each run. It sits between the pilot, where a human watches everything, and unattended operation, where the human only samples. The gate makes the hand-over an explicit decision by a named owner instead of something that drifts.
Twenty-four questions, each answered No (0), Partly (1) or Yes (2), for a maximum of 48. Six questions are hard stops: if any of them is not a full Yes, the verdict is capped at 'not cleared for unattended operation' regardless of the total. A high score with a failed hard stop means the foundations are good and one specific thing is missing.
About eight minutes if you know the agent. You need the person who owns the outcome and the person who built it in the same room, because half the questions are about intent and half are about mechanics. Nothing you enter leaves your browser.
On every change to the prompt, tools, model version or data sources, and on a fixed cadence — quarterly is a sensible default — even when nothing changed, because the world around the agent does.
No. It is an engineering and governance readiness check drawn from ImageFirm's own agent work. It maps naturally onto the human-oversight, logging and risk-management expectations that regulators are converging on, but it does not replace legal advice for your jurisdiction and sector.
The twenty-four conditions are distilled from ImageFirm's own agent builds and rollouts — the same questions we put to ourselves before an agent of ours runs without a hand on it. They are deliberately vendor-neutral: nothing here depends on a particular model, framework or cloud. The hard stops are the six failures we have seen cause real damage; the rest cause expensive embarrassment. Revised when our practice changes, not on a content calendar.
If the verdict was not the one you wanted, the gaps are now specific and named. We close them the same way we opened them: a person in final authority, the method in the open, and a written record of where the machine stops.
Start an enquiry