Ten people across six disciplines — structured as a delivery organisation rather than a pool of generalists who get assigned to whatever came in.
The distinction matters because it decides what happens when two projects collide. A pool reassigns whoever is free. An organisation has an owner for each stage who was always going to do that stage.
Process mapping, business cases, verdicts
System design, integration boundaries, the review gate
Agent construction, prompt and tool design
Connector development against the adapter interface
Golden sets, regression suites, the deploy gate
Baselines, outcome reports, retention
We would be poor advocates for AI workforce engineering if we did not run one ourselves. Ten internal agents cover intake through proof, each subject to the same autonomy gates and evaluation discipline we sell.
Six of the ten draft rather than act. That ratio is deliberate — it is what we would recommend to a client at our stage, so it is what we do.
Qualifies inbound, scores fit, routes to a pod
Turns interviews into a formal process map
Reviews the map with the client before anything is priced.
Three environments with three different sets of rules. The middle one is where most firms cut the corner — testing against a copy of production because it is easier than generating realistic synthetic data. We do not.
Development happens locally with branches, pull requests and CI. Production deploys only through CI with the evaluation gate passed. On-call is named from the first deployment rather than after the first two-in-the-morning failure.
Monitoring that alerts into a void is theatre. If nobody is holding the pager, the dashboard is decoration.