How do we know it will not make things worse?
You do not, at the start, and neither do we. That is why the boundary of what the agent may do alone is agreed before the build and written into the contract, why every consequential action passes an approval gate, why the rollback path is tested rather than assumed, and why accuracy is scored against a fixed set of your own cases. The question is not whether an AI system will be wrong. It is what happens in the minute after it is.
Why do you insist on measuring the process first?
Because Gartner's stated causes of agentic project cancellation are escalating costs, unclear business value and inadequate risk controls, and a baseline addresses two of the three directly. Without one, there is no way to defend the spend at renewal and no way to decide what to scale. The second project then gets approved on the same evidence as the first, which is none.
Can you use our existing model provider contract?
Yes, and it is usually preferable. We build on your contracts, in your tenancy, on your accounts. If you have no provider relationship yet we will help you choose one on the merits of your workload and residency requirements, and we hold no reseller agreement or commission on any of them.
Our data cannot leave our environment. Is this still possible?
Yes. Open-weight models run fully self-hosted in your own tenancy, which is slower to build and usually costlier to run than an API, and we will tell you the trade-off in numbers rather than in principle. Region is treated as a requirement rather than a preference, and every third party in the path is listed before work starts with a standing right for you to refuse any of them.
What happens when the model underneath is deprecated?
It is planned for, because it is certain. OpenAI announced retirements from ChatGPT effective 13 February 2026 and Anthropic's published policy is at least 60 days notice on publicly released models. The version is pinned, deprecation notices are tracked, and a successor is scored against your evaluation set before cutover. That handling sits in managed operation.
Will you tell us if we should not build this?
Yes, and it happens. Where the process is unstable, the data unreachable, or the volume too low to repay the build, a no-go with the reasoning attached is the outcome of the Diagnostic. You keep the report, the baseline method and the plan. We would rather lose the build than sell one that fails.