AI Integration
Eleven Things to Confirm Before Going Live
Last updated:
Quality
- An evaluation set of real cases, running automatically on every change
- A measured accuracy figure, above the threshold agreed before the build
- Confidence thresholds set from a shadow run, not from intuition
Failure behaviour
- Decided behaviour for outage, rate limit, timeout and bad output
- A tested switch that disables the AI and falls back to manual
- Quarantine for repeated failures, with the input preserved
Test the switch before launch, with the business owner operating it. A control nobody has used is a control nobody will reach for.
Cost and monitoring
- A hard cap per day and per user, with a decided behaviour at the cap
- Alerts on quality drift, cost per item and output volume
- Alerts routed to people who can act, not to a shared channel nobody reads
Access and records
- Credentials in your accounts, with expiry monitored
- Audit trail recording input, context, model version, output and reviewer
Ownership
- A named business owner for quality
- A named technical contact for changes
- A runbook covering each alert and the first three checks
- A ninety-day review booked, with the baseline attached
Eleven items. Most take hours rather than days, and the cost of skipping them is paid at the worst possible moment.
Frequently asked questions
Can we launch without all of these?
The audit trail can be lighter for low-stakes internal tools. Evaluation, the switch and the cost cap are not negotiable.
How long does this add?
A week or two if designed in from the start. Considerably more if it is all retrofitted after the build.
Who signs off?
The business owner of the process, having seen the numbers and used the switch. Not the supplier.
What about a soft launch?
Two or three users first, always. First impressions across a whole team are hard to revise.
About to put an AI feature live?
Run through these eleven first. Happy to review your readiness before you launch.