How to Pilot an AI Voice Agent: A Practical 6-Week Playbook for UK Contact Centres

Arkadas Kilic
Author: Arkadas Kilic, Founder & CEO, Rel8 CX

Most AI voice agent pilots fail before they start. Not because the technology does not work, but because the pilot is designed as a science experiment rather than a production deployment. Teams spend weeks on demos, skip compliance review, and never define what "good" looks like. Six months later, the initiative is shelved.

This playbook is different. It is the exact approach we use to take UK contact centres from kickoff to a live, enterprise-grade AI voice agent in 4 to 6 weeks. Every milestone is actionable. Every number is real.


Before You Start: Define the Problem, Not the Technology

The single biggest mistake contact centre leaders make is starting with the technology. Before you write a single line of configuration, answer these four questions:

1. Which call type are you automating? Pick one. Ideal candidates handle 500 or more calls per month, follow a predictable script, and have a clear resolution outcome (e.g. balance enquiry, appointment booking, payment arrangement, meter reading).

2. What does success look like at week 6? Set a number. A realistic target for a first pilot is 60 to 70% containment on the chosen call type, with a CSAT score within 5 points of your human agent baseline.

3. Who owns compliance sign-off? In UK regulated industries (financial services, utilities, healthcare), FCA, Ofgem, or CQC obligations apply to automated voice interactions. Identify your compliance stakeholder on day one, not week five.

4. What is your telephony stack? If you are already on Amazon Connect, your path to production is significantly shorter. If you are on a legacy PBX, budget an extra week for SIP trunk configuration.


Week 1: Discovery and Data Baseline

Goal: Understand the call you are automating at a forensic level.

Pull 90 days of call recordings for your chosen call type. You need a minimum of 200 calls. Listen to 50 of them. This is not optional.

From this analysis, document:

Deliver a one-page brief summarising these findings. This document governs every design decision for the next five weeks.


Week 2: Architecture and Compliance Design

Goal: Design the system before you build it.

For UK contact centres, we build on AWS. Amazon Connect is the telephony layer. Amazon Lex handles natural language understanding. AWS Lambda executes business logic. Amazon DynamoDB or your existing CRM stores session state. All of this runs inside your AWS account, in the eu-west-2 (London) region by default.

Compliance checkpoints to complete this week:

Draft your escalation logic at this stage. Every AI voice agent must have a clean handoff to a human agent. Define the exact triggers: caller requests a human, agent confidence score falls below threshold (we use 0.75 as a default), or three consecutive misunderstood utterances.


Week 3: Build Sprint One (Core Flow)

Goal: A working agent that handles the top three intents end to end.

Do not try to build everything. Build the core flow that covers the top three intents identified in week one. In a typical UK utility contact centre, this might be: balance enquiry, payment arrangement, and direct debit update. Together, these three intents usually account for 55 to 65% of total call volume on the chosen call type.

What to build this week: A note on voice persona: Choose a voice that matches your brand. AWS offers 20 or more UK English neural voices via Amazon Polly. Test at least five with internal stakeholders before locking in. This decision is harder to change post-launch than any technical component.

By end of week three, your agent should pass 85% of scripted test cases. If it is below that, do not move to week four. Spend an extra two to three days on utterance tuning.


Week 4: Build Sprint Two (Edge Cases and Integrations)

Goal: Handle the 35% of calls that are not the happy path.

This is the week most pilots skip, and it is the reason most pilots fail in production.

Work through your containment risk register from week one. For each of the five failure points you identified, build explicit handling:

Complete your end-to-end integration testing this week. Run 100 test calls across all intents and edge cases. Target: 90% pass rate before moving to live traffic.


Week 5: Controlled Live Traffic

Goal: Validate performance on real calls before full rollout.

Route 10 to 15% of your target call type to the AI voice agent. Keep 85 to 90% on your existing human agent queue. This is not a soft launch. It is a controlled production deployment with active monitoring.

Monitor these metrics daily:

Hold a daily 15-minute standup with your technical lead, operations manager, and compliance stakeholder throughout week five. Issues surface fast on live traffic. You need a decision-making loop that is equally fast.


Week 6: Scale and Handover

Goal: Full production deployment with documented runbook.

If your week five metrics hit target, scale to 100% of the chosen call type in week six. If you are 5 to 10 percentage points below containment target, spend the first three days of week six on utterance and flow tuning before scaling.

Deliverables for week six:

What Realistic Results Look Like

Here is what a well-executed 6-week pilot delivers for a UK contact centre:

MetricPre-Pilot BaselineWeek 6 Target
Containment rate (target call type)0%60 to 70%
Average handle time4 to 7 minutes90 to 150 seconds
Cost per interactionGBP 4.50 to GBP 7.00GBP 0.40 to GBP 0.80
CSAT (AI-handled calls)N/A3.8 to 4.2 out of 5.0
Human agent capacity freed0%15 to 25% on target queue

These numbers are based on deployments we have run in UK financial services and utilities. They are not projections. They are outcomes.


The Three Mistakes That Kill Pilots

1. Piloting on too many call types at once. One call type. One agent. One set of success metrics. Scope creep in week two is the fastest route to a failed pilot. 2. Treating compliance as a week-five activity. In regulated UK industries, compliance review takes 5 to 10 business days minimum. If you start this conversation in week five, you will not go live in week six. 3. Building for the demo, not for production. A voice agent that works in a controlled test environment and breaks on live caller accents, background noise, or unexpected utterances is not a pilot. It is a prototype. Build for production from day one.

Ready to Run Your Pilot?

We build enterprise-grade AI voice agents for UK contact centres. Our team has delivered production deployments in financial services, utilities, and healthcare, all on AWS, all with compliance built in from the first line of configuration.

If you want a structured assessment of your contact centre's readiness and a clear path to production in 4 to 6 weeks, start with our scorecard.

Book a discovery call

Is your pilot going to reach production?

Fifteen questions, three minutes, no cost. You get a score against the ten checks we run every deployment through, and a straight answer on what is blocking yours.

Find out what is blocking you