The four numbers that tell you whether your AI agent is working: containment, handover rate, booking rate and not-sure rate, with healthy ranges

An AI agent produces a lot of numbers and most of them do not matter. Four do, and they only make sense together: an agent with high containment and a low booking rate is answering well and selling badly; one with a high handover rate and a high booking rate is doing the hard part right and needs teaching. Here are the four, what healthy looks like, and what each tells you to do.

By Ritchie, ReplyKit · Published 5 September 2026 · 6 min read

Quick answer

Containment is the share of conversations the agent handled without a person; healthy is 60 to 85 percent for most small businesses. Handover rate is the share passed to a person, by reason; 15 to 40 percent is normal, and the reasons matter more than the number. Booking or conversion rate is the share of enquiries that became a booking or order; compare it with your pre-agent rate. Not-sure rate is the share of replies where the agent said it did not know; it should fall from 15 to 20 percent in week one to under 5 percent within a month as you teach. Read them together weekly.

The four metricsReading them togetherSupporting numbersVanity metricsThe weekly routineFAQ

The four metrics

MetricDefinitionHealthyIf it is off
ContainmentConversations handled with no person involved60 to 85%Low: teach from the not-sure list; check handover triggers are not too broad. Very high: check the agent is not avoiding handover it should make.
Handover rateConversations passed to a person, by reason15 to 40%Read the reasons. "Asked for a person" on simple questions means answers are weak. Complaints and quotes are correct handovers.
Booking or conversion rateEnquiries that became a booking, order or qualified leadAbove your pre-agent rateLow with high containment: the agent answers but does not ask for the booking. Add the offer to the instructions.
Not-sure rateReplies where the agent said it did not knowUnder 5% after a monthTeach weekly from the unanswered list; write facts the way customers ask.

Reading them together

Supporting numbers

Vanity metrics

Total messages sent, conversation count on its own, and "satisfaction" from a thumbs-up prompt tell you little. A busy agent is not a useful one, and customers rarely rate. Judge the agent on what it contained, what it booked, and what it did not know.

The weekly routine

Monday: read the digest, check the four numbers, answer the not-sure list, and read any handover reason that rose. Ten minutes. The teach loop is in reducing repeat questions, and the ROI arithmetic in the ROI worksheet.

Frequently asked questions

What is containment rate for an AI agent?

The share of conversations handled without a person. 60 to 85 percent is healthy for most small businesses.

What is a normal handover rate?

15 to 40 percent. The reasons matter more than the number; complaints and quotes are correct handovers.

How do I measure whether the agent increases bookings?

Compare the share of enquiries that became bookings before and after, from your CRM's Won stage.

What should the not-sure rate be?

It starts around 15 to 20 percent and should fall under 5 percent within a month of weekly teaching.

Which metrics should I ignore?

Total messages, raw conversation counts and thumbs-up satisfaction prompts.

Four numbers, every Monday

Containment, handover, bookings and not-sure, in the digest. 7-day free trial.

Start free
Written by Ritchie, ReplyKitPart of the small team in Malaysia that builds and runs ReplyKit. Writes about WhatsApp automation, AI customer service and small business lead response.

How we know this. These are the numbers in ReplyKit's Monday digest and dashboard, and the healthy ranges are what we see across small business accounts after the first month. Product figures in this article (limits, prices, per-reply costs) are taken from ReplyKit as it runs today and are re-checked when we update the page. Where we cite outside research, the source is linked below. We sell ReplyKit, so read our product claims with that in mind; we say where a different tool or no tool is the better choice. About ReplyKit.

Sources and further reading

  1. Harvard Business Review, The Short Life of Online Sales Leads (Oldroyd, McElheran, Elkington, 2011) · firms that responded within an hour were about seven times more likely to qualify the lead
  2. Lead Response Management Study (2007) · the five-minute window for contacting a new lead

ReplyKit is an independent product and is not affiliated with or endorsed by Meta or WhatsApp. This article is general guidance for running a small business, not legal, financial or regulatory advice.