Every Vendor Quotes 90% Containment. Here's Why the Number Is Lying to You.
In 2026 every voice-AI vendor quotes a containment rate of 80–90%. But a 90% containment rate can sit right on top of a 40% resolution rate — and most decks never show the second number. A 'contained' call that produces a callback tomorrow was deflected, not resolved. Here is the operator's filter for grading a voice agent on the paired metric that keeps it honest, not the headline on the slide.
ScaleVoice
July 8, 2026 · 6 min read
Direct answer
Containment is the share of calls an AI handles without escalating to a human; resolution is the share where the customer's problem is actually solved and stays solved. They are not the same, and a platform can show 90% deflection with only about 40% true resolution. A call the bot 'contained' that generates a callback tomorrow was deflected, not resolved. To grade a voice agent honestly, never read a containment number without its partner metric — repeat-contact rate (the share who call back within 24–72 hours; benchmark under 15% at 72 hours) — and measure cost per RESOLVED call, not cost per call. Audited 2026 deployments put true containment in roughly the 58–85% band, not the 90% on the slide.
# Every Vendor Quotes 90% Containment. Here's Why the Number Is Lying to You.
Every vendor now quotes 90% containment. You should stop believing the number.
A 90% containment rate can sit right on top of a 40% resolution rate, and most decks will never show you the second one. That gap is the most important thing to understand about voice AI in 2026, and it's the thing the buying process is least equipped to catch.
Containment and resolution are not the same word
Containment is the share of calls the AI handled without escalating to a human. Resolution is the share of calls where the customer's actual problem got solved and stayed solved. A call the bot "contained" — kept away from a human — that generates a callback tomorrow was not resolved. It was deflected. And a deflection that comes back is worse than an honest transfer, because you paid to annoy the customer twice and you booked a win in the dashboard for it.
How an honest 90% happens
None of it requires anyone to lie. The vendor counts a call as "contained" the moment it ends without a transfer. The customer hangs up confused, calls back four hours later, gets a different session, and that second call is also "contained." Two contained calls, one unsolved problem, and a containment rate that looks spectacular. Industry analysts auditing real deployments in 2026 put *true* containment — resolved, not just un-transferred — in roughly the 58–85% band depending on call type, language mix, and how mature the analytics are. The headline on the slide and the number the operation actually lives with are almost never the same calculation.
The partner metric that keeps containment honest
If you've deployed a voice agent — or you're about to — the discipline is simple to state and annoying to enforce: never look at a containment number that isn't paired with its partner metric. The partner for containment is repeat-contact rate: the share of customers who call back within 24 or 72 hours about the same thing. A widely used benchmark is under 15% repeat contact within 72 hours. If containment is 90% and repeat contact is 30%, you don't have a 90% agent. You have a deflection machine with a good marketing team.
Cost per call is the wrong denominator
"$0.40 a call versus $7 for a human" is the line, and the automated portion really is 80–90% cheaper. But cost-per-*call* is the wrong denominator. The number that matters is cost-per-*resolved*-call. If your $0.40 call has to happen twice because the first one deflected, your real unit cost doubled and your customer's patience halved. Cheap-per-call and expensive-per-outcome is the exact trap the headline economics walk you into. Divide by resolutions, not by attempts.
Three questions to grade any voice-AI result
1. How do you define "contained," and does a callback within 72 hours reverse it? If a repeat contact doesn't claw back the win, the metric is scoring itself. 2. Show me containment and repeat-contact rate on the same chart, by call type. Aggregate numbers hide the call types where the agent is quietly failing. 3. **What's the cost per *resolved* contact, not per contact?** That single division separates a genuine automation from an expensive way to make customers call twice.
None of these are about the model, the voice, or the latency — the things demos are built to win on. All of them are about what happens *after* the call, which is the only place the truth lives.
The reason this matters beyond voice AI: 2026 is the year every operations team gets handed a dashboard full of confident percentages, and the percentages are increasingly self-graded. The skill that's about to be scarce isn't building the agent. It's reading its scorecard without being fooled by it — knowing which number has a partner that keeps it honest, insisting the definition is written down before the first call is placed, and refusing to celebrate any metric that can be gamed by making the customer do the work twice.
Next step
Turn this workflow into a scoped demo.
Bring the call source, booking rules, system destination, and exception path. ScaleVoice will map the first workflow that can produce a measurable booked outcome.
Book a demoRelated pages
FAQ
Questions buyers ask before scoping the workflow
What is the difference between containment rate and resolution rate?
Containment is the percentage of calls handled without human escalation. Resolution is the percentage where the customer's issue is actually solved and doesn't recur. A call can be "contained" and unresolved at the same time if the customer calls back later about the same problem.
What is a good repeat-contact rate for a voice agent?
A commonly cited benchmark is under 15% repeat contact within 72 hours. Always read repeat-contact rate alongside containment; a high containment paired with a high repeat-contact rate signals deflection, not resolution.
Why is cost per call misleading?
Because a cheap call that has to happen twice isn't cheap. Measure cost per resolved contact so that deflected calls, which generate a second attempt, are counted honestly against the total.
What true containment should I expect in production?
Audited 2026 deployments tend to cluster in roughly the 58–85% band once callbacks are counted against containment, rather than the 90% often shown on vendor slides. The exact figure depends on call type, language mix, and analytics maturity.