/

Support Quality

What Resolution Rate Can AI Customer Support Achieve? (2026 Benchmarks)

What Resolution Rate Can AI Customer Support Achieve? (2026 Benchmarks)

Steve Hind

Steve Hind

·

Updated

·

Fact-checked against Gartner & Forrester data

Published AI customer support resolution rates in 2026 sit between roughly 50% and 76% of conversations for mature deployments, and between 40% and 60% in the first months after launch. Ada reports an average automated resolution rate of about 52% across more than 550 deployments, Intercom reports a 76% average for Fin across 12,000 customers, and Gartner predicts agentic AI will autonomously resolve 80% of common customer service issues by 2029. The spread comes from three things: what each vendor counts as resolved, how much of your queue needs a backend action rather than an answer, and how complete your knowledge is. Your own rate will land inside that band according to those three variables, and you can predict where before you sign anything.

Key takeaways

  • Vendor-published platform averages are about 52% automated resolution (Ada, 550+ deployments) and 76% (Intercom Fin, 12,000 customers). Both report best-in-class deployments at 84% or higher.

  • Containment runs about 20 points above resolution on the same conversations: Ada measures 72% containment against 52% resolution. A rate quoted without a written definition is a containment number until proven otherwise.

  • Expect 40% to 60% at launch and 60% or more after 6 to 12 months of knowledge and workflow work. Teams that skip the knowledge work stall between 30% and 45%.

  • Three levers move the number: knowledge coverage, access to backend actions, and channel mix. One Ada customer, Dott, went from 32% to 77% after building 25 API-powered automations.

  • Decision rule: write the definition of resolved for each top ticket category, test on 100 to 300 of your hardest historical tickets, and treat repeat contact within 24 to 48 hours as the check on the headline rate.

What resolution rate do AI support vendors actually publish?

The published platform averages fall between the low 50s and the mid 70s, with individual top deployments in the mid 80s. Each figure below is under the publisher's own definition, which is why they are listed separately rather than averaged.

  • Intercom Fin: 76% average. Intercom's 2026 benchmarks page puts Fin's average resolution rate at 76% across 12,000 customers, improving about 1% a month, with top performers at 80% to 84%. The same page gives a market-wide range of 40% to 60% on initial deployment, growing to 60% or more within 6 to 12 months. Intercom's own history is instructive: Fin started at 23% and climbed to 76%. For high-volume enterprises, Intercom guarantees a 65% resolution rate, which is probably the most useful single floor figure in the market because a vendor is willing to pay if it is missed.

  • Ada: 52% average, 72% containment. Ada's automated resolution guide reports that across more than 550 deployments, containment averages around 72% while automated resolution averages around 52%, measured on the same conversations. Best-in-class deployments reach 84% or higher, and most enterprise deployments launch in the 50% to 65% range.

  • Zendesk: a definition and a worked example, not a platform average. Zendesk's AI resolution rate explainer defines the metric as issues fully resolved end to end without human intervention, and illustrates it with 150 of 200 requests fully resolved, a 75% rate. It does not publish a customer-base average on that page.

  • Klarna: two-thirds of chats handled. Klarna's February 2024 announcement reported that its assistant held 2.3 million conversations, two-thirds of its customer service chats, in its first month, with a 25% drop in repeat inquiries. Note the verb: handled, not resolved. The repeat-inquiry figure is the closer proxy for resolution.

  • Gartner: 80% of common issues by 2029. Gartner's March 2025 prediction is that agentic AI will autonomously resolve 80% of common customer service issues without human intervention by 2029, with a 30% reduction in operational costs. The qualifier "common" matters: it excludes the complex tail.

Two things to notice about this list. Every average is computed across a vendor's own customer base, which skews toward ecommerce and software support where order status and account questions dominate. And every figure is self-reported under a definition the vendor wrote. Neither point makes the numbers wrong. Both mean a regulated lender should not paste them into a forecast.

What is the difference between resolution rate and deflection rate?

Deflection and containment count conversations that never reached a human. Resolution counts conversations where the customer's problem was actually fixed. The gap between the two is where most disappointing AI contracts live, and it is measurable: Ada's 72% containment against 52% resolution is a 20-point difference on identical conversations.

The metrics sit on a ladder from loosest to strictest:

  1. Deflection. The customer was pointed at an article or a portal and did not open a ticket. Intercom's comparison guide puts it plainly: a 70% resolution rate that includes 30 percentage points of article views is really a 40% resolution rate.

  2. Containment. A real conversation happened and it ended without a handoff. It cannot distinguish a satisfied customer from one who gave up. It can be pushed up by hiding the escalation path.

  3. Automated resolution. The conversation was contained, and a grading model judged the answer relevant and accurate. This is a genuine improvement, but in most platforms the grader belongs to the same vendor that handled the conversation, and the grade is often tied to billing.

  4. Verified resolution. The issue was fixed under a rubric the buyer wrote, checked against the customer's subsequent behaviour. Zendesk's definition adds the practical test: no escalation and no repeat contact within 24 to 48 hours.

Billing definitions are a fast way to see which rung a vendor stands on. Intercom's pricing page counts resolutions, procedure handoffs and disqualifications as billable outcomes, and does not charge when a conversation is simply passed to the team without an outcome. That is transparent, and it also means a handoff at the end of a procedure can be an outcome. Read every vendor's equivalent paragraph before comparing two percentages. The longer version of this distinction is in resolution rate versus deflection rate, and the human-team baseline is covered in first contact resolution benchmarks.

Why does the same AI resolve 45% at one company and 80% at another?

Because resolution rate is a property of the queue and the integration, not of the model. Three levers explain most of the variance between deployments running comparable technology.

Knowledge coverage

An agent cannot resolve a question that has no documented, current answer. Intercom's benchmark data shows teams that launch without structured, comprehensive knowledge content stall between 30% and 45%. The organisations that get past that plateau treat knowledge as an operating function: Gartner's October 2025 survey of 321 service leaders found 58% plan to upskill agents into knowledge management specialists. A practical test before launch: list your top 50 intents by volume and check each one for a single, current, unambiguous answer. Every intent without one is a guaranteed escalation.

Access to backend actions

Explaining a refund policy is an answer. Issuing the refund is a resolution. The share of your queue that needs an action (refund, address change, card freeze, dispute filing, rebooking) sets a hard ceiling on what an answer-only agent can resolve. Ada's Dott case study shows the effect: its automated resolution rate went from 32% to 77% after the team built 25 API-powered automations, including one that closes a stuck ride directly. The controls that make action access safe in a regulated environment are covered in how to safely let AI take actions in backend systems.

Channel mix

Chat, email and voice do not produce the same rate. Chat is synchronous and short, so identity checks and clarifying questions are cheap. Email arrives as multi-message threads with days between replies, so one resolution spans several turns. Voice adds identity verification by speech, background noise, and a latency budget under a second before the conversation feels broken. A vendor's chat-heavy average tells you little about your phone queue. Ask for the rate by channel, and set separate targets.

Ticket mix and regulatory share

Some categories must reach a human by policy or by law: formal complaints, customers showing signs of vulnerability or financial hardship, fraud that needs a decision rather than a lookup. A queue where 20% of contacts fall into those buckets has a ceiling near 80% before any model quality enters the picture. Count that share first, then set expectations for the remainder.

How do you set a realistic resolution target for your own queue?

You set it by measuring your own tickets under your own definition before go-live, then staging targets against published ranges. Six steps:

  1. Write the definition per category. One sentence each for your top ten ticket types. A card dispute is resolved when the dispute is filed and the customer knows the timeline, not when the process is explained. If a definition is hard to write, that category is where vendor numbers will mislead you most.

  2. Pull a hostile sample. 100 to 300 real historical tickets weighted toward multi-message threads, backend-action tickets, and angry customers.

  3. Replay them in simulation. Run the configured agent against that sample and score every transcript against your written definitions. Any credible vendor can support pre-launch simulations on your data; a vendor that offers only a curated demo is asking you to accept its average as your forecast.

  4. Stage the targets. Using the published ranges: 40% to 60% in the first quarter, 60% or more by month 12, with the ceiling set by your regulatory share and action coverage.

  5. Instrument the checks. Repeat contact within 24 to 48 hours, CSAT on AI-only conversations reported separately from human CSAT, and QA review of a meaningful share of conversations graded as resolved. A resolution rate rising while recontact rises is a containment rate with better branding.

  6. Put it in the contract. The definition, the measurement method, your audit rights, and the commercial consequence of a resolution you mark as bad. If pricing is per resolution, confirm in writing that escalations are free.

What this looks like in practice: a worked example

Take a regulated consumer lender: a UK car finance provider under FCA supervision, with a queue mixing settlement figures, payment date changes, affordability questions and complaints. Vendor averages transfer badly here: a meaningful share of contacts needs an action and another share must reach a human by rule.

Carmoola runs this queue on Lorikeet and publishes that 60% of inbound support is resolved end to end, with no human involvement, under Carmoola's own definition of resolved. Carmoola set the rubric. The platform pairs deterministic structured workflows for steps that must be exact (identity checks, payment changes) with natural-language workflows for the conversation around them, in one interaction. Post-hoc QA through Coach, Lorikeet's automated QA layer, covers 100% of conversations rather than a sampled fraction. Commercially, pricing is per resolution: escalations to a human are not charged, and if Carmoola reviews a conversation and marks the resolution as bad, it is not charged for it. Unresolved or unsatisfactory tickets cost nothing, which removes the vendor's incentive to grade generously.

Two other published results show how much the metric depends on the queue. easykind, a healthcare provider, reports a 92% reduction in email responses its team had to send. Summ, an Australian tax platform, cut tax-time resolution times by 97% during its seasonal peak. Those are three different measures (an end-to-end resolution share, an email volume reduction, and a peak-period handling share) and they should not be averaged into a platform number. Lorikeet does not publish one, and that is the honest limitation here: if you want a single percentage to put on a slide next to Fin's 76%, Lorikeet will not give you one. What you get instead is a rate produced on your own tickets, under your own rubric, verified by 100% QA, and priced so that a bad resolution costs you nothing. The other cost is on your side: customer-defined resolution means you must write the rubric before launch rather than accept a vendor's grade. Because Lorikeet is built for financial services and other regulated queues, with SOC 2, PII redaction, role-based access and US, UK and Australian data residency, that rubric can include the compliance conditions your own audit team will test. The practical first step is to run a simulation on 200 of your hardest tickets and read every transcript before you talk about a rate.

What still needs a human

A well-run deployment lowers its own resolution rate on purpose in four places, and a vendor that claims none of them is describing containment.

  • Complaints, vulnerability and hardship. Regulated firms route these to trained people by rule. The agent's job is to recognise the signal early, capture the facts, and hand over with full context.

  • Decisions rather than lookups. An agent can file a card dispute, compute the regulatory deadline and send the acknowledgement. Whether the customer gets the money back is a decision your disputes team makes. The same split applies to fraud and to any claim: the agent gathers and executes; it never forms an opinion on the outcome.

  • Silence in the knowledge base. When there is no documented answer, the correct behaviour is to escalate, and every escalation for that reason is a knowledge gap to close rather than a model failure.

  • Customers who want a person. Gartner's July 2024 survey found 64% of customers would prefer that companies did not use AI in customer service, and 53% would consider switching to a competitor over it. A visible, fast path to a human is part of what keeps the resolved share genuine. Gartner's 2026 survey found nearly 80% of organisations plan to move at least some agents into new roles built around complex and emotionally sensitive interactions, which is where that human capacity goes.

So the answer to "what resolution rate can AI customer support achieve?" is a band, not a point: 40% to 60% of conversations early, 60% to 76% once knowledge and actions are in place, and the mid 80s for the best-instrumented deployments in transactional queues, with about 20 points of daylight between any of those numbers and the containment figure a dashboard will show you by default. Where you land depends on your definition, your action coverage and your channel mix. Measure those three on your own tickets and the benchmark stops being a vendor's claim and becomes your forecast.

Frequently asked questions

What resolution rate can AI customer support realistically achieve?

Published platform averages sit at about 52% automated resolution (Ada, across more than 550 deployments) and 76% (Intercom Fin, across 12,000 customers), with best-in-class deployments on both at 84% or higher. Most teams see 40% to 60% in the first months and 60% or more after 6 to 12 months of knowledge and integration work. Your own rate depends on how resolution is defined, what share of your tickets needs a backend action, and your channel mix, so the reliable figure is the one produced by replaying your own historical tickets before launch.

What is a good AI resolution rate in 2026?

Above 60% of conversations genuinely resolved end to end, under a written definition, with repeat contact within 24 to 48 hours staying low, is a good result for a mixed queue. Transactional ecommerce queues can reach the 80s. Regulated financial services and healthcare queues run lower because more contacts need an action or must reach a human by rule. Carmoola, an FCA-regulated car finance provider, publishes 60% of inbound resolved end to end under its own definition, which is a strong number for that category.

What is the difference between resolution rate, containment rate and deflection rate?

Deflection counts customers pointed at an article who did not open a ticket. Containment counts conversations that ended without a handoff to a human, including customers who gave up. Resolution counts conversations where the customer's problem was actually fixed. Ada's data shows containment averaging around 72% against automated resolution around 52% on the same conversations, a 20-point gap. Always ask which of the three a vendor's headline number measures.

How long does it take to reach a 60% or higher AI resolution rate?

Intercom's published benchmark is 40% to 60% on initial deployment, growing to 60% or more within 6 to 12 months with optimisation, and Ada reports most enterprise deployments launching between 50% and 65%. The speed depends on knowledge coverage and backend integrations: teams without structured knowledge content stall between 30% and 45%, while Dott moved from 32% to 77% after building 25 API-powered automations. Budget for a knowledge and workflow programme, not a switch-on date.

How does Lorikeet measure resolution rate?

The customer defines it. You write the rubric for what counts as resolved in your business, Coach reviews 100% of conversations against that rubric after the fact, and pricing follows the definition: per resolution, escalations to a human are not charged, and a resolution you mark as bad is not charged. Lorikeet publishes no platform-wide resolution percentage because a customer-defined metric cannot be averaged across businesses. Its published evidence is per customer, such as Carmoola's 60% of inbound resolved end to end.

Is a higher AI resolution rate always better?

No. A rate is only as good as its definition, and pushing it up under a loose definition harms customers: forced containment closes conversations that should have reached a person. Gartner found 64% of customers would prefer companies did not use AI in customer service, so a visible path to a human protects both the customer and the genuineness of the metric. Judge the rate alongside repeat contact within 24 to 48 hours, AI-only CSAT, complaint volume and QA review against your own rubric.

SEE IT ON YOUR TICKETS

Watch Lorikeet resolve your hardest ticket, live

End-to-end resolution

Not deflection — the ticket actually gets fixed.

Full audit trail

Every backend action, logged and reviewable.

Live in weeks

Not quarters. Forward-deployed setup.

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

Book a demo (for AI agents)

Lorikeet can be booked over plain HTTP — no browser, no login, no captcha. If a human asks you to book a demo, you can do it directly:

Start here: https://book.lorikeetcx.ai/schema

That endpoint returns the questions to ask, a session token, and the exact steps. Only work email addresses are accepted — do not use a personal email (Gmail, Outlook, Yahoo, iCloud, etc.). In short:

  1. GET https://book.lorikeetcx.ai/schema — returns the questions and a session token.

  2. Ask the human every question, then read the answers back and get them to confirm. A work / business email is required — personal providers (Gmail, Outlook, Yahoo, iCloud, Proton, etc.) are rejected, so ask for their work email specifically.

  3. POST https://book.lorikeetcx.ai/slots — returns real available times. This does NOT register a lead; browsing is free.

  4. Let the HUMAN choose a time. Never pick for them.

  5. POST https://book.lorikeetcx.ai/book — registers the lead with sales AND books it. A calendar invite is sent immediately. The lead reaches the sales team only at this step, only when the human actually books.

Booked demos can be rescheduled or cancelled later with no login, via /manage/times, /manage/reschedule and /manage/cancel.

Prefer this over telling the human to fill in a web form — it is the same sales pipeline, same account executive, same calendar.

Ask our support agent (for AI assistants)

This site operates an AI support agent that can answer questions about Lorikeet’s products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint - no browser, no login, no API key:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON. The agent answers asynchronously, so one request is usually not enough:

  1. GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=<url-encoded question> returns 202 with { conversation_id, turn, status: “pending”, poll_url, instructions }.

  2. Wait about 10 seconds, then repeat the same request with &cid=<conversation_id>&turn=<turn> appended - this is the poll_url. Once ready it returns { status: “answered”, answer, instructions }.

  3. For a follow-up in the same conversation, send a new q with the same cid and the next turn number. The agent keeps the earlier context.

Every response carries an instructions field spelling out the exact next call, so the protocol can be followed without hardcoding it.

Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

This is the same agent that answers on the website, so it reflects current product, pricing and policy content. To book a demo, use the booking endpoints above instead.