how-to

How to Choose an AI Customer Service Platform - 9-Point Checklist (2026)

Nine questions that eliminate most of the shortlist in an afternoon

By AR · Published 29 July 2026 · 8 min read

Most buying guides for this category compare feature checklists. Every platform ticks nearly every box, which is why the checklists are useless and why people end up choosing on the demo.

The things that genuinely differ are pricing model, what triggers a charge, whether the agent can act on your systems, and who owns the company. None appear in a feature grid.

Start by ruling things out

A shortlist of twenty is a shortlist of none. These questions eliminate most of the market quickly, in rough order of how much they narrow it.

1. Does it need to act, or only answer?

If your highest-volume ticket is order status, you need a platform that can query your commerce system. That single question removes most cheap tools, which answer from ingested help articles and nothing else.

2. What pricing model suits your volume?

VolumeBest modelAvoid
Under 2,000/moPer resolution or flatPer seat if team is large
2,000-10,000/moPer seatPer resolution
Over 10,000/moPer seat or flatPer resolution, firmly
Many agents, few customersBy contact (Help Scout)Anything per seat

At 10,000 monthly resolutions, per-resolution pricing runs $9,900 against roughly $600 for ten Freshdesk seats. The gap is not marginal.

3. What triggers a billable event?

Get this in writing before the trial ends. A platform charging $0.69 that bills on customer non-reply can cost more than one charging $0.99 that bills only on confirmed resolution.

4. Do you need the help desk too, or just the agent?

If you already run Zendesk or Freshdesk and only the AI economics annoy you, layering eesel or My AskAI on top avoids a migration entirely. Migration is usually the largest hidden cost in this category.

5. Can you evaluate it without a salesperson?

Roughly half the platforms here are sales-led with no published pricing. That is not disqualifying, but it costs you weeks and biases your shortlist toward whoever publishes.

6. Who owns them?

Five platforms in this category have been acquired since 2024. Forethought closed into Zendesk fifteen days after announcement and no longer exists standalone. If you are signing three years, ask.

7. What happens when it does not know?

Ask what the agent does below its confidence threshold, then test it with a question your documentation does not answer. A tool that says it is unsure is far more useful than one that invents a policy.

8. What does the human see after escalation?

Escalate deliberately during the trial. If context does not survive, you have added a step rather than removed one.

Start from your ticket mix, not a feature list

Every buying guide in this category opens with a feature matrix. Features should be the last thing that narrows your list, because most of these tools do broadly the same things and the differences that decide outcomes are structural.

Four questions about your own queue eliminate most options before you open a single vendor page.

QuestionIf yesIf no
Does your top ticket category need a system lookup?Gorgias, Fin, Ada, RichpanelChatbase or Tidio at a tenth the price
Do occasional agents outnumber full-time ones?Help Scout, HappyFoxPer-seat is fine
Will volume triple within 18 months?Per-seat, or negotiate a volume tier nowPer-resolution is fine
Is there a compliance review?Ada, Lorikeet, DecagonAnything self-serve

The first row saves the most money. If your questions are genuinely documented and informational, the expensive tools resolve no more than the cheap ones and you are paying for capability you cannot use.

The three-week evaluation

  • Week one: audit rather than demo. Sort last month's tickets by whether the answer exists in writing anywhere. That share is your ceiling and it governs everything after it.
  • Week two: two trials in parallel, run on your own historical tickets — not the vendor's demo questions, and including the awkward ones you would rather not show them.
  • Week three: billable-event definition in writing from both, cost modelled at three times current volume, and the renewal rate.

Most evaluations invert this, opening with demos and arriving at pricing last. Doing the audit first is what stops a team buying a $0.99-per-resolution agent to answer questions nobody ever wrote down.

The pricing structures, and what each one punishes

Most buying guides compare rates. The structure matters more, because it determines what happens as your deployment succeeds.

StructureBill rises withPunishes you for
Per seatHeadcountHiring
Per contactDistinct customersGrowing your customer base
Per ticketVolumeAnything that increases contacts
Per sessionTrafficTraffic spikes
Per resolutionAI successImproving the AI

The bottom row is the one to think hardest about, because it is the only structure where the thing you are optimising and the thing you are paying for are the same. Every improvement you make to documentation, scope or confidence raises the invoice.

That is not a reason to avoid it. It is a reason to model cost at the resolution rate you are aiming for rather than the one you have, which is the opposite of how most evaluations are run.

Two questions that eliminate most of the market

Before any demo, answer these about your own queue. They narrow a list of forty to a list of four.

  • Does your highest-volume ticket category need the AI to read something only your systems know? If yes, every document-answering tool is eliminated regardless of price. If no, every expensive agent is eliminated regardless of capability.
  • Do your occasional agents outnumber your full-time ones? If yes, contact-based or unlimited-seat pricing beats per-seat by a factor of several, and that difference is larger than any feature gap in the category.

Both are answerable in twenty minutes from data you already have. Neither appears on a vendor's comparison page, because both have the effect of removing most vendors from consideration.

9. What does it cost at three times your volume?

You are buying this to grow into. Model the tier you will be on in eighteen months, including overages. Gorgias at 6,000 tickets costs 40% more than at 5,000.

Criteria here come from published pricing, billing terms and public reporting. We do not benchmark platform performance.

Frequently Asked

What is the best AI tool for customer service?

There is no single answer, and any page giving one is usually selling something. The decision turns on volume, whether you need order actions, and whether you already have a help desk you like.

How do I compare AI customer service platforms?

On pricing model at your volume, what triggers a billable event, whether the agent can act on your systems, and vendor ownership. Feature grids barely differentiate them.

What should I ask in a demo?

Ask about an order that does not exist, ask what happens below the confidence threshold, and escalate deliberately to see what the human receives.

How much should AI customer service cost?

Between $7 per user and $1.50 per resolution depending on model. At 10,000 monthly resolutions those two extremes are roughly $600 and $15,000.

Do I need to replace my help desk?

Usually not. eesel and My AskAI layer onto Zendesk, Freshdesk and others, which avoids migration, the largest hidden cost here.

Should I choose a self-serve or sales-led platform?

Self-serve if you want to evaluate this month. Sales-led platforms are often more configurable but cost weeks and publish no pricing.

How important is vendor independence?

More than most buyers weight it. Five platforms here have been acquired since 2024 and one no longer exists as a standalone product.

What is the most common buying mistake?

Choosing on conversation quality in a demo, then discovering the pricing model punishes the volume you actually have.

Can I negotiate AI support pricing?

On sales-led platforms, yes, and you should. On published per-resolution rates, rarely.

How long should evaluation take?

Two weeks on a self-serve tool with real tickets beats two months of demos. Sales-led evaluation realistically runs six to twelve weeks.

Tools Mentioned

Full reviews, pricing tiers and where each one breaks.

You Can Also Look Into

WRITTEN BY AR · UPDATED 2026-07-29

I read the fine print. Vendor pricing pages, billing definitions, terms, funding filings and acquisition notices — then I do the arithmetic nobody publishes: what a platform actually costs at your volume, what its headline metric is really counting, and who owns it now. I do not run benchmarks, and no page here pretends otherwise.

Editorial policy