guide
4 of 10 AI Support Vendors Won't Say If Your Data Trains Their AI
Two answer it properly, one has the answer behind a login, and the biggest names in the category do not address it in the documents where you would look.
Your support inbox is the most sensitive text your company holds. Order histories, addresses, complaints, refund arguments, occasionally a photo of somebody's damaged medication. It is other people's data, held on their behalf, and you agreed to look after it.
So: does any of it end up training your vendor's AI? On four of the ten platforms we checked, you cannot find out from anything they publish.
Two answer it thoroughly. Two say a flat no. One says its suppliers do not train on your data while never quite saying whether it does. One has written the answer down and put it behind a login. The remaining four do not address it in their AI terms, their DPA or their privacy policy.
There are two questions here and vendors answer the easier one
The first is whether the vendor trains its own models on your conversations. The second is whether the model provider underneath — OpenAI, Anthropic, Azure, Google — trains on them, because most support platforms do not run their own models and the AI layer is an API call to somebody else.
A vendor can truthfully say it does not train on your data while the layer it does not own remains an open question. Only three of the ten name a provider at all.
For fairness: the major API providers have defaulted to not training on business inputs since 2023. OpenAI's API terms say data sent through it is not used for training unless you opt in. So the exposure here is usually smaller than the silence suggests. The complaint is not that these vendors are secretly feeding your tickets into a model. It is that half of them make you take it on trust.
How this was read
For each platform, in this order: AI-specific terms, then the trust or security page, then the data processing agreement, then the privacy policy. We stopped at the first document that addressed AI training. Where none did, the row says not stated — an absence we went looking for, not one we gave up on.
No vendor was contacted. This is a study of what a buyer can establish alone, at eleven at night, before a call is booked. Several of these companies would no doubt answer immediately if asked, and that is the point: it should not require asking.
| Platform | Trains its own models on your data? | Model provider underneath trains on it? | Opt-out |
|---|---|---|---|
| Zendesk | Yes | No | None described |
| Freshworks | Yes | No | Yes — by emailing support. Opting out may reduce or disable some AI features. |
| HubSpot | Not stated | No | None described |
| eesel AI | No | No | None described |
| Chatbase | No | Not stated | None described |
| Ada | Behind a login | Behind a login | None described |
| Fin (formerly Intercom) | Not stated | Not stated | None described |
| Gorgias | Not stated | Not stated | None described |
| Tidio | Not stated | Not stated | None described |
| Front | Not stated | Not stated | None described |
READ FROM PUBLISHED AI TERMS, TRUST PAGES, DPAS AND PRIVACY POLICIES · ALL CHECKED 2026-08-06
Freshworks sets the standard, and it costs them nothing
Freshworks is the only one of the ten that offers a choice. It does not train third-party hosted models on customer data. It does train its own, on data from consenting customers, and splits that into two kinds: custom models trained only on your organisation's data and used only in your account, and collaborative models trained on de-identified data pooled across customers.
Either can be opted out of by contacting support, and the page states the trade-off rather than burying it: opting out may reduce performance in some features or turn them off.
That is what a complete answer looks like. It admits to training, explains what kind, offers an exit, and names the cost of taking it. It is also, notably, not a competitive disadvantage — nobody is switching away from Freshworks for being clear about this.
Zendesk is nearly as good, and the sanitisation detail is worth reading
Zendesk confirms it trains its own models on customer data and then does something unusual: it describes how the data is stripped first. Fields that are identifiers by definition, username and email, are excluded from the training set. A natural language pass removes further identifying information from free text. What survives is tokenised into vectors that are not human readable.
On the provider layer it is one sentence and it is the right sentence.
“No third-party will use your inputs to train their models or otherwise improve their services.”— Zendesk AI Data Use Information, read 6 August 2026
What is missing is a choice. Nothing describes an opt-out, so on this account training on your data is a condition of using the product.
HubSpot answers the question it was not asked
HubSpot is precise about its suppliers: it does not permit the AI service providers it engages to use your data for model training. It names OpenAI, AWS and Google, and commits to minimising retention, including zero-day retention where possible. That is better than most of this table manages.
It is also not an answer to whether HubSpot trains on your data. We could not find that addressed anywhere. The sentence is about providers, and it is easy to read as covering both if you are skimming — which is the problem with it.
Ada has written the answer down and made you ask for it
Ada's trust centre contains a document called AI Training Data. It requires authorisation to open.
This is a different failure from silence and in some ways a worse one. The work has been done, the position exists, somebody wrote it out — and reading it costs you a form and a wait. A buyer comparing five platforms on a Tuesday evening does not do that for one of them.
The flat noes, and the one that names its providers
eesel AI gives the cleanest answer of the ten: your data is never used for model training, and never included in any training data, period. It also names OpenAI, Anthropic and Google as the providers underneath, which almost nobody else does, and describes per-workspace isolation.
Chatbase says it does not train on your data and explains why it does not need to — it uses retrieval rather than fine-tuning, so your content is fetched at answer time instead of absorbed. It names no provider, so the layer underneath is unaddressed.
The four that say nothing
Fin, Gorgias, Tidio and Front do not address AI training in the documents where a buyer would look.
Fin's and Gorgias's DPAs are thorough about other things. Both prohibit selling personal data and restrict processing to delivering the service, which a lawyer might argue covers training by implication. Neither says so. Tidio's privacy policy contains exactly one sentence about model training and it is about Google Workspace APIs rather than Lyro. Front acknowledges that some of its service providers supply generative AI, then neither names them nor says what they may do with your data.
None of this is evidence of wrongdoing and this post does not suggest otherwise. It is evidence that the question was not anticipated, on products sold specifically for handling other people's personal information.
Zendesk
Publishes a dedicated AI data use page and confirms it trains its own models on customer data, with the sanitisation described step by step: identifier fields such as username and email are excluded from the training set outright, an NLP pass removes further identifying information from free text, and what remains is tokenised into vectors that are not human readable. On the layer underneath it is unambiguous — "no third-party will use your inputs to train their models or otherwise improve their services". No customer opt-out is described.
SOURCE: ZENDESK AI DATA USE INFORMATION, SUPPORT.ZENDESK.COM · CHECKED 2026-08-06
Freshworks
The fullest disclosure of the ten, and the only one offering a choice. Freshworks does not train third-party hosted models on customer data, and does train its own on data from consenting customers, splitting them in two: custom models trained only on your organisation's data and used only in your account, and collaborative models trained on de-identified data across customers. Either can be opted out of by contacting support, with the trade-off stated plainly.
SOURCE: FREDDY AI DATA AND SAFETY FAQS, FRESHWORKS SUPPORT · CHECKED 2026-08-06
HubSpot
Closes the provider question and leaves its own open: "We do not permit HubSpot AI service providers that we engage to provide the Subscription Service to use your data for model training." It names OpenAI, AWS and Google, and commits to minimising retention including zero-day retention where possible. What it does not say anywhere we could find is whether HubSpot itself trains on your data.
Providers named: OpenAI, AWS, Google.
SOURCE: HUBSPOT AI CLOUD INFRASTRUCTURE FAQ, KNOWLEDGE.HUBSPOT.COM · CHECKED 2026-08-06
eesel AI
The flattest no, and it names the layer underneath rather than talking around it: "Your data is never used for model training. It serves only your agents, only your team," and "your data is never included in any training data, period." Names OpenAI, Anthropic and Google as the providers, and describes per-workspace isolation with no cross-contamination between customers.
Providers named: OpenAI, Anthropic, Google.
SOURCE: EESEL.AI/SECURITY · CHECKED 2026-08-06
Chatbase
"We do not use your data to train AI models. We use Retrieval-Augmented Generation (RAG) to generate responses without compromising your data." No opt-out, because on this account there is nothing to opt out of. Names no model provider, so the layer underneath is unaddressed.
SOURCE: CHATBASE.CO PRIVACY POLICY · CHECKED 2026-08-06
Ada
Ada has an AI Training Data document. It sits inside a trust centre that requires authorisation, so the answer exists and a buyer cannot read it without asking. That is a different failure from silence and arguably a more frustrating one: the work has been done and the disclosure has been made conditional.
SOURCE: SECURITY.ADA.CX TRUST CENTRE · CHECKED 2026-08-06
Fin (formerly Intercom)
The DPA is detailed and prohibits selling personal data, sharing it for cross-behavioural advertising, and using it beyond the stated business purposes. None of that is a statement about model training, and nothing else in the document addresses it. Subprocessors are delegated to a separate page that changes over time.
SOURCE: INTERCOM.COM/LEGAL/DATA-PROCESSING-AGREEMENT · CHECKED 2026-08-06
Gorgias
Same shape as Intercom. Processing is restricted to providing the services, the subprocessor list lives elsewhere, and AI training is not addressed either way in the DPA.
SOURCE: GORGIAS.COM/LEGAL/DATA-PROCESSING-AGREEMENT · CHECKED 2026-08-06
Tidio
The privacy policy contains exactly one sentence about model training and it is about Google Workspace APIs, not Lyro. Whether visitor conversations feed model development is not addressed.
SOURCE: TIDIO.COM/PRIVACY-POLICY · CHECKED 2026-08-06
Front
The privacy notice acknowledges service providers that supply generative AI, which is more than most of the silent group manage, then neither names them nor says whether anything is trained on your data.
SOURCE: FRONT.COM/LEGAL/PRIVACY-POLICY · CHECKED 2026-08-06
What to ask, and what a good answer sounds like
Four questions, one email, before the trial ends:
- Do you train your own models on our conversation data? If yes, is it isolated to our account or pooled across customers?
- Which model provider handles inference, and is our data excluded from their training by contract rather than by their current default?
- Can we opt out, and what breaks if we do?
- How long is conversation data retained at every layer, including the provider's logs?
A good answer is Freshworks': yes to the first, here is the split, here is the opt-out, here is what it costs you. A bad answer is a link to a SOC 2 report, which describes how carefully data is handled and says nothing about what it is used for.
Check it yourself
Open the vendor's site and search its legal index for the word training. If the AI terms or trust page address it, the answer is usually one paragraph and easy to find. If you end up in a DPA reading about cross-behavioural advertising, the answer is not published, and you have learned something anyway.
Every row above is dated. If a vendor has published a clearer position since we read it, send it over and the record gets corrected and re-dated — that has already happened twice this month on other pages.
Frequently Asked
Does my customer data train the AI in my helpdesk?
It depends on the vendor, and on four of the ten we checked in August 2026 you cannot tell from anything they publish. Zendesk and Freshworks both confirm they train their own models on customer data, with Freshworks offering an opt-out by email and Zendesk describing how identifiers are stripped first. eesel AI and Chatbase say they do not train on your data at all. Fin, Gorgias, Tidio and Front do not address it in their AI terms, DPA or privacy policy.
Does OpenAI train on data sent through a helpdesk that uses it?
Not by default. OpenAI's API terms have said since 2023 that data sent through the API is not used to train its models unless you opt in, which is different from the consumer ChatGPT default. The gap worth closing is contractual: ask your vendor whether exclusion from provider training is written into their agreement, or whether they are relying on the provider's current default, which the provider can change.
Which AI support vendors say they do not train on customer data?
Of the ten we read, eesel AI and Chatbase state it flatly. eesel goes furthest by also naming OpenAI, Anthropic and Google as the providers underneath and stating that data is never included in any training data. Chatbase explains that it uses retrieval rather than fine-tuning, so content is fetched at answer time rather than absorbed into a model.
Can I opt out of my helpdesk vendor training on my data?
Only one of the ten publishes a route. Freshworks lets you opt out of either custom or collaborative model training by contacting support, and says plainly that doing so may reduce or disable some AI features. Zendesk describes no opt-out, so training appears to be a condition of use. The other eight either do not train, do not say, or do not publish the answer at all.
Is a SOC 2 report enough to answer this?
No. SOC 2 describes how carefully data is handled — access controls, encryption, monitoring — and says nothing about what the data is used for. A vendor can be fully SOC 2 compliant and train on every ticket you have ever received. Ask about purpose separately from security, and get the answer in the contract rather than in a trust badge.
Tools Mentioned
Full reviews, pricing tiers and where each one breaks.
Zendesk AI Agents
The default incumbent, now consolidating the category by acquisition. Strongest if you already live in Zendesk.
Freshdesk with Freddy AI
Cheaper than Zendesk with a comparable feature list. The trade shows up in depth rather than breadth.
eesel AI
Trains on your existing docs and tickets and works inside the help desk you already run, instead of replacing it.
Fin (formerly Intercom)
SalesforceThe most polished autonomous agent on the market, attached to the pricing model buyers complain about most — and now being bought by Salesforce.
You Can Also Look Into
29 of 78 AI Support Tools Can't Take PHI, or Won't Say
All 78 platforms classified by what they publish about HIPAA, plus what the Federal Register actually says about the encryption rule vendors are selling against.
35 of 47 Google Results for AI Support Are Vendor-Written
Six AI customer service buying searches, with every page-one result classified by who published it. Full result lists, method and screenshots.
20 Questions to Ask an AI Customer Service Vendor Before You Sign
The questions that determine your invoice, your exit, and whether the demo resembles what you will actually get — with what a good answer sounds like.
44 of 78 AI Support Tools Won't Sell to You Without a Call
We checked all 78 platforms in the directory for whether a buyer can sign up and pay without booking a demo. The answer splits the category almost exactly in half.
SOURCES
WRITTEN BY AR · UPDATED 2026-08-06
I read the fine print. Vendor pricing pages, billing definitions, terms, funding filings and acquisition notices — then I do the arithmetic nobody publishes: what a platform actually costs at your volume, what its headline metric is really counting, and who owns it now. I do not run benchmarks, and no page here pretends otherwise.