Part: One
Here’s an uncomfortable fact about the AI you use every day: most of it reads your chats for practice. Not by default-in-theory β by default-in-practice, on the plans people actually buy. And almost nobody ever touches the setting.
Open a fresh ChatGPT or Copilot or Gemini account and look at the fine print before the hype: in almost every case, your conversations are fair game for training the next model β unless you flip an opt-out well hidden in the settings. The industry calls it “training on-by-default.” It’s the single most important thing your AI provider’s privacy policy doesn’t want you to read.
This is the first of three pieces. Here we audit the big players β on their own words, not marketing. In Part 2, we walk the spectrum to the providers that genuinely don’t train on you, including Venice and the rare few built around privacy from the ground up. In Part 3, we get hands-on with the one option that is 100% private by nature: running the models on hardware you control.
One frame before we start. Every claim below is taken from the provider’s own privacy or help-center pages, and every price is taken from the provider’s pricing page (or a reputable current source where the page is login-walled), obtained August 2026. Pricing changes; the “who trains on you” defaults have been remarkably stable.
The core pattern: on by default, opt-out hidden
Consumer AI is overwhelmingly built on one design: training on your data is switched on for you, and turning it off is a chore. Read across the majors, a consistent picture emerges:
- The free tier and everyday paid plans train on your chats by default.
- Opting out is usually possible for the consumer tiers β but it’s a settings-page hunt, and most users never do it.
- The business and enterprise tiers are the exception: as a rule they do not train on your data (you’re paying for it).
- A handful of services are the genuine reverse: privacy by design, where your content is never stored or trained on in the first place (that’s Part 2).
That’s the consent gap. Not “we don’t use your data” β “we use your data unless you navigate four menus to stop us.”
The majors, on their own terms
OpenAI / ChatGPT β trains by default on consumer plans
OpenAI’s policy is blunt: for individuals using ChatGPT, Sora, or Operator, “we may use your content to train our models” β on by default, with an opt-out in the privacy portal (the ChatGPT pricing page’s own comparison table confirms “Content is used to train our models: Opt-out available” for all consumer plans). Temporary Chats don’t train; the business line (ChatGPT Team, Enterprise, API) is opted out by default. Prices obtained August 2026: ChatGPT Go $8/month, Plus $20/month, Pro $100 or $200/month on the ChatGPT pricing page.
Anthropic / Claude β the privacy outlier that softened its stance
Anthropic built its reputation as the privacy-conscious major: for years consumer conversations were not used to train its models and were deleted after 30 days. In 2025 that changed. Anthropic’s own policy language is now consent-based β it uses consumer chats and coding sessions “if you choose to allow us to” (policy effective September/October 2025), retaining that data for up to 5 years if you opt in, or 30 days if you don’t. The practical default is genuinely disputed: several legal and compliance analyses describe the consumer toggle as on by default (you’d have to turn it off in Settings β Privacy), while others describe it as an opt-in choice at signup. Bottom line: consumer Claude is no longer the clean opt-in exception it once was β check your own toggle.
What still genuinely differentiates Anthropic:
- Incognito chats are never used for training, even if Model Improvement is switched on.
- The safety carve-out is real and unusually transparent: conversations flagged by safety classifiers may still be used to improve internal trust-and-safety models even if you opt out β with flagged inputs/outputs reportedly retained for up to ~2 years and classification scores for up to ~7 years.
- The API offers one of the strongest defaults in the market: 7-day log retention with no training (reduced from 30 days in September 2025).
- Claude for Work and the API are covered by commercial terms and do not train by default.
- Claude was also among the first majors to attach machine-readable AI watermarks to its output as the EU’s AI transparency rules took effect in August 2026.
Prices obtained August 2026: Claude Pro $20/month ($17/month billed annually), Max $100 or $200/month on the Claude pricing page.
Google Gemini β activity on by default, human reviewers, 18-month auto-delete
Google’s Gemini Apps Privacy Hub says Keep Activity is on by default and chats may be used to train generative AI models, reviewed by human reviewers (retained up to 3 years, disconnected from your account). Turn Keep Activity off and future chats won’t train; the default auto-delete is 18 months. Even with it off, Google still uses chats to respond and protect users. Prices obtained August 2026: Google AI Plus $4.99/month, Pro $19.99/month, AI Ultra $99.99 or $199.99/month (the top tier was cut from $250 at I/O 2026) on the Google AI plans page.
Microsoft Copilot β consumer on-by-default, business never
Consumer Copilot conversations are saved by default (18-month storage) and used for AI training on by default, with an opt-out for logged-in users. The enterprise Copilot (Microsoft 365 Copilot, Entra ID) is not used to train AI models. Worth knowing the line moved: the standalone consumer Copilot Pro was retired and folded into Microsoft 365 Premium, and the business Copilot tiers start around $18β21/user/month billed annually (the $18 entry rate is promotional through September 2026 β regular $21; up to ~$32/user/month for the fullest tier) on the Microsoft 365 Copilot pricing page (prices obtained August 2026).
Meta AI β trains on what you share with it; your DMs stay private
Here’s the surprisingly good news from Meta: when you use Meta AI inside WhatsApp, Meta can only read messages that mention @Meta AI or that you choose to share with it β your personal messages and calls stay end-to-end encrypted. Your interactions with Meta AI itself, plus data from across your Accounts Center, may be used to improve AI at Meta by default. There’s no paid consumer tier β Meta AI is free.
xAI / Grok β trains on your content by default
Grok’s privacy policy (effective April 2026) says your User Content and outputs are used to train its models by default, along with public X posts and internet data. The genuinely private escape hatch is Private Chat, which won’t appear in history and is deleted from xAI systems within 30 days. Prices obtained August 2026: Free, SuperGrok $30/month, SuperGrok Plus $100/month, and Business at $30/user/month for teams (Enterprise custom) on the xAI pricing page.
DeepSeek β trains by default, stores in China
DeepSeek collects prompts, files and chat history to “train and improve” its models, and its policy states personal data is processed and stored in the People’s Republic of China for as long as you keep an account. That also means PRC legal jurisdiction applies to how that data is handled β material for many readers. There’s a right to opt out and a reduced ability to control processing compared to Western providers. DeepSeek has no paid consumer tier β the app is free.
Mistral / Vibe β cheaper, and trains by default below the business line
The French provider (now branding its assistant “Vibe”) is a budget pick: Le Chat Pro runs $14.99/month (obtained August 2026, Mistral pricing page). But the privacy story splits by plan: Free, Pro, and Education plans use your input and output to train by default unless you opt out, while Team and Enterprise plans are not trained on.
Why you should care, not just feel annoyed: every time your chats feed training, they train on you β your private questions, your confidential documents, your half-formed ideas. For a business, using a train-on-by-default tool can leak what should stay inside the company. The privacy difference isn’t cosmetic; it’s a real risk pass and a real trust question for whoever’s AI you choose.
The quick reference: who trains, who opts out, who never
| Provider | Consumer default | Business / enterprise |
|---|---|---|
| OpenAI / ChatGPT | Trains by default (opt-out available) | No training (Team/Enterprise/API) |
| Anthropic / Claude | Consent-based; default disputed (pricing page labels consumer tiers “Opt-out”) | No training (Claude for Work/API) |
| Google Gemini | Trains by default (Keep Activity on) | β |
| Microsoft Copilot | Trains by default (opt-out for logged-in) | No training (M365 Copilot) |
| Meta AI | Trains on Meta AI interactions | β |
| xAI / Grok | Trains by default | No training (Business/Enterprise) |
| DeepSeek | Trains by default (stored in China) | β |
| Mistral / Vibe | Trains by default (Free/Pro/Education) | No training (Team/Enterprise) |
The pattern is unmistakable: consumer tiers train on you by default; the paid business tiers are where the privacy lives. If you’re an individual, the default is against you. If you’re a business, you’re paying to opt out by default.
The takeaway
- The default across consumer AI is training-on-by-default with a buried opt-out. Treat “we respect your privacy” on a majors’ page as a claim to check, not a fact.
- Anthropic’s consumer stance is now consent-based and its default is disputed β the clean opt-in exception it once was is gone, so treat Claude like the rest and check the toggle.
- The business/enterprise tier is your privacy shortcut at the majors: those plans don’t train on you, and you’re paying for the privilege.
- If you want the model to literally not hold your conversations, the providers that start there are few β and that’s exactly what Part 2 covers.
Information as of August 2026; product and pricing details change β verify against the linked pages before relying on them. Nothing here is legal or security advice.
Sources
- OpenAI β How your data is used to improve model performance: openai.com/policies
- OpenAI β ChatGPT pricing: chatgpt.com/pricing
- Anthropic β Is my data used for model training: privacy.claude.com
- Anthropic β Updates to our privacy policy (consumer training + 5-yr retention): privacy.claude.com
- Anthropic β Incognito chats: support.anthropic.com
- Anthropic β Claude pricing: claude.com/pricing
- Google β Gemini Apps Privacy Hub: support.google.com
- Google β AI plans: one.google.com
- Microsoft β Copilot Privacy FAQ: support.microsoft.com
- Microsoft β 365 Copilot pricing: microsoft.com
- Meta β WhatsApp Meta AI FAQ: faq.whatsapp.com
- xAI β Privacy policy (effective Apr 2026): x.ai/legal ; pricing: x.ai/pricing
- DeepSeek β Privacy policy: cdn.deepseek.com
- Mistral β Do you use my data to train: help.mistral.ai ; pricing: mistral.ai/pricing