Of course they train on literally everything they get their hands on, like everyone else. If you need privacy, that's what local models are for.
Whether you trust them is different, but there ARE knobs on other hosted AI companies.
Would you trust this?
Anthropic paid 1.5 billion for illegally training on stuff and that's even without a judge forcing them to do this. This seems to be their modus operandi
When these companies prove dishonest, I'll adopt skepticism.
You can say they stole from everyone to train their models in the first place and that's valid, but this isn't that. You are saying they are actively ex-filtrating data from any company using their services and lying about it.
Google/Apple/Microsoft or all of the dozen trillion dollar companies in the US would absolutely crush them in litigation. Neither OpenAI or Anthropic would be able to survive it. It's just not worth the risk.
But unless you are one of those you mentioned and a few others you probably aren't notable enough to care about. Everyone who uses their services directly, paying or not, is surely ignored in that sense. I wouldn't be surprised if there's a team of their own lawyers ready to interpret their EULA in fascinating ways.
And out of those three I'd only probably assume Apple is the only one who doesn't use the data given that they've built up privacy as a selling point, MS and Google probably train their own models on it themselves.