> Anthropic, OpenAI, and even Grok have "Don't train on my data" knobs.
Would you trust this?
Anthropic paid 1.5 billion for illegally training on stuff and that's even without a judge forcing them to do this. This seems to be their modus operandi
Do you have proof it's not? There are no laws they have to follow in regards to it, and it's practically impossible for them to go against their core self interest. More data is literally a direct component to better models and more revenue. If someone proves it's fake worst they get is a month of bad PR and then people will move on, as they always do.
> If someone proves it's fake worst they get is a month of bad PR and then people will move on.
You can say they stole from everyone to train their models in the first place and that's valid, but this isn't that. You are saying they are actively ex-filtrating data from any company using their services and lying about it.
Google/Apple/Microsoft or all of the dozen trillion dollar companies in the US would absolutely crush them in litigation. Neither OpenAI or Anthropic would be able to survive it. It's just not worth the risk.
Of course they train on literally everything they get their hands on, like everyone else. If you need privacy, that's what local models are for.