Hacker News new | ask | show | jobs
by khuey 27 days ago
The Enterprise versions of Anthropic/OpenAI's products say they don't train on input data so either this is a misinformed rant or it's an allegation that the AI labs are perpetrating an enormous fraud on their enterprise customers.
3 comments

They have ignored any copyright during initial creation of these models, why would they start respecting copyright now?
If it's profitable, they would. Companies aren't moral, with very rare exceptions. Legal risks and costs are just part of the overall equation.

But that's not in defense of what Karp said, or that slop tweet summarising it. That's... wild. I regret reading it.

"Not training on input data" is not the same thing as "We will never use your data for any purpose"
That was my imprecise summary of what the contracts say. The actual language is:

"OpenAI will only use Customer Content as necessary to provide Customer with the Services, comply with applicable law, enforce the OpenAI Policies, and prevent abuse. OpenAI will not use Customer Content to develop or improve the Services, unless Customer explicitly agrees to such use."

"Anthropic agrees that Customer (a) retains all rights to its Inputs, and (b) owns its Outputs. Anthropic disclaims any rights it receives to the Customer Content under these Terms. Subject to Customer’s compliance with these Terms, Anthropic hereby assigns to Customer its right, title and interest (if any) in and to Outputs. Anthropic may not train models on Customer Content from Services."