|
|
|
|
|
by bloppe
5 days ago
|
|
They're attitude is more "this is fair use" which, according to all precedent, is probably true in most cases (unless the models actually start regurgitating huge parts of the Copyrighted material without a license). Of course, distillation is also fair use under Copyright law. Copyright was never meant to be a moral framework. It was always a practical framework designed to incentivize publishing that would ultimately pass into the public domain. Everybody seems to want to attribute some sort of moral weight to it though; the idea that people are naturally entitled to certain rights over things they've published. That idea would be totally alien to the people who designed the Copyright system in the first place. |
|
I am completely willing to accept that "this is fair use" for any company that publishes the LLM weights, i.e. the result of processing all the copyrighted work, because they have performed a public service with this.
But when the so-called "fair use" was a method to transform public data into private data that they guard and claim that any access to it would now be IP theft and which they use to obtain huge profits, that does not look like fair use to me.