Hacker News new | ask | show | jobs
by ACCount37 5 days ago
Niche text. It's not impossible that there was only ever under a thousand of them printed and released into circulation.

A digital copy would exist somewhere, of course. But for us, that only matters if we can buy or download it. And for AI companies, that only matters if they can get a digital copy DRM-free and licensed permissively enough.

1 comments

This argument doesn't make any sense. All manner of AI companies just ingest whatever random text they can find on the internet to train their data, including copyrighted publications. Why would DRM on a digital copy of a book matter?
> Why would DRM on a digital copy of a book matter?

Because DRM is just a way to make "breaking copyright" more practically cumbersome. What's easier, breaking digital DRM for each and every E-book you find, or just establishing a single pipeline for scanning physical books?

breaking DRM is so easy my generation was doing it as kids, there is no technical obstacle there.
It's a manual process.

More importantly, it's also explicitly illegal. Destructive format-shifting is not. Thank copyright laws.

Why would it be a manual pass? There's only so many DRM schemes, and you can even use your previous generation AI to help with the breaking.

> More importantly, it's also explicitly illegal.

Agreed!

Again, what's more technically feasible? Setting up a DRM-cracking LLM and manually verifying that it's actually succeeded for each and every DRM scheme you encounter, or just throwing the books in a regular old office scanner and being sure it works without even checking...
There might be special rules around DRM that go beyond normal copyright?
Ladies and gentlemen, the Digital Millennium Copyright Act

(which is terrible, but I would be delighted if they breached it and got thoroughly spanked)

I would be delighted if AI companies got together and thoroughly dismantled DMCA.

That atrocity of a law was a blight upon digital freedom since the day it came to exist. DRM should never have been given any legal protection - and I would push for numerous forms of DRM to be outlawed instead.

so if I base64 encode my blog, have some Javascript that 'validates' an authorized viewer and then decodes the base64 into HTML which is added to the DOM does that constitute DRM ?
Yes.