AI companies should pay billions to wayback machine for access

I think that'd raise serious copyright concerns, if the Wayback machine started selling other people's intellectual property.

Shouldn’t it follow that it’s illegal for the AI labs to profit off of all of that stolen copyrighted data too?

It should, but it doesn't.

[deleted]

[flagged]

They wouldn't be paying for the content, just the bandwidth. Like buying a linux OS on a CD ROM was about the cost of media not profiting off of the software.

Isn't that already a big part of reddit's business model?

It's time for copyright to end anyhow; that's what's gumming up the whole project in the first place.

I'd have a lot less of a problem with AI if everything that went into their training was public domain and made easily available to anyone for any use. It'd feel less like AI companies were just stealing the work of others and charging for it.

Seems like a reasonable norm:

* If you train AI on it, you have to afford public access to it.

* Nobody can exact violence against anybody else in response to that person providing public access to any data anymore (ie, all bytestrings are public domain).

That's the world I'd like to try in the coming years.