Comment on AI labs buy, scan, shred millions of rare books
Cethin@lemmy.zip 3 days agoIt’d be such an easy PR win to do this, but of course they aren’t. I guess for things they don’t have the rights to they probably couldn’t easily, but they could for older things. They could also create a database of it and release it when they become public domain too, which wouldn’t be that hard.
Kissaki@beehaw.org 3 days ago
If they can publish a dataset of all of GitHub* then they can publish a dataset of all books.
* main branch head state, + some filtering and manual opt-out-apply logic