26
you are viewing a single comment's thread
view the rest of the comments
[-] FaceDeer@fedia.io 2 points 1 hour ago

"How dare AI companies pirate books to train their AIs! They should be paying for their training data!"

...

"Not like that!"

This bit of clickbait outrage has been making the rounds for weeks now. They're using the standard approach to bulk scanning books. The "rare editions" they're scanning are rare because nobody is interested in them. It's stuff like old guidebooks from a particular town fair in 1970 or a book on crochet in early colonial Kansas or whatever, they were sitting in warehouses and would have eventually been pulped if not bought in bulk. This is their one shot at preservation, frankly. And they're doing this because the law requires them to do it.

[-] mindbleach@sh.itjust.works 2 points 35 minutes ago

And in fairness, they should be able to share their scans - with each other, with the public, and with Archive.org.

But that means relaxing copyright, not getting even more tightassed about it.

this post was submitted on 02 Aug 2026
26 points (96.4% liked)

Hacker News

5245 readers
371 users here now

Posts from the RSS Feed of HackerNews.

The feed sometimes contains ads and posts that have been removed by the mod team at HN.

Source of the RSS Bot

founded 2 years ago
MODERATORS