@YourNetworkIsHaunted@awful.systems
Post #4229828
2026-07-30 06:22 UTC
What kills me is that there are so many obvious ways to be less wasteful about this. Nondestructive book scanners exist. They’re expensive but it’s not like AI companies are averse to throwing money down a fucking hole. Even if that’s not an option, it’s possible to rebind the pages and return the books to the market. And regardless of what happens to the physical books, the scans they create could be archived in a format that could be available as a digital library rather than just being fed into the statistical meat grinder to keep the trough topped off with slop. Even if they’ve got to fight with copyright holders it would cost them basically nothing to leave the option open and given the PR battle they’ve been losing it feels like doing so would be incredibly obvious. Hell, even Google Books was able to navigate this in a way that was less cartoonishly evil than this because they could point to their digitization effort as a public good in ways that Anthropic here just fucking can’t.
Replies (1)
-
@corbin@awful.systems 2026-07-30 18:37
The reason that they are destroying the books is to avoid the accusation that the scanning constituted copying, an irony that we’ve discussed previously, on Awful while trying to understand which court cases are relevant. Anthropic actually hired the same guy, Tom Turvey, who designed Google Books’ ingestion process, so I’m thinking of this as a sequel to Google Books; hopefully the courts won’t take a decade this time.