{"id":844177,"date":"2026-08-03T10:26:23","date_gmt":"2026-08-03T10:26:23","guid":{"rendered":"https:\/\/www.newsbeep.com\/ca\/844177\/"},"modified":"2026-08-03T10:26:23","modified_gmt":"2026-08-03T10:26:23","slug":"suppliers-scramble-as-rare-books-destroyed-to-train-ai","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ca\/844177\/","title":{"rendered":"Suppliers Scramble As Rare Books Destroyed To Train AI"},"content":{"rendered":"<p>This week, heavy readers and casual posters alike were aghast at emerging reports of rare books being mulched in bulk as a sacrifice for LLMs. Word traveled that book suppliers were receiving abnormally large orders, believing that AI firms sought cleaner source material while skirting copyright laws. At the center of the controversy was ISBNdb, a book database who not only encouraged the trend but seemed to facilitate it. The optics bit back, <a href=\"https:\/\/www.404media.co\/ai-company-training-scanning-books-database-isbndb\/\" rel=\"nofollow noopener\" target=\"_blank\">as the site now tries to distance itself<\/a> from the controversy, scrubbing its own posts on the subject.<\/p>\n<p>\u201cWe\u2019ve seen the recent coverage about a marketing landing page on our site, and we understand the concern it raised,\u201d <a href=\"https:\/\/isbndb.com\/news\" rel=\"nofollow noopener\" target=\"_blank\">writes ISBNdb in an update<\/a>. \u201cWe don\u2019t train AI models, and we never have. The page was a test of market interest; no such service was ever brought to life. We\u2019ve taken the page down.\u201d<\/p>\n<p>A few days earlier, <a href=\"https:\/\/www.404media.co\/ai-companies-are-buying-tons-of-old-books-because-theyre-free-of-ai-slop\/\" rel=\"nofollow noopener\" target=\"_blank\">404 Media reported<\/a> about the concerning trend of suppliers suddenly hit with bulk book orders. With schools and libraries lean on resources, the decreased business has left them vulnerable. It was suspected that AI firms were making these new orders, but putting themselves in front of the crosshairs was ISBNdb, a database who seemingly introduced a service to patch AI firms through to depositories.<\/p>\n<p>\u201cThe world\u2019s best AI training data is sitting on a shelf,\u201d <a href=\"https:\/\/web.archive.org\/web\/20260727225205\/https:\/\/isbndb.com\/print-books-for-ai-training\" rel=\"nofollow noopener\" target=\"_blank\">read the now deleted landing page<\/a>. \u201cBooks represent curated, peer-reviewed, domain-specific human knowledge, structured in a way no web crawl can replicate. Dense, edited, authoritative.\u201d<\/p>\n<p>The outrage hit a fever pitch this week but <a href=\"https:\/\/www.washingtonpost.com\/technology\/2026\/01\/27\/anthropic-ai-scan-destroy-books\/\" rel=\"nofollow noopener\" target=\"_blank\">news of the practice broke last January<\/a>, when court documents exposed \u201cProject Panama,\u201d a program within Anthropic to scan then destroy as many books as they can get a hold of. Anthropic settled with authors for $1.5 billion, but the unsealed filings showed that the practice is considered legal, just bad publicity, and the company was willing to break the bank to keep it under wraps. Now the juice is out of the tube, and suspicious rare book orders are under intense scrutiny.<\/p>\n<p>\u201cIt benefits me financially as well as by clearing out old inventory that is otherwise unlikely to sell,\u201d one anonymous seller told 404\u2018s Samantha Cole. \u201cOn the other hand, I don\u2019t like the end-use, and I don\u2019t like that uncommon books are being pulped.\u201d<\/p>\n<p>Confronting the hell they\u2019ve rendered for themselves, AI firms are struggling to find \u201cclean\u201d source material to train LLMs on. Whether it\u2019s low quality web content, or material already generated by AI that threatens negative feedback loops, these companies are trying to siphon higher quality stuff without raising too many alarms. Earlier this summer <a href=\"https:\/\/kotaku.com\/a24s-bad-excuse-for-its-controversial-google-ai-partnership-isnt-impressing-fans-wed-rather-have-a-seat-at-the-table-2000711181\" rel=\"nofollow noopener\" target=\"_blank\">A24 announced a partnership with Google<\/a> to assist training DeepMind, hoping to help their media generators escape the bog of <a href=\"https:\/\/kotaku.com\/studio-ghibli-ai-lord-rings-lotr-openai-sora-1851773155\" rel=\"nofollow noopener\" target=\"_blank\">muddy brown, cursed Ghibli slop<\/a>.<\/p>\n<p>As for the destruction part, it\u2019s unlikely being done in attempts to cover their tracks or some Brianiac-style rare knowledge obsession. The Project Panama documents didn\u2019t specify why they trashed the books, but it\u2019s most likely just trying to save a buck. Archiving and scanning books <a href=\"https:\/\/www.youtube.com\/watch?v=QThaHpkFVzw\" rel=\"nofollow noopener\" target=\"_blank\">doesn\u2019t have to destroy the source material<\/a>, but being done cheaply and quickly, it likely will. Ripping apart the spines for clearer scans and disposing of the crumpled heap.<\/p>\n","protected":false},"excerpt":{"rendered":"This week, heavy readers and casual posters alike were aghast at emerging reports of rare books being mulched&hellip;\n","protected":false},"author":2,"featured_media":844178,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[62,276,277,353,49,48,61],"class_list":["post-844177","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-books","tag-ca","tag-canada","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/844177","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/comments?post=844177"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/844177\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media\/844178"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media?parent=844177"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/categories?post=844177"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/tags?post=844177"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}