{"slug": "the-devouring-of-all-knowledge", "title": "The Devouring of All Knowledge", "summary": "AI companies, including Anthropic, are buying, scanning, and shredding rare physical books at an unprecedented scale to train large language models, according to a Futurism report citing a settled lawsuit. The practice destroys unique copies of out-of-print works and risks sequestering all human knowledge behind corporate walled gardens, with critics warning of a dystopian scenario where history is gatekept by tech firms.", "body_md": "I was horrified but unsurprised to read reports 1 this week of the\nAI/LLM\n\ncompanies buying, butchering, scanning, and shredding rare books at an unprecedented scale.\n\n[2](#fn:2)With all online forums, PDFs, articles, and books consumed, the next\nbest source of human writing is the abundance of *physical books*\npublished before the advent of LLMs. The year 2021 3 and the\nemergence of LLM\n\ntext generation is coarsely termed\n\n[2](#fn:2)*“the ensh**tification of all knowledge”*. Writing published after this point is potentially LLM-written and therefore unsuitable to be used as training data. Whatever we wrote before this historic point is exclusively the work of human hands and minds.\n\nThe race is now on to consume and destroy (at least one copy of) every physical book on the planet first - and I wouldn’t rule out a scenario similar to the global memory shortage of 2025. At this time, purchases became weapons used to strangle the supply chains of competitors and gain an exclusive material advantage. Memory prices shot far out of reach for normal people as AI/LLM companies rushed to secure the remaining global supply of DRAM chips. Will we see the same thing happen with rare, old, and out-of-print books?\n\n“Many companies, including Anthropic, have turned to ingesting physical books instead, which they can buy countless used copies of on the cheap. According to the settled lawsuit, Anthropic used a hydraulic powered cutting machine to neatly remove the pages from the books it procured from book resellers and then scanned them using industrial-grade imaging equipment. In other words, it was literally ripping off authors’ books to train its AI.”\n\n–\n\nFrank L.Futurism[1]\n\nThe problem here is not just the destruction of rare and priceless\nbooks which, though currently unused, are presently accessible human\nknowledge - **hoarding this rare knowledge behind walled gardens** is\nthe real issue. Countless dystopian sci-fi visions of humanity’s\ndarkest paths describe the gathering and destruction of knowledge so\nthe record can be set straight by a central authority. Ray Bradbury’s\n*Fahrenheit 451* and George Orwell’s *Nineteen Eighty-Four* both come\nto mind as keen examples of governments and their proxies working\ntirelessly to destroy the past and rewrite history.\n\nUltimately, this leads to a nightmare situation where all knowledge is gathered\nby the tech companies feeding these insatiable large language models.\nEvery copy of every book will be frantically secured to keep them out\nof the hands of competitors. Knowledge will then become accessible\n*exclusively* through the censored and monitored outputs from these\nmodels. Information deemed ‘unsafe’ will be hidden forever underneath\nthe safety, compliance, and ethics layers of the largest models -\ngoverning and correcting the output of every individual response.\nWhatever is deemed harmful by regulators will be unavailable for the\ngreater good. History will be gatekept.\n\nAn argument could be made that these books are much better put to use\nin this manner. By digitizing and using the books as training data,\nthe knowledge is made more accessible to the everyman who uses\n*ChatGPT* as their primary knowledge source. I would be more excited\nabout this if the scans and content were not to be permanently\nsequestered in each LLM company’s knowledge silo. No guarantee is\nprovided that any of this material will be presented in an impartial\nmanner, or ever released to the public - this would remove the\nexpensive knowledge advantage the particular LLM company gained\nthrough the acquisition and scanning process.\n\n“One small book seller said that in April, he suddenly went from selling no more than 20 books a week to\n\nhundreds, and he’s almost certain that the customers are AI labs, noting the random selection of the books and how theyall have ISBNs. He added that his inventory is full with rare and out-of-print books, meaning that an AI company could be destroying some of the few remaining copies that can be found.”–\n\nFrank L.- Futurism[1]\n\n“Obviously, printed books represent a treasure trove of information for AI. However, many debate the ethics of removing books from circulation since it is uncertain whether AI companies filter the rare or even out-of-print books from the common titles during digitalization. The other major issue is that scanned books go directly into a private database to train AI, which the general public does not have access to. True, we will have smarter AI, but at the cost of the information not being available to future generations.”\n\n–\n\nZhiye L.- Tom’s Hardware[4]\n\n**What can you do about it?** *Learn to read, then buy and hold paper\nbooks.* Your children will thank you, and paper books are entirely\nyours to mark up, borrow, trade, and love - entirely out of the\nsubscription prison most companies are hurriedly constructing today.\nPaper books can’t be revoked, altered, or monetized past the point of\nacquisition, and therefore are unattractive to these businesses who\nseek recurring revenue above all other things.\n\n**Is the danger real?** *Not to a dystopian extent,* but just as we’ve\nall seen the deterioration in the quality of web results - **allowing\nthe American technological establishment to act as a middleman between\nus and all knowledge is a dangerous game**. Convenience comes at the\ncost of narrative injection. Summarized responses at the cost of\nviewing history through Silicon Valley’s frame. 5 I don’t think Meta\nand Anthropic are actually bent on buying every paper book in\nexistence, but this month is certainly the opening of an interesting\nnew frontier of knowledge collection, segregation, and protection.\n\nAnything we can collectively do to keep old, beautiful, and rare books out of the scanners and shredders of these LLM leviathans (in particular history books and encyclopedias written before 1945), is a great blessing to our future generations.\n\n*Here’s to a future with books!*\n\n*“AI Companies Are Buying Antique Books, Ingesting Their Contents to Train Models, and Then Destroying Them at Incredible Scale”*- Frank Landymore, Futurism, Pub. July 25th 2026,[futurism.com/artificial-intelligence/ai-companies-destroying-rare-books](https://futurism.com/artificial-intelligence/ai-companies-destroying-rare-books)[↩︎](#fnref:1)[↩︎](#fnref1:1)[↩︎](#fnref2:1)**LLM**:*Large language model.*A generic term for the algorithms that power ChatGPT, Claude, Microsoft Copilot, Grok, Llama, and many other*generative pre-trained transformer*neural network algorithms. These models are**not AI**, but they are branded with the term.[↩︎](#fnref:2)[↩︎](#fnref1:2)Some say 2022, but my first conversations with OpenAI’s Davinci model were around August 2021.\n\n[↩︎](#fnref:3)*“AI companies are reportedly shredding millions of books after using them to train AI models — tech giants outsource to middlemen to secretly buy up books for training material”*- Zhiye Liu, Tom’s Hardware, Pub. July 28, 2026,[tomshardware.com](https://www.tomshardware.com/tech-industry/artificial-intelligence/ai-companies-are-reportedly-shredding-millions-of-books-to-train-models-tech-giants-outsource-to-middlemen-to-secretly-buy-up-books-for-training-material)[↩︎](#fnref:4)**Narrative injection** and**frame** in these sentences both refer to the retrieval of textual information colored with the perspectives and worldview of the silicon valley technology companies. Research performed by Cripps (reported by the[Society for Computers and Law:](https://www.scl.org/llms-are-left-leaning-liberals-the-hidden-political-bias-of-large-language-models/)) revealed that LLMs like ChatGPT and Copilot return*“LLMs are Left-Leaning Liberals: The Hidden Political Bias of Large Language Models”***exclusively left-leaning answers** when the neutral option is excluded, and it is difficult to find a mainstream model that will even mention a right-of-center opinion.[↩︎](#fnref:5)", "url": "https://wpnews.pro/news/the-devouring-of-all-knowledge", "canonical_source": "https://ryanfleck.ca/2026/the-devouring-of-all-knowledge/", "published_at": "2026-07-28 15:43:32+00:00", "updated_at": "2026-07-28 15:52:13.640906+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-ethics", "ai-policy"], "entities": ["Anthropic", "Futurism", "ChatGPT", "Ray Bradbury", "George Orwell"], "alternates": {"html": "https://wpnews.pro/news/the-devouring-of-all-knowledge", "markdown": "https://wpnews.pro/news/the-devouring-of-all-knowledge.md", "text": "https://wpnews.pro/news/the-devouring-of-all-knowledge.txt", "jsonld": "https://wpnews.pro/news/the-devouring-of-all-knowledge.jsonld"}}