Executive Overview
In a stark illustration of the lengths to which technology companies will go to secure pristine, human-authored data, artificial intelligence giants are quietly buying up millions of physical books—spanning everything from rare, out-of-print volumes to forgotten paperbacks—only to feed them into industrial cutting machines. Known colloquially as "destructive scanning," this practice has triggered a profound ethical, cultural, and legal firestorm among authors, historians, independent booksellers, and intellectual property advocates worldwide.
Major artificial intelligence developers, including heavyweights like Anthropic and retail titan Amazon, have been quietly orchestrating massive, global data-acquisition operations. Their primary objective? To acquire long-form, high-quality human writing published before the digital saturation of the internet and the modern era of synthetic "AI slop." By purchasing physical books outright and digitizing them locally, these tech firms exploit legal doctrines surrounding physical ownership and fair use, ostensibly bypassing the treacherous legal minefield associated with scraping copyrighted digital texts or utilizing pirated "shadow libraries."
Read Also
However, this relentless pursuit of machine intelligence has a physical cost: the permanent erasure of cultural artifacts. From dusty antiquarian shops in rural Ireland to sprawling second-hand marketplaces online, booksellers are witnessing a bizarre surge in automated, eclectic bulk orders. The trail of destruction recently culminated in an explosive investigation involving an Apple AirTag, which tracked an anonymous bulk order directly to an Amazon warehouse in Nevada where human workers systematically slice the spines off books to feed high-speed scanners. As multi-billion-dollar legal battles and hefty settlements underscore the high stakes of data acquisition, the literary world is left grappling with an unsettling new reality: in the age of generative AI, physical books are no longer treated as vessels of knowledge or culture, but as disposable raw material.
Detailed Chronology: From Whispers in the Trade to the Nevada Warehouse
The systemic acquisition of physical print for machine learning did not happen overnight; it represents an escalation in the ongoing data crunch faced by large language model (LLM) developers. As the internet becomes increasingly saturated with AI-generated content—creating a feedback loop that degrades the quality of future models—engineers have increasingly looked backward to pre-2022 human-authored texts to preserve grammatical nuance, stylistic complexity, and factual coherence.
Spring 2025: The Global Bookseller Anomalies Begin
Independent booksellers across the globe began noticing unusual ordering patterns as early as May. Bookshops in the United Kingdom, Ireland, continental Europe, Canada, the United States, and Australia reported a sudden influx of sporadic, highly eclectic bulk orders.
Jim Shaughnessy, owner of MW Books in Claregalway, Ireland, highlighted the bizarre nature of these transactions to The Guardian. "They are sporadic, presumably machine-driven and have little in terms of thematic currency," Shaughnessy observed, noting that orders ranged wildly from esoteric treatises on 18th-century agricultural implements in Africa to biographies of 1950s racing car drivers. While booksellers are generally hesitant to police customer reading habits, the sheer randomness of the selections raised immediate red flags within the antiquarian and used-book community.
Late 2025: The Third-Party Shell Game
As vendors compared notes across international borders, Australian and North American booksellers discovered that many of these orders were being routed through sophisticated third-party intermediaries. Companies such as Zoom Books—a Canadian enterprise marketing itself as a "book recycler"—frequently served as fronts or logistical nodes.
Orders typically materialized through major secondhand marketplaces like Abebooks and Biblio. Acquisitions ran the gamut from low-cost paperbacks to rare, expensive collector’s items. A single automated purchasing batch might pair a 1970s technical manual on soil mechanics with Born to Thunder: Champions of New Zealand Cycling, an out-of-print anthology of early Australian poetry, and a localized historical account of the Melbourne suburb of Hawthorn.
Early 2026: The AirTag Investigation and the Nevada Discovery
The true mechanics behind these bulk acquisitions were brought to light by an investigative report published by 404 Media. A vigilant bookseller, suspicious of an anonymous order for 1,000 titles placed via the Biblio marketplace, decided to take matters into their own hands. Concealing an Apple AirTag discreetly inside the pages of one of the shipped books, the bookseller watched the tracking data as the package journeyed across the North American logistics network.
The tracking signal terminated at a massive Amazon fulfillment center in Nevada (specifically designated as VGT3). Through subsequent interviews with employees working at the facility, the investigative journalists uncovered a chilling industrial workflow. Rather than cataloging or reselling the literature, the primary responsibility of a dedicated team inside the warehouse was to systematically decapitate incoming books—detaching their bindings using heavy-duty industrial guillotine cutters—to lay the loose pages flat for rapid, automated high-speed optical character recognition (OCR) scanning.
"I work at VGT3 here in Vegas, and our main task is scanning books," an anonymous Amazon employee revealed on an internal worker forum cited by 404 Media. "Some team members are assigned to cut the books, while others handle receiving and scanning the barcodes. We didn’t have targets before, but now we do, although it’s not really stressful."

Supporting Context & Metrics: The Scale of "Destructive Scanning"
The economic and logistical dimensions of destructive scanning highlight a multi-million-dollar industrial complex dedicated entirely to text extraction. To understand why tech giants are willing to purchase and destroy millions of physical books, one must examine the legal and computational pressures driving the artificial intelligence industry.
The Legal Loophole of Physical Ownership
Copyright law has long wrestled with the digital reproduction of text. When AI companies scrape the internet or ingest unauthorized digital archives (such as pirate "shadow libraries"), they expose themselves to widespread copyright infringement lawsuits from creators and publishers.
However, purchasing a physical book grants the buyer certain rights under the First Sale Doctrine and fair use principles regarding personal transformation, digitization for internal analysis, or archival preservation. By acquiring physical copies legally through open markets and reducing them to digitized text in secure, private facilities, tech companies construct a legal fortress around their training datasets. They can argue in court that every byte of training data originated from a legally purchased, lawfully owned copy of the book—even if the physical artifact was discarded or destroyed in the process.
The Financial Magnitude: Project Panama
The scale of these operations involves staggering amounts of capital. Investigative reporting by The Washington Post previously uncovered that AI pioneer Anthropic poured tens of millions of dollars into an operation dubbed "Project Panama." This covert initiative involved acquiring, scanning, and subsequently destroying millions of books to train iterations of Anthropic’s flagship Claude language model.
The operation eventually surfaced not through voluntary corporate transparency, but via a massive copyright class-action lawsuit filed by aggrieved authors. The legal pressure culminated in a landmark $1.5 billion settlement extracted from Anthropic—a figure that underscores just how much value these tech enterprises place on uncorrupted, long-form human literature. Recently, courts ordered the unsealing of key documents related to the case, offering the public a rare glimpse into the corporate machinery behind large-scale text harvesting.
| Metric / Dimension | Detail |
|---|---|
| Primary Driver | Acquisition of clean, pre-2022 human prose free from synthetic "AI slop." |
| Key Methods | Industrial guillotine cutting ("spine-slicing") followed by high-speed OCR scanning. |
| Notable Entities | Anthropic, Amazon (VGT3 facility in Nevada), Zoom Books, and various third-party brokers. |
| Financial Impact | Multi-million-dollar acquisition budgets; settlements reaching up to $1.5 billion (e.g., Anthropic litigation). |
| Primary Channels | Secondhand marketplaces including Abebooks, Biblio, and direct international bookstore orders. |
Official Statements and Industry Backlash
The revelation that centuries of human thought, regional history, and literary artistry are being shredded for the sake of parameter optimization has drawn fierce condemnation from cultural heritage sectors.
The Bookseller Perspective
Independent booksellers, who have long functioned as the stewards of print culture, find themselves in an ethical bind. While businesses rely on revenue from book sales to stay afloat—particularly in an era dominated by digital e-readers and online retail monoliths—the realization that their inventory is being systematically obliterated leaves a bitter taste.
"Not every rare book is worth a fortune," one anonymous bookseller noted to 404 Media, emphasizing that many volumes carry irreplaceable historical, regional, or sentimental value. "AI companies don’t care about that. They just want the content as a bunch of words strung together."
Jim Shaughnessy echoed this sentiment while expressing the philosophical resignation felt by many merchants. While sellers cannot legally dictate what a paying customer chooses to do with private property once a transaction is complete, the commodification of literature into mere computational fodder represents a sad milestone in the commodification of culture.
The Author and Creator Outrage
Literary guilds, authors’ unions, and intellectual property watchdogs have slammed the practice as an egregious violation of the spirit, if not always the exact letter, of copyright law. Creators point out that purchasing a book grants reading and ownership rights, but utilizing industrial machinery to mass-destroy cultural works solely to train competing machine intelligence models exploits legal loopholes designed for libraries and individual preservationists.
Writers argue that destroying books to feed AI models creates a perverse irony: technologies built on the eradication of physical literature are effectively consuming the very foundation of human creativity that made their existence possible in the first place.
Future Outlook: What Lies Ahead for Print and AI?
As courtrooms continue to unpack the unsealed documents from landmark copyright cases and investigative journalists shed light on covert operations like Amazon’s Nevada scanning facilities, the intersection of physical publishing and artificial intelligence stands at a critical crossroads.
- Stricter Vendor and Marketplace Oversight: In response to the influx of opaque, machine-driven bulk orders, many independent bookstores and secondhand marketplaces are beginning to implement stricter monitoring systems. Some sellers are actively screening buyers suspected of scraping inventory for corporate data harvesting, seeking to protect rare regional histories and out-of-print editions from permanent destruction.
- Evolving Legal Precedents: The $1.5 billion settlement involving Anthropic has set a high financial watermark for unauthorized text acquisition. Future litigation will likely test whether "destructive scanning" under the guise of physical ownership truly holds up as a protected fair use defense, or if courts will rule that mass-scale industrial destruction of copyrighted works constitutes willful infringement.
- The Search for Synthetic Solutions: As physical book acquisition becomes legally hazardous and publicly controversial, AI developers are under intense pressure to develop synthetic data generation techniques that can match the quality of human prose without requiring the physical liquidation of global literature. Until then, however, the shadowy pipeline connecting dusty antiquarian bookshelves to high-speed industrial shredders remains operational—a sobering testament to the resource-hungry nature of the artificial intelligence boom.

Comments
Facebook App ID not configured. Please add it in the Customizer.