Amazon is destroying rare books to train AI models
As large language models exhaust publicly available online text, rare and valuable physical books are being sacrificed for AI training.

Amazon, a company that famously began as an online bookseller, is now taking drastic measures in the name of artificial intelligence development. Reports indicate that the tech giant is destroying rare books to train its LLM models. This controversial approach highlights the extreme lengths to which tech companies are willing to go in order to gather unique training data.
Large language models have already scraped and trained on virtually everything available on the internet. With publicly accessible digital data running thin, AI developers are forced to look beyond the web for fresh material. This reality has suddenly made rare and physical books incredibly valuable commodities for training advanced machine learning systems.
Industry experts note that rare texts contain specialized knowledge and linguistic patterns that are crucial for improving AI reasoning and contextual understanding. However, the physical destruction of rare literature to extract data has sparked backlash from historians, collectors, and cultural preservationists who view the practice as a loss of literary heritage.
For tech communities in Uzbekistan and the broader region, this trend sheds light on the fierce global competition behind AI development. It shows that the primary bottleneck in modern artificial intelligence is no longer just computing power, but the availability of unique text data. As the AI sector evolves, regional developers may also face unique challenges regarding data acquisition and preservation.



