AI Companies Are Buying Rare Books, Scanning Them for Training Data, Then Shredding Them
Multiple AI companies are buying rare books, scanning them for training data, then shredding the originals — including irreplaceable texts.
As TechCrunch first reported, the Model Context Protocol — the open standard that lets AI agents connect to external data sources and tools — is receiving a significant usability overhaul built around a new stateless approach to session management. For developers who have tried and abandoned MCP implementations, the change directly addresses the most common complaint: that MCP demanded too much state-tracking overhead compared to how standard web APIs are typically designed.
MCP matters because it defines how AI agents reach beyond their training data — pulling live information from databases, files, calendars, and third-party services. More capable and reliable agents depend on well-implemented MCP connections, but the protocol's complexity has kept adoption narrower than its backers hoped.
The new stateless session model brings MCP architecture closer to REST and other familiar API patterns, reducing the specialist knowledge required to build a working implementation. That lowers the bar meaningfully for developers who are not protocol experts but want to build agents that interact with real-world data.
Whether adoption follows will depend on how quickly toolchains, SDKs, and documentation catch up to the updated spec.
All comments are reviewed before appearing. Keep it respectful.
Multiple AI companies are buying rare books, scanning them for training data, then shredding the originals — including irreplaceable texts.
Adam Mosseri says Instagram no longer requires engineers to code 40–60% of the time and has scrapped the traditional technical interview loop.
Sam Altman says AI won't shorten the workweek because humans enjoy staying busy and will just create more work to fill the time.