> ## Content Index
> Fetch the complete content index at: https://adjacent.media/llms.txt
> Use this file to discover other available public pages before exploring further.

# AI Labs Hoard Pre-2022 Books to Dodge Their Own Slop
- URL: https://adjacent.media/signals/ai-labs-hoard-pre-2022-books-to-dodge-their-own-slop/
- Published: 2026-08-22T16:09:02.000Z
- Updated: 2026-08-22T16:09:02.000Z
- Description: As AI training data decays into self-referential garbage, frontier labs are treating pre-internet-collapse content as scarce resource—paying premiums for books published before the feedback loop poisoned web text.
- Author: Jonathan Greene
- Tags: #signal, theme-ai, training data quality, ai model development, content watermarking

Source: [Search Engine Journal](https://www.searchenginejournal.com/the-house-doesnt-publish-its-tells/586359/?ref=adjacent.media)

As AI training data decays into self-referential garbage, frontier labs are treating pre-internet-collapse content as scarce resource—paying premiums for books published before the feedback loop poisoned web text. The simultaneous investment in watermarking infrastructure reveals the actual concern: preventing competitors from identifying and harvesting proprietary training sets, turning content provenance into a competitive advantage.