# Training-Data Copyright Litigation Sets the Price of Content

Publishers whose content is directly substituted by assistant answers keep bringing infringement claims against model developers, gradually converting training data from a free input into a licensed and priced one and raising the cost floor for every lab that trains on the open web.

- Conviction: 36 / 100 (weakening)
- Horizon: Emerging (watchlist)
- Tracking since: 2026-08-26T00:00:00.000Z
- Last updated: 2026-08-28T06:28:38.436Z
- Canonical: https://polylog.news/ai/trends/ai-training-data-copyright-litigation
- Publisher: Polylog
- Affected regions: United States, Europe

## Recent score history

- 2026-08-27: 38
- 2026-08-28: 36

## Recent evidence

- [confirms] wikiHow Sues OpenAI, Citing 148,529 Crawler Visits and Verbatim Reproduction of Its Guides (2026-08-26): wikiHow sued OpenAI citing 148,529 crawler visits and verbatim reproduction of its guides, covering 1,211 registered copyrights across 11,211 articles, and argues the outputs substitute for the originals rather than transform them. The substitution framing plus logged crawl counts is the fact pattern most likely to convert scraped how-to content into a licensed, priced input, and how-to content is among the most directly displaced by assistant answers.
