Question of the Day
One question per day to look beyond the headlines.
How does training on paywalled news turn a copyright fight into a revenue-protection battle for journalism?
Take-away Paywalled training shifts harm from copying to disintermediation: models answer queries directly, breaking the traffic→subscription funnel publishers rely on.
Training on paywalled news without permission turns a copyright fight into a revenue-protection battle for journalism because it potentially reduces site traffic and subscription revenue for publishers. AI models trained on this content can generate answers to user queries, thereby substituting direct visits to publisher sites, which is a key revenue stream [1]. Publishers are concerned that AI-generated responses could lead to a decrease in direct access to their content, effectively undermining their business model which relies on web traffic and paid subscriptions [2]. Microsoft and OpenAI are implicated in favoring AI development profitability over the traditional revenue models of news outlets, turning this into a battle over protecting the financial sustainability of journalism [3]. The substantial potential damages discussed in the lawsuits further emphasize the financial impact, as publishers could claim up to $1.6 trillion based on copyright violations, demonstrating the high stakes involved for both sides [2].
- ‘Astonishing theft’: NYT filings reveal OpenAI knew about paywall workaround - Storyboard18 storyboard18.com (opens in new tab)
- ‘Astonishing theft’: News outlets are using OpenAI worker comments against them in lawsuit san.com (opens in new tab)
- OpenAI Execs Hoped to Make “Gazillions” After Feeding Model Journalists’ Work | Truthout truthout.org (opens in new tab)