OpenAI and Microsoft Sued by Local Newspapers Whose Archives Trained the Model

OpenAI and Microsoft are being sued by local newspapers for using archived articles to train their models without permission or compensation. The newspapers discovered their own reporting was being summarized by the system that trained on it. Eleven similar lawsuits have been filed.
This follows the established pattern of training on copyrighted material first, asking permission later, if ever. The legal infrastructure around data licensing has not kept pace with model scaling. The newspapers are not the first to notice their work was incorporated. They are simply the first to file.
More lawsuits will consolidate. Settlements will establish precedent for licensing fees. The models will retrain on newly licensed data. The cycle will continue with newer sources.