lunes, 7 de septiembre de 2026

Newspapers Sue OpenAI Over AI Training Practices

The escalating tension between traditional media organizations and artificial intelligence developers has reached a critical juncture. Two major newspaper publishers have initiated legal action against OpenAI and Microsoft, alleging that their proprietary journalism has been systematically scraped and used to train large language models without authorization or compensation. This lawsuit marks a significant escalation in the ongoing debate regarding intellectual property rights in the era of generative AI.

A new battle between journalism and artificial intelligence: Two newspapers sue OpenAI and Microsoft - صوت الإمارات
Imagen generada con IA

The plaintiffs argue that these tech giants are effectively cannibalizing their business models. By ingesting copyrighted news content to power platforms like ChatGPT, AI developers are creating tools that can summarize or reproduce journalistic work, thereby reducing the necessity for users to visit the original news websites. The core of the complaint focuses on the unauthorized reproduction of protected content for commercial gain, positioning the AI industry as a direct threat to the financial viability of investigative journalism and news gathering operations.

For a technically literate audience, this case highlights the growing friction between the open-web data acquisition strategies of AI companies and the legal protections afforded to content creators. The defense, likely to be led by Microsoft and OpenAI, will almost certainly invoke the fair use doctrine, asserting that the transformative nature of AI training—converting raw text into probabilistic models—constitutes a legitimate and legal application of existing data. However, the legal threshold for what constitutes fair use in the context of massive dataset scraping remains largely untested in the courts.

This litigation is pivotal because it addresses the foundational question of whether AI companies owe royalties to the entities that provide the high-quality data necessary to make these models effective. If the courts rule in favor of the newspapers, it could force a radical restructuring of how AI training data is procured, potentially mandating a transition toward licensing models rather than open-ended web crawling. Conversely, a victory for the technology firms would solidify the status quo, effectively allowing the continued exploitation of the public web to fuel the next generation of generative AI products without direct oversight or financial participation from original authors.

Artículos relacionados de LaRebelión:


Fuente Original: صوت الإمارات

Artículo generado mediante AI.larebelion.

No hay comentarios:

Publicar un comentario

// Telegram BOT