UA RU EN

Court Unveils OpenAI and Microsoft Communications in Lawsuit Over Copyright Issues in ChatGPT’s News Training

Суд розкриває переписку між OpenAI та Microsoft у справі, пов'язаній з правами авторства щодо новинної підготовки ChatGPT. Photo: НВ — Техно

Legal Battle Between OpenAI and Microsoft

A court has released confidential exchanges between OpenAI and Microsoft connected to a lawsuit filed by The New York Times in 2023 alongside five authors. The case centers on allegations that these tech giants trained ChatGPT using millions of news articles without proper copyright authorization, raising significant concerns about intellectual property rights and the future of journalism.

Documents reveal that OpenAI and its collaborators bypassed paywalls and stripped copyright notices from the training data. A 2023 internal Microsoft memo warned that 'millions of people worldwide' might perceive the use of their work by AI as theft. Brand Gecht, Microsoft’s Applied Sciences Director, described the situation as 'the greatest labor theft in human history, potentially triggering a death spiral.'

'This presents an existential threat to publishers.' - Nick Turley, OpenAI

Judges involved in the case noted that AI training laws remain underdeveloped. This lawsuit is seen as a pivotal moment for clarifying whether AI training qualifies as fair use. Meanwhile, Microsoft has sought to distance itself from internal remarks, with company spokesperson Alex Haurek stating that 'Copilot is not a replacement for publishers’ journalism.'

Implications for Journalism

Despite these reassurances, internal documents highlight the potential for large AI models to significantly undermine traditional journalism. An OpenAI engineer commented, 'No matter how clearly we show citations, users won’t click on them.' This underscores ongoing tensions surrounding the use of news content to train AI systems and the broader impact on the media industry.

The outcome of this case could have far-reaching consequences for the integration of AI in media, addressing crucial issues of copyright, ethical use, and fairness in content utilization. It may also prompt new regulations that define the boundaries of fair use for AI algorithms trained on existing informational sources.

As the legal landscape evolves, the European Union has recently classified ChatGPT as a large-scale online search engine, raising important questions about regulation and accountability in AI usage. This development could further complicate the ongoing debate over copyright and the role of AI in journalism. To explore how this classification might influence future legal battles surrounding AI technologies, read more about the EU's recognition here.