
Microsoft and OpenAI’s Secret War Over The Times
Newly unsealed court filings reveal that Microsoft privately referred to OpenAI’s data practices as “theft” during a period when both companies were scraping paywalled content from The New York Times. The documents show that both companies were building datasets from The Times’ content, despite internal warnings that the practice would harm publishers. Microsoft and OpenAI have both been accused of using copyrighted material without proper licensing, raising questions about their ethical and legal responsibilities. The filings also show that both companies were aware of the potential damage their actions could cause to news organizations. The Times has been vocal about its opposition to the unauthorized use of its content by AI companies.
Microsoft’s Internal Warnings
Internal communications from Microsoft show that the company was aware of the potential negative impact of scraping The Times’ content. The company warned that the practice could “gut publishers” and damage the media industry. These warnings came as both Microsoft and OpenAI were actively building datasets from The Times’ paywalled content. The internal documents suggest that Microsoft was not entirely surprised by the legal and ethical backlash that followed. The company’s public stance on OpenAI’s data practices appeared to contradict its private concerns.
OpenAI’s Data Practices Under Scrutiny
OpenAI has faced increasing scrutiny over its data practices, particularly regarding the use of copyrighted material from The New York Times. The company has been accused of using The Times’ content without proper licensing, which has led to legal challenges. The newly unsealed court filings show that OpenAI was aware of the potential consequences of its actions. Despite this, the company continued to build datasets from The Times’ content, raising questions about its commitment to ethical AI development. The controversy has fueled debates about the role of AI companies in the media landscape.
The Times’ Legal Fight
The New York Times has been actively fighting against AI companies that use its content without permission. The company has filed lawsuits against OpenAI and other firms, accusing them of copyright infringement. The newly unsealed court filings provide insight into the internal discussions between Microsoft and OpenAI regarding the risks of scraping The Times’ content. The Times has argued that its content should not be used without proper licensing or compensation. The legal battle highlights the growing tension between AI companies and traditional media outlets.
A Growing Trend in AI Data Acquisition
The use of paywalled content by AI companies is not unique to Microsoft and OpenAI. Many tech firms have been accused of scraping content from news websites to train their models. The practice has raised concerns about the sustainability of the media industry and the ethics of AI development. The Times’ legal actions have drawn attention to the broader issue of how AI companies source their data. As AI continues to evolve, the debate over data rights and ownership is likely to intensify.









