Modelwire
Subscribe

OpenAI and Microsoft's internal warnings about web damage surface in court

Illustration accompanying: OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web

Newly unsealed litigation documents from the New York Times case reveal that OpenAI and Microsoft internally acknowledged creating a self-reinforcing cycle of web degradation through their training data practices. The companies' own records characterize their model training as extractive at scale, raising questions about the sustainability of current LLM development models. This disclosure shifts the liability calculus around AI training practices from theoretical concern to documented corporate knowledge, potentially influencing how courts and regulators assess fair use claims and labor impact in the AI sector.

Modelwire context

Analyst take

The critical detail the summary underplays is the word 'internally acknowledged.' This isn't a whistleblower or an outside critic's characterization. It's the companies' own language in their own records, which is a qualitatively different evidentiary problem than disputed expert testimony about harm.

This is largely disconnected from recent activity in our archive, as we have no prior coverage to anchor it to. But it belongs to a longer-running story about the legal infrastructure being built around AI training practices, specifically whether fair use arguments can survive proof of actual corporate awareness of harm. The NYT case has been the most watched test of that question, and these disclosures suggest the defendants' best arguments may be harder to sustain than their public posture implied.

Watch whether the presiding judge allows this internal characterization into the fair use analysis directly, rather than treating it as peripheral to the technical copying claims. If it enters the fair use record, it will almost certainly be cited in every parallel training-data suit filed in the next 18 months.

This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.

MentionsOpenAI · Microsoft · New York Times · GPT

MW

Modelwire Editorial

This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.

Modelwire summarizes, we don’t republish. The Verge - AI originally reported this story as OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web”. The full content lives on theverge.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.

OpenAI and Microsoft's internal warnings about web damage surface in court · Modelwire