Court filing reveals Microsoft, OpenAI knew AI training on journalists' work posed existential threat
Unsealed documents in copyright lawsuit show internal memos calling the practice theft; companies proceeded anyway for profit.
What to know
- Internal memos from Microsoft and OpenAI acknowledged that training AI on journalists' work constituted theft and posed an existential threat to publishing and journalism employment.
- OpenAI obtained over 1.8 million New York Times articles from a corpus licensed only for non-commercial use, violating the licensing agreement.
- Despite recognizing these harms, both companies proceeded with commercialization, driven by the enormous financial returns the products would generate.
- The New York Times and 11 other news organizations are suing for copyright infringement; the companies do not dispute using the stolen content.
“has started a 'doom loop,' one that will ironically 'hurt the performance of our models and the entire web at the same time'”
Microsoft memo, Internal analysis · Truthout ↗
OpenAI AI company defendant in copyright lawsuitMicrosoft AI company defendant in copyright lawsuitNick Turley Led ChatGPT development team
Sam Altman OpenAI co-founder and CEOThe New York Times and 11 other news organizations Plaintiffs in copyright infringement lawsuit
How it unfolded 1 development · click the chart to see its coverage articlesposts
-
1
Documents show OpenAI and Microsoft pursued commercialization despite internal warnings
The court filing demonstrates that despite recognizing the harms posed by their AI systems, Microsoft and OpenAI continued developing the products driven by financial motivations. The filing notes that although OpenAI was founded with a goal of "benefit[ing] humanity as a whole," by 2017 CEO Sam Altman was planning to form a for-profit entity to carry forward AI research. Microsoft recognized that "millions of people around the world will soon consider large models 'hoovering up' all their work to be an astonishing theft of unprecedented proportions."
“millions of people around the world will soon consider large models 'hoovering up' all their work to be an astonishing theft of unprecedented proportions…”
— Microsoft (internal recognition) -
background
Internal memos show Microsoft and OpenAI recognized harms of using journalists' work — Documents detail internal memos where Nick Turley, who led ChatGPT development, called AI an "existential threat" to publishers. A senior Microsoft employee described the AI models as "an astonishing theft of unprecedented proportions" and possibly the "largest theft of labor in human history." Another memo warned that AI could "significantly disrupt[] the employment of the very people who generated the data on which the foundation model was trained" and created a "doom loop" that would "hurt the performance of our models and the entire web at the same time."
-
background
Court filing reveals Microsoft and OpenAI internal warnings about AI training on news content — Unsealed court documents in a copyright infringement lawsuit by The New York Times and 11 other news organizations show that Microsoft and OpenAI employees expressed alarm that their jointly developed AI systems could pose an "existential threat" to the publishing industry. The documents reveal that OpenAI used over 1.8 million New York Times articles published between 1987 and 2007 from The New York Times Annotated Corpus, which had a user license restricting use to "non-commercial" purposes.