OpenAI's Data Deletion Raises Concerns in NY Times Copyright Lawsuit

OpenAI is facing a lawsuit from The New York Times and Daily News for allegedly using their copyrighted content to train its AI models without permission. During the legal proceedings, OpenAI engineers accidentally deleted data on a virtual machine that was being used to search for the plaintiffs' content in OpenAI's training sets. Although OpenAI managed to recover most of the data, the loss of folder structures and file names rendered it unusable for the plaintiffs' purposes. OpenAI denies intentional deletion, attributing the issue to a system misconfiguration requested by the plaintiffs. The incident highlights ongoing tensions over AI training practices and copyright infringement.
- OpenAI accidentally deleted potential evidence in NY Times copyright lawsuit (updated) TechCrunch
- OpenAI Accidentally Deletes ChatGPT Training Data Amid Publisher Copyright Claims, Sparking Concerns Over Evidence Retention In Legal Cases Wccftech
- New York Times lawyers claim OpenAI accidentally deleted evidence in copyright case The Register
- Oops! OpenAI just deleted important legal data in a lawsuit from The New York Times Business Insider
- The Intercept’s Lawsuit Against OpenAI Advances on Claim It Removed Reporters’ Bylines The Intercept
Reading Insights
0
15
2 min
vs 4 min read
83%
602 → 104 words
Want the full story? Read the original article
Read on TechCrunch