OpenAI Hid Billions of Records for the NYT Copyright Suit
OpenAI is accused by the NYT in a lawsuit of hiding 78M ChatGPT chat logs and an internal system called “Project Giraffe” for measuring the amount of news articles in its output.
The NYT and Daily News claim in their lawsuit against OpenAI that it hid 78M chat logs from its ChatGPT chatbot and an internal system called “Project Giraffe” that measures the amount of news articles in its output. The company only gave 20M logs out of the 120M requested and heavily redacted them, making them useless. They also broke a court order by deleting “billions” of model outputs, according to testimony from OpenAI engineer Vinnie Monaco.
The NYT is moving for sanctions and seeking to have the 20M logs produced inadmissible. OpenAI claims they are protected under the privacy of users and fair use. If OpenAI is found to be in breach of a court order, this will have far reaching implications in the AI industry. Any models built on top of third party data can be sued. There’s also a potential to have a huge licensing market for training data.
Source: techcrunch.com
Free course
Stop reading about AI — start building with it
The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.
Start free →
Author
Evgenii Arsentev
PhD · Chief Executive Officer, digital health
Articles · Latest articles