Editorial illustration for Britannica sues OpenAI, alleging ChatGPT copied its content verbatim
Britannica Sues OpenAI for ChatGPT Content Plagiarism
Britannica sues OpenAI, alleging ChatGPT copied its content verbatim
After 255 years, Encyclopedia Britannica has stopped publishing its arguments and started filing them in court. The publisher sued OpenAI this week with a shocking, simple allegation: ChatGPT didn't learn from its content. It memorized it.
The lawsuit claims the AI spits out whole passages from Britannica and its Merriam-Webster dictionary, verbatim. This isn't inspiration. It's a digital photocopier.
The lawsuit accuses OpenAI of outputting near-identical copies of Britannica and Merriam-Webster's content. According to Britannica, OpenAI repeatedly copied its content without permission, stating, "GPT-4 itself has 'memorized' much of Britannica's copyrighted content and will output near-verbatim copies of significant portions on demand. The memorized examples are unauthorized copies that [OpenAI] used to train their models, including GPT-4." The lawsuit goes on to include examples of responses from OpenAI's models side by side with Britannica's text, in which entire passages appear to match word for word. Britannica also claims that OpenAI has been "cannibalizing" its web traffic by generating responses that "substitute, or directly compete" with Britannica's content, rather than directing users to its website the way a traditional search engine would.
Every major AI lawsuit grapples with fair use. Britannica’s filing adds a brutal twist: the crime is in the output. Its side-by-side comparisons aim to prove GPT-4 regurgitates, plain and simple.
That shifts the legal debate from how data is ingested to whether paragraphs are stolen. The real killer for publishers, though, is the traffic claim. Why click through to a site when the chatbot serves the answer?
A business model built on facts evaporates. OpenAI scanned the library. Now it wants to be the library.
Britannica survived the internet. It may be the plaintiff that finally forces a judge to name the price.
Common Questions Answered
What specific copyright allegations does Britannica make against OpenAI's ChatGPT?
Britannica alleges that ChatGPT's GPT-4 model is reproducing its copyrighted encyclopedia entries word for word without authorization. The lawsuit claims that the AI system has 'memorized' substantial portions of Britannica's content and can output near-verbatim copies of entire passages on demand.
How does Britannica characterize OpenAI's content reproduction in the lawsuit?
Britannica describes OpenAI's content reproduction as more of a 'copy-and-paste operation' rather than genuine learning or transformation of text. The complaint suggests that GPT-4 is essentially creating unauthorized copies of their copyrighted material, going beyond typical machine learning practices.
Which other organization has joined Britannica in the lawsuit against OpenAI?
Merriam-Webster has joined Britannica in the lawsuit against OpenAI, supporting the claim that the AI model is reproducing copyrighted dictionary and encyclopedic content without permission. Together, they are challenging OpenAI's training data practices and content generation methods.
Further Reading
- Encyclopedia Britannica Inc v OpenAI - Complaint — US District Court Filing
- Encyclopedia Britannica Sues OpenAI Over Alleged Use of Content to Train ChatGPT — ChatAI
- Britannica sues OpenAI over alleged misuse of reference materials — Investing.com
- Encyclopedia Britannica sues OpenAI over AI training — Global Banking and Finance Review
- Encyclopedia Britannica and Merriam-Webster sue OpenAI — The Next Web