Skip to main content
Britannica vs. OpenAI: ChatGPT lawsuit over verbatim content copying, with Britannica logo and ChatGPT interface.

Editorial illustration for Britannica sues OpenAI, alleging ChatGPT copied its content verbatim

Britannica Sues OpenAI for ChatGPT Content Plagiarism

Britannica sues OpenAI, alleging ChatGPT copied its content verbatim

Updated: 3 min read

After 255 years, Encyclopedia Britannica has stopped publishing its arguments and started filing them in court. The publisher sued OpenAI this week with a shocking, simple allegation: ChatGPT didn't learn from its content. It memorized it.

The lawsuit claims the AI spits out whole passages from Britannica and its Merriam-Webster dictionary, verbatim. This isn't inspiration. It's a digital photocopier.

The lawsuit accuses OpenAI of outputting near-identical copies of Britannica and Merriam-Webster's content. According to Britannica, OpenAI repeatedly copied its content without permission, stating, "GPT-4 itself has 'memorized' much of Britannica's copyrighted content and will output near-verbatim copies of significant portions on demand. The memorized examples are unauthorized copies that [OpenAI] used to train their models, including GPT-4." The lawsuit goes on to include examples of responses from OpenAI's models side by side with Britannica's text, in which entire passages appear to match word for word. Britannica also claims that OpenAI has been "cannibalizing" its web traffic by generating responses that "substitute, or directly compete" with Britannica's content, rather than directing users to its website the way a traditional search engine would.

Every major AI lawsuit grapples with fair use. Britannica’s filing adds a brutal twist: the crime is in the output. Its side-by-side comparisons aim to prove GPT-4 regurgitates, plain and simple.

That shifts the legal debate from how data is ingested to whether paragraphs are stolen. The real killer for publishers, though, is the traffic claim. Why click through to a site when the chatbot serves the answer?

A business model built on facts evaporates. OpenAI scanned the library. Now it wants to be the library.

Britannica survived the internet. It may be the plaintiff that finally forces a judge to name the price.

Common Questions Answered

What specific copyright allegations does Britannica make against OpenAI's ChatGPT?

Britannica alleges that ChatGPT's GPT-4 model is reproducing its copyrighted encyclopedia entries word for word without authorization. The lawsuit claims that the AI system has 'memorized' substantial portions of Britannica's content and can output near-verbatim copies of entire passages on demand.

How does Britannica characterize OpenAI's content reproduction in the lawsuit?

Britannica describes OpenAI's content reproduction as more of a 'copy-and-paste operation' rather than genuine learning or transformation of text. The complaint suggests that GPT-4 is essentially creating unauthorized copies of their copyrighted material, going beyond typical machine learning practices.

Which other organization has joined Britannica in the lawsuit against OpenAI?

Merriam-Webster has joined Britannica in the lawsuit against OpenAI, supporting the claim that the AI model is reproducing copyrighted dictionary and encyclopedic content without permission. Together, they are challenging OpenAI's training data practices and content generation methods.

LIVE03:06Microsoft Confirms Copilot 'Super App' for This Year