Home
Encyclopedia Britannica Sues OpenAI, Alleging Knowledge Exploitation Amid Technological Change

As copyright disputes heat up across the AI landscape, traditional guardians of knowledge are no longer staying on the sidelines. This Friday, the globally respected Encyclopedia Britannica and its subsidiary Merriam-Webster officially filed a lawsuit, accusing OpenAI of using their copyrighted material without permission for "massive" AI model training.
This marks another major legal move after the two institutions sued AI search engine Perplexity last year. According to the complaint, Encyclopedia Britannica alleges that OpenAI illegally copied nearly 100,000 online articles, encyclopedia entries, and dictionary definitions to train its GPT series of large language models.
Traffic Drain and Near-Verbatim "Plagiarism"
The plaintiff provides several examples in the suit, noting that ChatGPT generates responses that are almost identical to content from Encyclopedia Britannica when answering user queries. What's more troubling for publishers is that AI-generated summaries directly answer users' questions within the chat interface, siphoning off traffic that once went to the encyclopedia's website — directly undermining its traffic-dependent revenue model.
False Attribution: A New Claim Under the Lanham Act
Beyond copyright infringement, the lawsuit also invokes trademark provisions under the Lanham Act. The plaintiffs claim that ChatGPT sometimes fabricates facts (the so-called "hallucination" phenomenon) and incorrectly attributes those fabricated facts to Encyclopedia Britannica. This misleading behavior not only damages the encyclopedia's authoritative reputation but also leads the public to mistakenly believe that its content use has been officially authorized or endorsed.
AI Industry's Future Amid a Legal Storm
Currently, AI giants like OpenAI and Anthropic face a wave of lawsuits from authors, publishers, and news organizations. While some judges have previously considered AI training to have "transformative" characteristics, using pirated materials remains illegal. For example, Anthropic once paid a $1.5 billion settlement for using pirated e-books to train its models.
Now that traditional knowledge authorities are taking legal action, the "black box" practices of generative AI companies — which have long refused to disclose their training data sources — are facing unprecedented scrutiny. The outcome of this lawsuit will directly shape the power balance between the future AI industry and traditional copyright holders.
Related article
ByteDance’s Seed launches global campus drive, offering virtual shares to win top large model talent
In the competitive landscape of large language models, securing top-tier talent remains the most critical strategic asset.On April 1st, ByteDance announced the launch of its Seed global campus recruitment initiative, part of its large model talent de
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a
Related Special Topic Recommendations
Comments (0)
0/500

As copyright disputes heat up across the AI landscape, traditional guardians of knowledge are no longer staying on the sidelines. This Friday, the globally respected Encyclopedia Britannica and its subsidiary Merriam-Webster officially filed a lawsuit, accusing OpenAI of using their copyrighted material without permission for "massive" AI model training.
This marks another major legal move after the two institutions sued AI search engine Perplexity last year. According to the complaint, Encyclopedia Britannica alleges that OpenAI illegally copied nearly 100,000 online articles, encyclopedia entries, and dictionary definitions to train its GPT series of large language models.
Traffic Drain and Near-Verbatim "Plagiarism"
The plaintiff provides several examples in the suit, noting that ChatGPT generates responses that are almost identical to content from Encyclopedia Britannica when answering user queries. What's more troubling for publishers is that AI-generated summaries directly answer users' questions within the chat interface, siphoning off traffic that once went to the encyclopedia's website — directly undermining its traffic-dependent revenue model.
False Attribution: A New Claim Under the Lanham Act
Beyond copyright infringement, the lawsuit also invokes trademark provisions under the Lanham Act. The plaintiffs claim that ChatGPT sometimes fabricates facts (the so-called "hallucination" phenomenon) and incorrectly attributes those fabricated facts to Encyclopedia Britannica. This misleading behavior not only damages the encyclopedia's authoritative reputation but also leads the public to mistakenly believe that its content use has been officially authorized or endorsed.
AI Industry's Future Amid a Legal Storm
Currently, AI giants like OpenAI and Anthropic face a wave of lawsuits from authors, publishers, and news organizations. While some judges have previously considered AI training to have "transformative" characteristics, using pirated materials remains illegal. For example, Anthropic once paid a $1.5 billion settlement for using pirated e-books to train its models.
Now that traditional knowledge authorities are taking legal action, the "black box" practices of generative AI companies — which have long refused to disclose their training data sources — are facing unprecedented scrutiny. The outcome of this lawsuit will directly shape the power balance between the future AI industry and traditional copyright holders.
ByteDance’s Seed launches global campus drive, offering virtual shares to win top large model talent
In the competitive landscape of large language models, securing top-tier talent remains the most critical strategic asset.On April 1st, ByteDance announced the launch of its Seed global campus recruitment initiative, part of its large model talent de
Suno to Watermark Songs Amid Legal Battles
Suno, the platform enabling users to generate AI-created music, has unveiled new features to label platform-produced tracks, restrict downloads, and update community standards to curb unauthorized replicas. These updates arrive as Suno confronts mult
Musk Admits Grok Build Leaked User Code, Promises to Erase All Historical Data
Elon Musk directly addressed the privacy controversy surrounding Grok Build, beginning with a simple "True" to confirm the incident's validity. He pledged that all user data previously uploaded to SpaceXAI would be permanently erased, stating, "not a











