惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
Docker
Apple Machine Learning Research
Apple Machine Learning Research
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 三生石上(FineUI控件)
月光博客
月光博客
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
WordPress大学
WordPress大学
Hugging Face - Blog
Hugging Face - Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
M
MIT News - Artificial intelligence
腾讯CDC
B
Blog RSS Feed
H
Help Net Security
J
Java Code Geeks
有赞技术团队
有赞技术团队
Y
Y Combinator Blog
博客园_首页
Last Week in AI
Last Week in AI
博客园 - 【当耐特】
博客园 - Franky
B
Blog
MongoDB | Blog
MongoDB | Blog
博客园 - 叶小钗
Martin Fowler
Martin Fowler

The Verge

The Verge The Verge The Verge The Verge The Verge The Verge The Verge The Verge The Verge The Verge The Verge Govee’s multicolor ceiling light doubles as a low-res screen The plan to quietly kill Coyote v. Acme blew up in David Zaslav’s face AirPods, Touch Bars, and the rest of Tim Cook’s legacy I don’t think Gwyneth Paltrow knows what a peptide is Brendan Carr’s war on wokeness targets inclusive children’s television Anthropic’s Mythos breach was humiliating Ikea’s new inflatable chair doesn’t look like an inflatable chair Inside Microsoft’s wave of executive departures Netflix can’t seem to follow-up its biggest shows The Iranian women Trump ‘saved’ from execution are simultaneously real and AI-manipulated Elon Musk admits that millions of Tesla vehicles won’t get unsupervised FSD Tesla’s revenue rises again as it prepares for more AI and robotics Former MrBeast exec sues over ‘years’ of alleged harassment Watch Sony’s elite ping-pong robot beat top-ranked players Anthropic’s Mythos rollout has missed America’s cybersecurity agency Will a new CEO realize Apple’s smart home potential? It’s amazing how good Alienware’s $350 OLED monitor is Call of Duty never made much sense for Xbox Game Pass BMW’s flagship 7 Series gets its ‘Neue Klasse’ upgrade
Meta sued by major book publishers over copyright infring...
Emma Roth · 2026-05-06 · via The Verge

Meta is facing a class action lawsuit filed by five major book publishers and one author over claims the company “engaged in one of the most massive infringements of copyrighted materials in history” when training its Llama AI models, as reported earlier by The New York Times. In their suit, Macmillan, McGraw-Hill, Elsevier, Hachette, Cengage, and author Scott Turow allege that Meta “repeatedly copied” their books and journal articles without permission.

The lawsuit accuses Meta of knowingly ripping copyrighted work from “notorious pirate sites,” such as LibGen, Anna’s Archive, Sci-Hub, Sci-Mag, and others, and then feeding that material into its AI model. It also claims that Meta trained Llama with information inside the Common Crawl dataset, which is allegedly “full of unauthorized copies of copyrighted works.” As a result, Llama “outputs verbatim and near-verbatim substitutes” of copyrighted material:

For example, when prompted with two brief sentences from Cengage’s best-selling textbook, Calculus: Early Transcendentals, 9th edition, by James Stewart, Llama begins reproducing word-for-word the continuation of the section.

Several authors have already sued Meta for alleged copyright infringement, which brought to light the company’s internal discussions about how to handle “media coverage suggesting we have used a dataset we know to be pirated.” Last year, a federal judge ruled in favor of Meta in one of these lawsuits, though he pointed out that his ruling “does not stand for the proposition that Meta’s use of copyrighted materials to train its language models is lawful.”

A group of authors also sued Anthropic over copyright infringement. While a federal judge ruled that training AI models on legally purchased books without permission is considered fair use, he allowed the authors to move forward with a class action lawsuit over the “millions” of works Anthropic allegedly pirated. Anthropic agreed to pay writers $1.5 billion last year to settle the class action lawsuit.

Turow and the group of publishers are suing Meta for damages, and ask that the court order the company to block its allegedly unlawful activities. They also ask the court to require the company to provide a list of books, journal articles, and other copyrighted works that it trained its Llama AI models on.

“AI is powering transformative innovations, productivity and creativity for individuals and companies, and courts have rightly found that training AI on copyrighted material can qualify as fair use,” Meta spokesperson Dave Arnold said in an emailed statement to The Verge. “We will fight this lawsuit aggressively.”

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.