this post was submitted on 09 Aug 2025

510 points (98.5% liked)

News

37724 readers

1215 users here now

Welcome to the News community!

Rules:

1. Be civil

Attack the argument, not the person. No racism/sexism/bigotry. Good faith argumentation only. This includes accusing another user of being a bot or paid actor. Trolling is uncivil and is grounds for removal and/or a community ban. Do not respond to rule-breaking content; report it and move on.

2. All posts should contain a source (url) that is as reliable and unbiased as possible and must only contain one link.

Obvious biased sources will be removed at the mods’ discretion. Supporting links can be added in comments or posted separately but not to the post body. Sources may be checked for reliability using Wikipedia, MBFC, AdFontes, GroundNews, etc.

3. No bots, spam or self-promotion.

Only approved bots, which follow the guidelines for bots set by the instance, are allowed.

4. Post titles should be the same as the article used as source. Clickbait titles may be removed.

Posts which titles don’t match the source may be removed. If the site changed their headline, we may ask you to update the post title. Clickbait titles use hyperbolic language and do not accurately describe the article content. When necessary, post titles may be edited, clearly marked with [brackets], but may never be used to editorialize or comment on the content.

5. Only recent news is allowed.

Posts must be news from the most recent 30 days.

6. All posts must be news articles.

No opinion pieces, Listicles, editorials, videos, blogs, press releases, or celebrity gossip will be allowed. All posts will be judged on a case-by-case basis. Mods may use discretion to pre-approve videos or press releases from highly credible sources that provide unique, newsworthy content not available or possible in another format.

7. No duplicate posts.

If an article has already been posted, it will be removed. Different articles reporting on the same subject are permitted. If the post that matches your post is very old, we refer you to rule 5.

8. Misinformation is prohibited.

Misinformation / propaganda is strictly prohibited. Any comment or post containing or linking to misinformation will be removed. If you feel that your post has been removed in error, credible sources must be provided.

9. No link shorteners or news aggregators.

All posts must link to original article sources. You may include archival links in the post description. News aggregators such as Yahoo, Google, Hacker News, etc. should be avoided in favor of the original source link. Newswire services such as AP, Reuters, or AFP, are frequently republished and may be shared from other credible sources.

10. Don't copy entire article in your post body

For copyright reasons, you are not allowed to copy an entire article into your post body. This is an instance wide rule, that is strictly enforced in this community.

founded 3 years ago

MODERATORS

JonsJava@lemmy.world

gedaliyah@lemmy.world

little_cow@lemmy.world

510

AI industry horrified to face largest copyright class action ever certified (arstechnica.com)

submitted 9 months ago* (last edited 9 months ago) by MicroWave@lemmy.world to c/news@lemmy.world

86 comments fedilink hide all child comments

Copyright class actions could financially ruin AI industry, trade groups say.

AI industry groups are urging an appeals court to block what they say is the largest copyright class action ever certified. They've warned that a single lawsuit raised by three authors over Anthropic's AI training now threatens to "financially ruin" the entire AI industry if up to 7 million claimants end up joining the litigation and forcing a settlement.

Last week, Anthropic petitioned to appeal the class certification, urging the court to weigh questions that the district court judge, William Alsup, seemingly did not. Alsup allegedly failed to conduct a "rigorous analysis" of the potential class and instead based his judgment on his "50 years" of experience, Anthropic said.

you are viewing a single comment's thread
view the rest of the comments

[–] N0t_5ure@lemmy.world 173 points 9 months ago (56 children)

"If we have to pay for the intellectual property that we steal and repackage, our whole business model will be destroyed!"

[–] errer@lemmy.world 82 points 9 months ago (2 children)

One thing this whole AI training debacle has done for me: made me completely guilt-free in pirating things. Copyright law has been bullshit since Disney stuck their finger in it and if megacorps can get away with massively violating it, I’m not going to give a shit about violating it myself.

[–] bss03@infosec.pub 28 points 9 months ago

For me it was Disney floating the idea of asking the wrongful death suit be dismissed because of the liability waiver in a Disney+ free trial.

I have the $$$, but I don't agree with the terms for any of the streaming services, so I'll just sail the seven seas and toss a doubloon (coin) to independent creators (my witchers) when I can.

[–] makyo@lemmy.world 12 points 9 months ago

I'm pretty much there too, the whole industry consolidates on the new things and charges us as they make it worse. And there can be some arguments to be made over the benefits of AI but we all know that it will not be immune to the entshitification that has already ruined all the things before it

[–] aramis87@fedia.io 72 points 9 months ago

If I downloaded ten movies to watch with my nephew in the cancer ward, they'd sue me into oblivion. Download tens of millions of books and claiming your business model depends on doesn't make it okay. And sharing movies with my sick nephew would cause less harm to society and to the environment than AI does.

[–] ThePantser@sh.itjust.works 22 points 9 months ago* (last edited 9 months ago) (2 children)

I started my own streaming service with pirated content. My business model depends on that data on my server.

Same thing but for some reason it's different. They hate when we use their laws against them. Let's root they rule against this class action so we can all benefit from copyright being thrown out. Or alternatively it kills AI companies, either way is a win.

[–] Knock_Knock_Lemmy_In@lemmy.world 1 points 9 months ago

They hate when we use their laws against them

YSK. They, we and them in this sentence mean different things to different people.

[–] FauxLiving@lemmy.world 2 points 9 months ago (1 children)

“If we have to pay for the intellectual property that we steal and repackage, our whole business model will be destroyed!”

They are very likely to be civilly liable for uploading the books.

That's largely irrelevant because the judge already ruled that using copyrighted material to train an LLM was fair use.

The judge did so in a summary motion, which means that they have to read all of the evidence in a manner most favorable to the plaintiff and they still decided that there is no way for the plaintiff to succeed in their copyright claim about training LLMs because it was so obviously fair use.

[–] N0t_5ure@lemmy.world 2 points 9 months ago (1 children)

Read the Order, which is Exhibit B to Antrhopic's appellate brief.

Anthropic admitted that they pirated millions of books like Meta did, in order to create a massive central library for training AI that they permanently retained, and now assert that if they are held responsible for this theft of IP it will destroy the entire AI industry. In other words, it appears that this is common practice in the AI industry to avoid the prohibitive cost of paying for the works they copy. Given that Meta, one of the wealthiest companies in the world, did the same exact thing, it reinforces the understanding that piracy to avoid paying for their libraries is a central component of training AI.

While the lower court did rule that training an LLM on copyrighted material was a fair use, it expressly did not rule that derivative works produced are protected by fair use and preserved the issue for further litigation:

Again, Authors concede that training LLMs did not result in any exact copies nor even infringing knockoffs of their works being provided to the public. If that were not so, this would be a different case. Authors remain free to bring that case in the future should such facts develop.

Emphasis added. In other words, Anthropic can still face liability if it's trained AI produces knockoff works.

Finally, the Court held

The downloaded pirated copies used to build a central library were not justified by a fair use. Every factor points against fair use. Anthropic employees said copies of works (pirated ones, too) would be retained "forever" for "general purpose" even after Anthropic determined they would never be used for training LLMs. A separate justification was required for each use. None is even offered here except for Anthropic's pocketbook and convenience. ... We will have a trial on the pirated copies used to create Anthropic's central library and the resulting damages, actual or statutory (including for willfulness). That Anthropic later bought a copy of a book it earlier stole off the internet will not absolve it of liability for the theft but it may affect the extent of statutory damages. Nothing is foreclosed as to any other copies flowing from library copies for uses other than for training LLMs.

Emphasis in original.

So to summarize, Anthropic apparently used the industry standard of piracy to build a massive book library to train it's LLMs. Plaintiffs did not dispute that training an LLM on a copyrighted work is fair use, but did not have sufficient information to assert that knockoff works were produced by the trained LLMs, and the Court preserved that issue for later litigation if the plaintiffs sought to bring such a claim. Finally, the Court noted that Anthropic built it's database for training it's LLMs through massive straight-up piracy. I think my original comment was a fair assessment.

load more comments (1 replies)

[–] Bakkoda@sh.itjust.works 2 points 9 months ago* (last edited 9 months ago)

That's unfair. They also have to sue people who infringe on "their" IP. You just don't understand what it's like to a content creator.

load more comments (51 replies)