Article may be outdated

This article is 12 days old. Some details may have changed since publication.

TechCrunch·4 min read·medium

Is it legal to train AI models on copyrighted books? It’s complicated

A
Amanda Silberling
Is it legal to train AI models on copyrighted books? It’s complicated
AI Summary

This article explores the legal complexities surrounding the use of copyrighted works to train AI models, citing a recent court ruling involving Anthropic. While the court penalized the company for using pirated content, it affirmed that training AI on copyrighted data is generally analogous to human learning and thus lawful.

Why it matters

This legal precedent significantly impacts the future of generative AI development and the rights of content creators.

Dive DeeperCreate a free account to unlock

You probably know by now that the AI models powering ChatGPT, Gemini, Claude, and other chatbots are trained on seemingly infinite databases of published works, containing hundreds of millions of books, online articles, academic papers, and basically anything you can find on the internet. Most published authors have, without their knowledge or consent, contributed to the development of the same AI tools that threaten to undermine their livelihoods. That seems illegal, right?

“I think one of the issues with this entire area of law and this entire area of technology is there’s a lot going on,” Cathy Gellis, an attorney with expertise in intellectual property, copyright, and technology, told TechCrunch. “It’s very complex and there are a lot of raw feelings about what is happening, both for and against.”

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologyai

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in