Is it legal to train AI models on copyrighted books? It’s complicated

https://techcrunch.com/wp-content/uploads/2018/02/tc-backlight-e1689786273147.png?w=1200

You probably know by now that the AI models powering ChatGPT, Gemini, Claude, and other chatbots are trained on seemingly infinite databases of published works, containing hundreds of millions of books, online articles, academic papers, and basically anything you can find on the internet. Most published authors have, without their knowledge or consent, contributed to the development of the same AI tools that threaten to undermine their livelihoods. That seems illegal, right?

The reality isn’t that simple.

“I think one of the issues with this entire area of law and this entire area of technology is there’s a lot going on,” Cathy Gellis, an attorney with expertise in intellectual property, copyright, and technology, told TechCrunch. “It’s very complex and there are a lot of raw feelings about what is happening, both for and against.”

Last year, in one of the first rulings of its kind, Judge William Alsup ordered Anthropic...

Copyright of this story solely belongs to techcrunch.com. To see the full text click HERE

Read more