a magnifying glass sitting on top of a piece of paper
Photo by Vlad Deep on Unsplash
Policy & Regulation

Is it legal to train AI models on copyrighted books? It’s complicated

Original source: TechCrunch 8/23/2026
🤖 This summary was written by AI based on public reporting from TechCrunch. It is not a reproduction of the original article. Read the original →

Millions of published authors have likely had their books incorporated into AI training datasets without ever being asked — or paid. This has sparked a wave of lawsuits and a heated public debate about whether such practices constitute copyright infringement or fall within legal doctrines like fair use.

The answer, frustratingly, isn't straightforward. Copyright law was built around human creativity and traditional forms of reproduction, not the way machine learning systems digest and synthesize vast libraries of text. Courts are only beginning to weigh in, and different judges have reached different preliminary conclusions.

Until clearer legal standards emerge — either through landmark rulings or new legislation — the publishing and AI industries remain in a tense standoff, with authors caught in the middle, uncertain whether they have meaningful recourse or any right to share in the profits their work may have helped generate.

Advertisement