WorkAugust 23, 2026via TechCrunch AI

Is it legal to train AI models on copyrighted books? It’s complicated

Why it matters

A foundational legal question that will shape AI development and creator economics: whether AI labs can use copyrighted works without permission or compensation. The answer determines whether authors and publishers can extract licensing revenue from model training.

Key signals

  • Authors trained AI models without consent or knowledge
  • Legal status of copyrighted training data remains unsettled in courts
  • Question frames AI-creator livelihoods and IP compensation models
  • Fair-use doctrine vs. training-at-scale tension unresolved
  • Most published authors contributed to AI training datasets without knowledge or consent
  • Legal status of training on copyrighted books remains unresolved in courts
  • Direct impact on author livelihoods and future book publishing economics
  • Affects downstream AI model development practices and data sourcing

The hook

Authors are suing — but the courts haven't settled whether training AI on copyrighted books is fair use or theft.

Most published authors have, without their knowledge or consent, contributed to the development of the same AI tools that threaten to undermine their livelihoods. That seems illegal, right?

The week's key stories, every Friday.

ONE BRIEFING · EVERY FRIDAY · FREE

Free. Unsubscribe anytime.