Is it legal to train AI models on copyrighted books? It’s complicated
Legal and ethical questions surrounding the use of copyrighted literature to train generative AI models remain unresolved. AI developers have routinely used large corpora of scraped and digitized books without explicit licensing or author consent. While creators argue this practice infringes on intellectual property rights and disrupts publishing livelihoods, technology companies often cite fair use doctrines to defend training practices across the AI industry.