Training AI on Copyrighted Books: Still a Legal Gray Zone
In one line: Whether it's legal to train AI on copyrighted books still has no clear answer.
Key points
- Most published authors reportedly had their work used as AI training data without their knowledge or consent.
- The central question is whether such training qualifies as "fair use" under U.S. copyright law.
- Authors argue that the AI tools threatening their livelihoods were built on their own writing.
- Related lawsuits continue, but legal outcomes appear to vary by case and jurisdiction.
Why it matters
How courts treat AI training data will shape both future model-development costs and how rights holders are compensated. Until clear precedent or legislation emerges, friction between creators and AI companies is likely to persist.
Read more
How this story unfolded
- LLM
- — Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.