Dataset Origin vs. Training Fair Use数据集来源与训练合理使用
Courts distinguished between training AI models on lawful purchases and acquiring pirated texts from shadow libraries.法院对在合法购买文本上训练AI模型与从盗版网站获取图书作出了区别认定。
Pirated Books Covered涵盖盗版图书数量
482,000 titles48.2万本
2024-20262024-2026年
- Retired Judge William Alsup ruled that training AI models on lawfully purchased books constitutes fair use.已退休的阿尔苏普法官曾裁定在合法购买的图书上训练AI模型属于合理使用。
- Liability stemmed specifically from Anthropic acquiring millions of copyrighted works through unauthorized pirate websites.法律责任专门源于Anthropic通过未经授权的盗版网站获取数百万本受版权保护的作品。