AI News AI资讯 6h ago Updated 1h ago 更新于 1小时前 49

Harry Potter publisher to receive millions in Anthropic copyright settlement 《哈利·波特》出版商将在Anthropic版权和解中获得数百万英镑赔偿

Anthropic agreed to a $1.5 billion copyright settlement with thousands of authors and publishers regarding the use of their protected works to train AI models like Claude. Bloomsbury, publisher of Harry Potter, is a major beneficiary, with 14,087 titles listed in the agreement, receiving approximately $19 million after fees. The settlement covers 482,000 works, with 91% already claimed, marking the largest known copyright recovery in history and setting a precedent for AI training data compensat Anthropic达成15亿美元版权和解协议,出版商Bloomsbury因旗下14,087部作品获赔约1900万美元。 该案件由作家Andrea Bartz等人于2024年提起,被视为历史上规模最大的版权赔偿案之一。 和解金在扣除约10%律师费后分配,Bloomsbury将分期收款并与作者分成。 超过91%的受案作品权利人已认领份额,标志着AI训练数据版权争议的重要里程碑。 此事件凸显了AI公司使用受版权保护内容进行模型训练所面临的巨大法律与财务风险。

75
Hot 热度
65
Quality 质量
70
Impact 影响力

Analysis 深度分析

TL;DR

  • Anthropic agreed to a $1.5 billion copyright settlement with thousands of authors and publishers regarding the use of their protected works to train AI models like Claude.
  • Bloomsbury, publisher of Harry Potter, is a major beneficiary, with 14,087 titles listed in the agreement, receiving approximately $19 million after fees.
  • The settlement covers 482,000 works, with 91% already claimed, marking the largest known copyright recovery in history and setting a precedent for AI training data compensation.

Why It Matters

This settlement establishes a critical financial precedent for how AI companies must compensate creators whose intellectual property is used in training datasets, challenging the "fair use" defense previously relied upon by many US-based AI firms. For AI practitioners and researchers, it signals a shift toward mandatory licensing or payment structures for high-quality textual data, potentially increasing operational costs but ensuring more sustainable and legally compliant data sourcing.

Technical Details

  • Settlement Scope: The agreement involves a total value of $1.5 billion (£1.12 billion) covering 482,000 works across various genres, including novels, news articles, and academic texts.
  • Compensation Structure: Publishers like Bloomsbury receive bulk payouts based on the number of titles listed (e.g., ~$3,000 per title), with proceeds split between the publisher and the authors after deducting approximately 10% for attorney fees and expenses.
  • Legal Context: The lawsuit was initiated in 2024 by authors including Andrea Bartz, arguing that AI companies failed to seek permission or pay for copyrighted material used in training generative AI chatbots.
  • Adoption Rate: The judge noted that over 91% of the covered works have been claimed by rights holders, indicating broad acceptance of the settlement terms among the affected community.

Industry Insight

  • Shift from Fair Use to Licensing: AI developers should anticipate increased regulatory pressure and legal risks associated with unlicensed data scraping, necessitating a transition toward explicit licensing agreements with content creators and publishers.
  • Cost Implications for Model Training: The financial scale of this settlement suggests that high-quality, copyrighted text data will become a premium asset, likely driving up the cost of training large language models and encouraging investment in synthetic data or openly licensed alternatives.
  • Publisher-AI Partnerships: Traditional publishing houses are positioning themselves as gatekeepers for AI training data, creating new revenue streams through licensing deals (as seen with Bloomsbury's prior AI licensing announcement), which may influence future collaborations between tech firms and media entities.

TL;DR

  • Anthropic达成15亿美元版权和解协议,出版商Bloomsbury因旗下14,087部作品获赔约1900万美元。
  • 该案件由作家Andrea Bartz等人于2024年提起,被视为历史上规模最大的版权赔偿案之一。
  • 和解金在扣除约10%律师费后分配,Bloomsbury将分期收款并与作者分成。
  • 超过91%的受案作品权利人已认领份额,标志着AI训练数据版权争议的重要里程碑。
  • 此事件凸显了AI公司使用受版权保护内容进行模型训练所面临的巨大法律与财务风险。

为什么值得看

对于AI从业者和内容创作者而言,此案确立了AI巨头为训练数据付费的先例,打破了“合理使用”的绝对主导地位。它预示着未来AI商业化必须建立合规的数据获取机制,直接影响模型训练成本结构和版权授权模式的发展。

技术解析

  • 和解规模与分配:Anthropic支付15亿美元(约11.2亿英镑),其中Bloomsbury作为受益方获得约1900万美元。每部作品提议赔偿约3000美元,资金将在财政年度下半年分期支付。
  • 案件背景与覆盖范围:诉讼始于2024年,涉及48.2万部作品,目前91%的作品权利人已认领赔偿。法官认定该和解为受影响作者和出版商提供了“有意义的救济”。
  • 法律争议焦点:核心在于AI公司是否有权在未获许可的情况下使用受版权保护的文本(如小说、新闻)训练聊天机器人(如Claude)。美国AI公司主张“合理使用”,而创作者要求事先许可或付费。
  • 行业应对策略:Bloomsbury此前已推出AI授权计划,允许学术作品用于训练并收取版税,作者可选择加入。这表明大型出版机构正从被动诉讼转向主动授权商业模式。

行业启示

  • 版权合规成为AI基础设施成本:AI公司无法再忽视数据版权问题,必须将版权许可费用纳入模型训练的成本结构,这可能推高高端AI模型的边际成本。
  • “先授权后使用”模式兴起:随着类似和解案的增多,建立合法的数据采购渠道(如与出版社、媒体集团合作)将成为AI企业获取高质量训练数据的标准路径。
  • 创作者议价能力增强:头部内容和知名IP持有者(如Bloomsbury)在AI生态中获得了更强的谈判筹码,未来可能出现更多基于数据使用的分层授权和收益分成机制。

Disclaimer: The above content is generated by AI and is for reference only. 免责声明:以上内容由 AI 生成,仅供参考。

Claude Claude Policy 政策 Regulation 监管