AI News AI资讯 16h ago Updated 15h ago 更新于 15小时前 56

US Department of Justice backs fair use for AI training in landmark copyright case 美国司法部在标志性版权案件中支持AI训练的合理使用

The US Department of Justice filed a brief siding with AI companies in the New York Times v. OpenAI/Microsoft copyright case, arguing that training LLMs on copyrighted material qualifies as fair use The DOJ draws a sharp legal distinction between copying for training purposes and model outputs, noting that training copies are never publicly distributed and outputs "often if not always lack substantial similarity" to originals The DOJ invokes a Joan Didion analogy, arguing that requiring payment 美国司法部在纽约时报诉OpenAI版权案中明确支持AI公司,主张AI训练使用受版权保护的材料构成合理使用 DOJ核心论点:训练过程中的复制与模型输出存在法律区分,训练数据不会公开,输出通常与原作缺乏实质性相似 特朗普政府解雇版权局局长Perlmutter引发争议,因其拒绝为AI训练版权材料提供合法性背书 此案被视为AI版权领域的"风向标"案件,判决结果将深刻影响整个AI行业的版权政策走向

82
Hot 热度
72
Quality 质量
85
Impact 影响力

Analysis 深度分析

TL;DR

  • The US Department of Justice filed a brief siding with AI companies in the New York Times v. OpenAI/Microsoft copyright case, arguing that training LLMs on copyrighted material qualifies as fair use
  • The DOJ draws a sharp legal distinction between copying for training purposes and model outputs, noting that training copies are never publicly distributed and outputs "often if not always lack substantial similarity" to originals
  • The DOJ invokes a Joan Didion analogy, arguing that requiring payment whenever creators draw on prior works would stifle the creativity copyright law aims to protect
  • The filing directly challenges the US Copyright Office's earlier report that rejected blanket fair use for AI training, dismissing it as carrying no binding legal authority
  • Political overtones emerge as former Copyright Register Shira Perlmutter was fired by the Trump administration after her report, with Democrats alleging the dismissal was tied to her refusal to legitimize AI training on copyrighted works

Why It Matters

This DOJ filing represents the most significant federal government intervention in favor of AI companies in an ongoing copyright battle that could reshape the entire generative AI industry. A ruling favoring the DOJ's fair use interpretation would lock in the current training paradigm for billions of dollars in AI development, while a rejection could force companies into costly licensing deals or fundamentally alter how models are built.

Technical Details

  • The case centers on allegations that OpenAI and Microsoft used millions of New York Times articles without permission to train GPT-4 and competing products, with the Times seeking billions in damages and model destruction
  • The DOJ's fair use argument hinges on the four-factor test, particularly emphasizing that training copies are non-expressive and never made publicly available, and that model outputs lack substantial similarity to source works
  • The DOJ directly contests the US Copyright Office's 2023 report, arguing it ignored case law on case-by-case fair use analysis and misidentified the types of market harm recognized under copyright statute
  • The Joan Didion/Hemingway analogy is used to frame AI training as analogous to human creative learning processes, though critics note the scale difference between individual study and corporate mass reproduction
  • The filing references the Kadrey ruling to argue against treating a learning process and subsequent creative output as a single infringing use

Industry Insight

  • The DOJ's intervention signals strong administrative support for the AI industry's current training practices, but fair use remains an affirmative defense decided case-by-case—this does not establish a legal precedent, only strengthens the companies' position in litigation
  • AI companies should prepare for continued legal pressure regardless of this filing; the scale argument raised by the Copyright Office and critics will likely dominate future courtroom debates and legislative efforts
  • The political dimension—Perlmutter's firing and the Trump administration's pro-AI stance—suggests copyright policy may become increasingly partisan, creating uncertainty for long-term planning and encouraging companies to pursue licensing deals as a hedge against shifting legal winds

TL;DR

  • 美国司法部在纽约时报诉OpenAI版权案中明确支持AI公司,主张AI训练使用受版权保护的材料构成合理使用
  • DOJ核心论点:训练过程中的复制与模型输出存在法律区分,训练数据不会公开,输出通常与原作缺乏实质性相似
  • 特朗普政府解雇版权局局长Perlmutter引发争议,因其拒绝为AI训练版权材料提供合法性背书
  • 此案被视为AI版权领域的"风向标"案件,判决结果将深刻影响整个AI行业的版权政策走向

为什么值得看

本文揭示了美国政府在AI版权争议中的明确立场转向,对AI企业的合规策略和商业模式具有直接影响。同时,DOJ与版权局之间的分歧反映了政策制定层面对于AI发展的不同态度,值得从业者密切关注。

技术解析

  • 案件背景:纽约时报于2023年底在曼哈顿联邦法院起诉OpenAI和微软,指控数百万篇NYT文章被未经许可用于训练GPT-4等模型,索赔数十亿美元并要求销毁相关模型
  • DOJ法律论点:合理使用四要素分析中,DOJ强调训练复制与最终输出存在本质区别——训练时复制完整作品但从不公开,输出"通常甚至总是缺乏实质性相似"
  • 类比论证:DOJ引用作家琼·狄迪恩青少年时期抄写海明威作品学习的例子,论证要求创作者为借鉴付费将扼杀版权法旨在保护的创造力
  • 版权局反对意见:美国版权局报告明确拒绝 blanket fair use,指出AI以完美复制和远超人类的速度规模运作,商业应用已超出合理使用范畴
  • 政治因素:前版权注册官Shira Perlmutter在报告发布后被特朗普政府解雇,民主党议员称其因拒绝为AI训练版权材料背书而被解职,目前正挑战解雇决定

行业启示

  • 政策风向明确:特朗普政府已明确采取亲AI立场,司法部直接挑战版权局报告,表明行政分支对AI发展的支持态度,企业可据此调整合规预期
  • 版权风险仍需管理:尽管DOJ支持AI公司,但版权局报告和法院的最终判决仍存在不确定性,建议AI企业建立数据授权机制作为风险缓冲
  • 行业格局将重塑:此案作为"风向标"案件,其判决结果将确立AI训练版权的法律边界,影响所有依赖大规模文本数据的AI公司商业模式

Disclaimer: The above content is generated by AI and is for reference only. 免责声明:以上内容由 AI 生成,仅供参考。

LLM 大模型 Training 训练 Policy 政策 Regulation 监管 Legal AI 法律AI