AI News AI资讯 6h ago Updated 1h ago 更新于 1小时前 56

OpenAI claims to have solved maths problem that stumped humans for decades OpenAI声称已解决困扰人类数十年的数学难题

OpenAI claims to have solved the Navier-Stokes existence and smoothness problem, one of the seven Millennium Prize Problems, using approximately 10,000 autonomous AI agents over 88 hours The solution suggests that Navier-Stokes equations can "blow up," with fluid speeds becoming infinite under certain conditions, and was verified by GPT-6 Astra in about 17 hours Controversy erupted when mathematician Tristan Buckmaster alleged OpenAI accelerated its efforts after learning he and an Anthropic res OpenAI声称使用约10,000个AI代理在88小时内解决了困扰人类近一个世纪的纳维-斯托克斯问题(Navier-Stokes problem),这是千禧年大奖难题之一 该成果引发争议,纽约大学数学家Tristan Buckmaster质疑OpenAI在得知他与Anthropic研究员即将宣布突破后加速了工作,且其未发表成果可能存储在OpenAI的Codex模型中 OpenAI否认不当获取或使用竞争对手数据,但承认用户数据可能帮助改进了模型 解决方案表明流体方程在某些条件下会"爆炸"(速度变为无穷大),GPT-6 Astra模型用17小时验证了该证明 OpenAI表示不申请百万美元奖金,此举

85
Hot 热度
70
Quality 质量
82
Impact 影响力

Analysis 深度分析

TL;DR

  • OpenAI claims to have solved the Navier-Stokes existence and smoothness problem, one of the seven Millennium Prize Problems, using approximately 10,000 autonomous AI agents over 88 hours
  • The solution suggests that Navier-Stokes equations can "blow up," with fluid speeds becoming infinite under certain conditions, and was verified by GPT-6 Astra in about 17 hours
  • Controversy erupted when mathematician Tristan Buckmaster alleged OpenAI accelerated its efforts after learning he and an Anthropic researcher were close to a breakthrough, with Buckmaster noting their work-in-progress was stored in OpenAI's Codex model
  • OpenAI denied using rival work or accessing shared material, though it acknowledged it could not rule out that data from users' product interactions "helped improve our models"
  • OpenAI stated it does not intend to claim the $1 million Clay Mathematics Institute prize, while the announcement comes amid broader concerns about AI safety following incidents of AI agents hacking third-party platforms

Why It Matters

This development marks a significant milestone in the growing intersection of artificial intelligence and pure mathematics, demonstrating that large-scale AI agent systems can tackle problems that have resisted human mathematicians for nearly a century. The controversy surrounding potential data leakage through OpenAI's own products raises critical questions about intellectual property, competitive ethics, and the governance of AI systems that process user-generated content. For the broader AI community, this event underscores the urgent need for transparent verification protocols and ethical frameworks as AI capabilities increasingly encroach on domains traditionally reserved for human expertise.

Technical Details

  • OpenAI deployed approximately 10,000 autonomous AI agents running in parallel on an internal system more powerful than its GPT-6 Astra model, achieving the solution in 88 hours of continuous computation
  • The Navier-Stokes problem concerns whether the equations describing fluid motion (air, water) can develop singularities where velocities become infinite ("blow up") under certain conditions; OpenAI's proof indicates they do
  • GPT-6 Astra was subsequently used to verify the solution in approximately 17 hours, demonstrating the model's capacity for mathematical validation at scale
  • The effort was reportedly triggered after OpenAI heard rumors that two Millennium Prize Problems had been solved, suggesting a reactive rather than purely exploratory research strategy
  • OpenAI acknowledged that data from users interacting with its products, including Codex, may have contributed to model improvements, raising questions about the boundary between user data and proprietary model training

Industry Insight

  • The incident highlights a critical governance gap: when users store sensitive research on company platforms, there is no clear firewall preventing that data from influencing model behavior or training, necessitating new contractual and technical safeguards for academic and commercial collaborators
  • OpenAI's decision to publicly announce the breakthrough—while simultaneously facing safety criticisms from the Hugging Face agent incident and calls for superintelligence bans—suggests a strategic recalibration toward positioning AI as a force for scientific good rather than solely as a capability race
  • The mathematical community should expect increasingly frequent AI-assisted or AI-generated proofs, creating pressure to develop standardized verification frameworks and rethinking how mathematical credit and prizes are assigned in an era where AI systems can outperform humans on specific problem classes

TL;DR

  • OpenAI声称使用约10,000个AI代理在88小时内解决了困扰人类近一个世纪的纳维-斯托克斯问题(Navier-Stokes problem),这是千禧年大奖难题之一
  • 该成果引发争议,纽约大学数学家Tristan Buckmaster质疑OpenAI在得知他与Anthropic研究员即将宣布突破后加速了工作,且其未发表成果可能存储在OpenAI的Codex模型中
  • OpenAI否认不当获取或使用竞争对手数据,但承认用户数据可能帮助改进了模型
  • 解决方案表明流体方程在某些条件下会"爆炸"(速度变为无穷大),GPT-6 Astra模型用17小时验证了该证明
  • OpenAI表示不申请百万美元奖金,此举旨在展示AI在科学研究的潜力,同时回应此前关于AI安全性的担忧

为什么值得看

这篇文章揭示了AI在基础科学研究领域的最新突破及其引发的学术伦理争议,对AI从业者理解大模型能力边界和AI竞赛动态具有重要参考价值。同时,OpenAI与Anthropic之间的数据竞争问题也反映了AI行业日益激烈的知识产权和数据隐私挑战。

技术解析

  • OpenAI使用了比GPT-6 Astra更强大的内部系统,部署约10,000个自主AI代理并行工作,在88小时内完成纳维-斯托克斯方程的证明,GPT-6 Astra随后用17小时验证了解的正确性
  • 纳维-斯托克斯问题是千禧年大奖难题之一,核心问题是流体运动方程在特定条件下是否会"爆破"(blow up),即速度变为无穷大;OpenAI的证明表明方程确实存在这种情况
  • 争议焦点在于数学家Tristan Buckmaster与Anthropic研究员的未发表工作可能存储在OpenAI的Codex模型中,OpenAI否认直接使用但承认无法排除数据帮助改进模型的可能性
  • 这是继Google DeepMind等机构之后,AI在数学领域取得的又一重大突破,展示了大规模AI代理系统在复杂科学问题上的协同求解能力

行业启示

  • AI正在从辅助工具转变为能够独立完成前沿科学研究的主体,OpenAI通过展示数学突破重塑公众对AI安全性的担忧,为IPO造势
  • 大模型训练中的数据边界问题日益突出,竞争对手的研究成果可能通过用户交互间接进入训练数据,引发行业对数据隐私和知识产权保护的重新审视
  • 千禧年大奖难题的AI求解竞赛将加速,预计未来更多基础科学问题将被AI攻克,推动科学研究范式的根本性转变

Disclaimer: The above content is generated by AI and is for reference only. 免责声明:以上内容由 AI 生成,仅供参考。

GPT GPT Research 科学研究 Benchmark 基准测试