OpenAI releases another batch of AI mathematics research results, solving hundreds of outstanding problemsOpenAI 发布又一批 AI 数学研究成果,攻破数百个悬而未决难题
In a series of documents containing 722 manuscripts and covering 372 result families, OpenAI announced a number of long-standing mathematical problems solved by an unpublished cutting-edge model. IT House notes that this initiative continues its series of breakthroughs, which...
IT之家 10 月 7 日消息,OpenAI 在一批包含 722 份手稿、涵盖 372 个结果家族的文件中,公布了某款未发布的前沿模型解决的多项长期存在的数学难题。 IT之家注意到,这一举措延续了其一系列突破性成果,这些…
Key Points要点速览
In a series of documents containing 722 manuscripts and covering 372 result families, OpenAI announced a number of long-standing mathematical problems solved by an unpublished cutting-edge model. IT House notes that the move continues a series of groundbreaking results that have both impressed and unsettled some in the mathematical community, as well as raising questions about research ethics and academic behaviour. According to AGMAI, the newly formed independent advisory group of elite mathematicians responsible for communicating these results in a responsible manner, this announcement contains answers to “hundreds” of unsolved questions. The results have been expected for weeks, but until then, OpenAI had not specified what problems the model solved, nor given a precise release time. The company said in September that its model “has solved more than a hundred long-standing open puzzles covering the vast majority of branches of mathematics.” The manuscript and related materials released this time also contain a part of the summary of the model inference process, the estimation of computing power consumption, and the statistical data of the number of problems that have been tried to be solved. According to OpenAI, the amount of computing power consumed by an ordinary result is roughly equivalent to the level of computing power that ChatGPT Pro thinks about for three hours in a row. The Advisory Group on Mathematics and Artificial Intelligence (AGMAI) released its first recommendations in late September. The group calls on artificial intelligence laboratories to release mathematical research results in a timely manner through mature academic channels as much as possible, and at the same time, they should disclose detailed information such as the names, prompts, and computing power costs of the models used. The panel also exhorted AI companies “not to use publishing mathematical results as a marketing tool to promote their own models”; the agency believes that such practices will cause serious harm to the mathematical community as a whole. OpenAI describes the workflow of this release as follows: For this release, we published the results in a GitHub code repository, and at the same time developed a document revision and citation related processing specifications. We will also continue to explore other publishing options hosted by the academic community to align them with the guidance provided by the Advisory Board. For the release of the results of subsequent editions, we promise to further improve the quality of the manuscript, optimize the citation situation, mathematical content elaboration, and the format of the results for everyone to understand. Mathematicians also need to evaluate and assimilate these results, so the full impact of these studies may not be visible for some time. Previously, competitor laboratories such as OpenAI and Anthropic have produced a rapidly growing number of mathematical results. The entire mathematical community is still in the digestive stage, including the results of a Millennium Prize puzzle, which is the most well-known unsolved problem in mathematics. This year, the rapid entry of artificial intelligence laboratories into the discipline of mathematics, especially the way they publish information to the outside world, has provoked heated debate on scientific research paradigms, ethical norms, and how companies should thank human mathematicians. Artificial intelligence systems are built on the existing work of human mathematicians, and it is possible to rely on these predecessors to generate new results.
IT之家 10 月 7 日消息,OpenAI 在一批包含 722 份手稿、涵盖 372 个结果家族的文件中,公布了某款未发布的前沿模型解决的多项长期存在的数学难题。 IT之家注意到,这一举措延续了其一系列突破性成果,这些成果既给数学界部分人士留下了深刻印象,也引发了不安,同时还带来了关于研究伦理和学术行为的疑问。据负责以负责任的方式传达这些成果、新成立的精英数学家独立咨询小组 AGMAI 介绍,此次公布的内容包含对“数百个”未解决问题的解答。 数周以来外界一直在期待这批成果,不过在此之前,OpenAI 始终没有说明模型究竟解决了哪些问题,也没有给出确切的发布时间。该公司曾在九月份称,自家模型“已经解决了一百余个长期悬而未决的公开难题,覆盖数学的绝大多数分支领域”。本次对外发布的文稿以及相关资料还包含一部分模型推理过程摘要、算力消耗估算值,以及尝试求解过的问题数量统计数据。OpenAI 称,一项普通成果所消耗的算力,大致相当于 ChatGPT Pro 连续思考三小时的算力水平。 数学与人工智能咨询小组(Advisory Group on Mathematics and Artificial Intelligence,AGMAI)于九月下旬发布首批建议。该小组呼吁人工智能实验室尽可能通过成熟的学术渠道及时对外发布数学研究成果,同时应当披露所使用模型的名称、提示词、算力开销等详细信息。小组还恳切规劝人工智能企业,“不要把发布数学成果当作推广自家模型的营销手段”;该机构认为,这类做法会给整个数学社群带来严重损害。 OpenAI 是这样描述本次发布的工作流程的: 针对本次发布,我们将成果发布在一处 GitHub 代码仓库当中,同时制定了文稿修订与引用相关的处理规范。我们还会继续探寻其他由学术社群托管的发布方案,使之符合该咨询委员会提出的各项指导意见。对于后续版本的成果发布,我们承诺会进一步完善文稿质量,优化引用情况、数学内容阐述以及成果展示形式,便于大家理解。 数学家还需要对这批成果开展评估、消化吸收,因此这批研究带来的完整影响或许尚需一段时间才能够显现。 此前 OpenAI 以及 Anthropic 等竞争对手实验室已经产出数量迅速增长的一批数学成果,整个数学界仍处在消化阶段,其中甚至包含一道千禧年大奖难题的相关结果,该难题属于数学领域最知名的未解问题。 今年各家人工智能实验室快速闯入数学这门学科,尤其是它们对外发布消息的方式,已经激起激烈争论,议题涉及科研范式、伦理规范,以及企业应当如何向人类数学家致谢;人工智能系统正是建立在人类数学家的已有工作之上,并有可能依托这些前人工作生成新结果。
想马上用起来?去「AI工具」栏挑一个直接下载,或在「AI教程」里跟着图文步骤做一遍。
去 AI工具 → 看 AI教程 →阅读与点赞数据保存在你的浏览器本地,欢迎留下你的想法。
今日正能量学一点,用一点;今天种下的种子,会长成明天的能力。去免费下载专区 →广告
评论 文明发言,让讨论更有价值