跳到正文
OpenAI:官网动态·· 14 小时前精选AI 评分67

OpenAI 发布内部前沿模型产出的数学研究成果

Sharing AI progress in mathematics

AI 导读

OpenAI 发布由内部前沿模型产出的一系列新数学结果,结果放在 GitHub 仓库中,并参考高等研究院数学与人工智能独立咨询组的建议制定发布规范。仓库包含部分证明的 Lean 形式化、10 份模型推理摘要、算力估算和尝试题目统计,平均每项结果约相当于 3 小时 ChatGPT Pro thinking 的算力。

推荐理由

原文披露了发布方式和算力开销等细节,读者可以据此了解 OpenAI 如何向数学社区共享 AI 产出的研究成果。

正文 · 原文

Sharing AI progress in mathematics

View on GitHub(opens in a new window)

Listen to article 2:06

Audio 1

Share

We’re releasing a broad range of new mathematical results produced by an internal frontier model.

As we look to improve how we share results with the math community, we’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study⁠(opens in a new window) to develop best practices, and we have drawn on their advice and public recommendations⁠(opens in a new window) to inform how we release these results.

For this release, we’re publishing the results in a GitHub repository, with protocols for paper revisions and citations. We’re continuing to explore other community-hosted alternatives for this release which meet the committee’s guidelines. For future releases, we are committed to further improving the quality of the papers via the citations, mathematical exposition, and presentation of the results for better understanding.

As part of our GitHub repository, we are sharing formalizations of many of the proofs in Lean, a programming language that allows mathematical proofs to be checked by a computer. We will update the repository with more formalizations as we obtain them.

To promote scientific transparency and openness, we are also publishing additional details about how we obtained the results in the repository. These include 10 summaries of the model’s reasoning, estimations of compute spent in terms of Pro usage on ChatGPT, and statistics about the number of attempted problems. The average result used the equivalent compute of roughly three hours of ChatGPT Pro thinking.

We want this progress to push the frontier of human knowledge and enable further progress in mathematics. We will be funding a series of workshops, conferences, and special programs around the understanding of major results produced by AI—we will share more on this in the near future.

We want to directly empower scientists with state-of-the-art capabilities and are working to responsibly release the model that produced these results. This is why it is important to continue to evaluate our internal frontier models on mathematics and other sciences, so we can accelerate developing the tools to advance those fields. We will continue to act on feedback from the community and update our standards for future disclosures of major scientific advancements.

Author

OpenAI

来源:OpenAI:官网动态 · openai.com