OpenAI 开源数学研究仓库,含 722 篇模型手稿与 Lean 形式化证明
Mathematical manuscripts and supporting proof artifacts produced by OpenAI
OpenAI 在 GitHub 开源数学研究仓库,收录 722 篇由未发布内部模型生成的数学手稿与证明成果,按 372 个家族分类。平均每个结果使用约 3 小时 ChatGPT Pro 思考算力,评估过程共向模型提出约 4000 道题目;部分成果附带 Lean 形式化证明,涵盖 Riemann zeta 函数 Re(s)>11/12 无零点区域、CM 阿贝尔簇上的 Hodge 猜想等方向。
Readme
This repository contains mathematical manuscripts and supporting proof artifacts produced by an internal OpenAI model.
As part of model development, we evaluate our models on open research problems. We expanded these evaluations after performance on our existing mathematical evaluations saturated. Some outputs build upon earlier results produced by the models.
This collection includes results at different stages of verification. Not all have accompanying Lean formalizations. We will continue to update this repository with Lean formalizations as we obtain them.
Some of the unformalized results could have issues. We will endeavor to fix any such issues quickly. We are also exploring community-hosted repositories for these materials.
Navigating the collection
The current catalogue contains 722 manuscripts organized into 372 families. A family groups related papers, which may include a principal result, companion arguments, consequences, or alternative proofs. Each family is classified by mathematical discipline.
- Start with the overview for descriptions of the families.
- Use the manuscript map to find individual papers and their supporting materials.
- The
preprints/directory contains PDFs, source files, and manuscript-specific citation and build instructions. - The Lean library and formalization catalogue describe the available formal proofs, their associated papers, and verification configurations. See the Comparator instructions for additional checking instructions. Many, but not all, of the manuscripts have been formalized.
Reasoning summaries
We are also releasing abridged summaries of the model's reasoning, covering the following results:
How the results were produced
The vast majority of results were obtained with the same procedure using an unreleased internal OpenAI model. On average, each result used three hours of ChatGPT Pro thinking compute with that model. Over the course of the evaluation, the model was posed approximately 4,000 problems. Aggregating the output into result families and manuscripts and requiring an appropriate level of significance led to the catalog outlined above.
Exceptions to this fixed procedure include work on a zero-free region for the Riemann zeta function and proof of the Hodge Conjecture for CM abelian varieties. Additionally, the writeup for the Re(s) > 11/12 zero-free region for the Riemann zeta function was human edited for readability.
Versions and citations
We will preserve the public release history of this collection. Corrections and revisions will be recorded as new versions, with previously released versions remaining accessible.
To cite the individual manuscript, use the BibTeX block in its directory.
来源:Hacker News · github.com