OpenAI Publishes 722 AI-Generated Math Manuscripts, With Verification Still Uneven
The collection includes computer-checkable proofs and selected reasoning summaries, but OpenAI warns that some unformalized results could contain errors.
Loading page…
The collection includes computer-checkable proofs and selected reasoning summaries, but OpenAI warns that some unformalized results could contain errors.
Listen to this story
OpenAI’s October 6, 2026 release gives researchers a repository to inspect work from its still-unreleased internal mathematics model, but not a uniformly verified corpus: many papers have Lean formalizations, while others do not and may contain problems. The 722 manuscripts are organized into 372 families, and the collection preserves earlier versions as corrections arrive. OpenAI says the work came from an evaluation of about 4,000 open problems; the practical value is access to papers and checkable artifacts, not access to the model or assurance that every result is correct.
OpenAI estimates each average result used compute equivalent to roughly three hours of ChatGPT Pro thinking; this does not give Pro users access to the model.
Only 10 selected results received abridged reasoning summaries, spanning topics from π’s irrationality exponent to a relativistic Vlasov–Maxwell system.
Manuscript families can include companion arguments, consequences, and alternative proofs, so 722 papers do not represent 722 distinct problems.
OpenAI’s internal mathematics model now has a public body of work to inspect: 722 manuscripts, grouped into 372 related families. Published on October 6, 2026, the GitHub collection pairs papers with supporting proof artifacts. But its contents are at different stages of verification, and the model that produced most of them remains unreleased.
Many manuscripts have accompanying formalizations in Lean, a programming language that lets computers check mathematical proofs. Those files provide a different way to examine an argument than reading the paper alone. OpenAI says it will add more formalizations as it obtains them; not every manuscript currently has one.
The repository explicitly warns that some unformalized results could have issues and says OpenAI will endeavor to fix them quickly. The release therefore comes with a verification distinction: it offers manuscripts and, for many, formal proof artifacts, rather than presenting the entire collection as equally checked.
Nor does each paper represent a separate problem solved. A family can collect a principal result, companion arguments, consequences or alternative proofs. Families are classified by mathematical discipline, giving readers a subject-level route into the catalogue before they examine individual manuscripts.
Individual papers in the current catalogue.
Groups of related papers, including companion arguments and alternative proofs.
For readers looking beyond the headline count, the repository supplies several entry points into the papers and their supporting materials:
OpenAI says it expanded its evaluations on open research problems after performance on existing mathematics evaluations saturated. Over the evaluation, the model received approximately 4,000 problems. Outputs were grouped into families and manuscripts, with an appropriate level of significance required for inclusion in the catalogue.
The vast majority of results followed the same procedure using the unreleased internal model. OpenAI estimates that the average result consumed compute equivalent to roughly three hours of ChatGPT Pro thinking. That is the company’s comparison for computing effort, not an announcement that ChatGPT Pro users can access this model.
Alongside the papers, OpenAI released 10 abridged summaries of the model’s reasoning. The selected subjects range from the irrationality exponent of π to a three-dimensional relativistic Vlasov–Maxwell system. These are summaries of selected results, not reasoning accounts accompanying all 722 manuscripts.
Corrections and revisions will appear as new versions, while earlier releases remain accessible. Each manuscript directory supplies a BibTeX citation block. That approach preserves a record of what was published even as papers or supporting artifacts change.
In its release announcement, OpenAI says it consulted the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study. The company drew on the group’s advice and public recommendations and is exploring community-hosted alternatives for the collection.
OpenAI also promises improvements to citations, mathematical exposition and presentation in future releases. It plans to fund workshops, conferences and special programs to help researchers understand major AI-produced results. Separately, it says it is working toward responsibly releasing the model itself; this publication does not launch it.
Loading discussion...
Join the conversation
Explain whether early scrutiny outweighs the risk of circulating errors.
Be the first to share a perspective or an experience.
Reader comments
Newest comments first. Replies stay oldest first.