Skip to content
Sıfırıncı Dakika
Latest

OpenAI publishes 722 math papers produced by AI

AINew2 min read

In brief

OpenAI has published 722 mathematics papers generated by an unreleased internal frontier model, grouped into 372 result families on GitHub. According to the independent AGMAI group, the work includes results on hundreds of long-open problems, but the papers still need review and verification by the math community. OpenAI says some proofs have been checked with Lean and that the repository will be

  • 722 math papers produced by an internal frontier model
  • Public GitHub collection grouped into 372 result families
  • Evaluation across roughly 4,000 problems
  • Average result uses compute equal to ~3 hours of ChatGPT Pro thinking
  • Part of the proofs verified with Lean formal system
  • Reasoning summaries and compute estimates shared for 10 problems

What happened?

OpenAI has shared 722 mathematics papers produced by an unreleased internal frontier model in a public GitHub repository. The work was filtered by significance and organized into 372 result families.

The model worked on roughly 4,000 problems

According to OpenAI, the model was confronted with about 4,000 math problems during evaluation. The company says producing an average result took compute roughly equivalent to three hours of ChatGPT Pro thinking capacity.

Verification is still ongoing

The independent Advisory Group on Mathematics and Artificial Intelligence (AGMAI) says the collection contains results on hundreds of open problems that had long resisted solution. However, OpenAI itself notes the results are at different verification stages and that some not yet formalized papers may contain errors. A significant portion of the proofs was checked with Lean, a system that uses computers to verify the logical validity of mathematical proofs; not all papers have Lean verification yet.

Transparency and academic publishing debate

OpenAI also shared condensed reasoning summaries, compute estimates and the number of problems attempted for 10 selected problems. Examples include the irrationality measure of π, Mahler's conjectures, Kaplansky's direct finiteness conjecture and the three-dimensional relativistic Vlasov-Maxwell system. AGMAI, meanwhile, wants AI companies to publish mathematical results through academic channels where possible, be more transparent about the model, cost and methods used, and avoid turning mathematical discoveries into marketing tools.

What's next?

OpenAI says it will keep updating the GitHub repository as new formal verifications are completed, and that it will fund workshops, conferences and special programs to help people understand significant AI-derived mathematical results. The company adds that it is working to release the internal model behind the work responsibly.

Why it matters

A key moment for readers who want to see how far AI has come in scientific production and how that output should be verified.

Sources

Related stories