Skip to main content
OpenAI logo on a screen, symbolizing AI solving complex math problems, representing their new solutions.

Editorial illustration for OpenAI Publishes Solutions to 'Hundreds' of Open Math Questions

OpenAI Solves Hundreds of Unsolved Math Problems

OpenAI Publishes Solutions to 'Hundreds' of Open Math Questions

• 3 min read

OpenAI has put out 722 manuscripts covering 372 result families, presented as solutions produced by an unreleased frontier model to a batch of long-standing math problems. The company first flagged this work in September, when it said the

OpenAI has revealed solutions to a number of long-standing mathematics problems produced by an unreleased frontier model in a batch of 722 manuscripts, covering 372 result families that group related papers.

Why this matters

For developers and researchers watching the AI-for-math space, the numbers here matter more than the headline. Seven hundred twenty-two manuscripts, 372 result families, produced by a model nobody outside OpenAI can test or verify. That's the part we keep coming back to.

An unreleased system generating hundreds of solutions is either a genuinely significant moment for automated reasoning or a case study in how hard it is to audit claims when the underlying model stays locked away. AGMAI's formation tells us the mathematical community isn't taking this on faith, and the "research ethics and academic conduct" questions mentioned alongside the release suggest working mathematicians are uneasy about how credit, verification, and peer review function when a closed model floods the field with output faster than anyone can check it. For founders building on top of frontier models, the lesson is practical: impressive output without reproducibility isn't a result, it's a claim.

Watch whether AGMAI's review confirms these solutions hold up, and whether OpenAI ever opens the model itself to outside scrutiny.

Common Questions Answered

How many manuscripts did OpenAI publish as solutions to open math problems?

OpenAI published 722 manuscripts covering 372 result families that group related mathematical papers. These solutions were produced by an unreleased frontier model working on long-standing mathematics problems that had previously remained unsolved.

What is significant about the unreleased frontier model generating these math solutions?

The unreleased frontier model produced hundreds of solutions to open math questions, but since the model remains locked away and inaccessible to external researchers, it is difficult to independently verify or audit these claims. This raises important questions about how to properly evaluate AI breakthroughs when the underlying system cannot be tested by the broader scientific community.

When did OpenAI first announce this mathematical breakthrough work?

OpenAI first flagged this work in September, when the company initially announced that an unreleased frontier model had produced solutions to long-standing math problems. The company subsequently published the full batch of 722 manuscripts covering the 372 result families.

Why do the specific numbers of manuscripts and result families matter more than the headline?

For developers and researchers in the AI-for-math space, the scale of the output (722 manuscripts across 372 result families) is more meaningful than the headline because it demonstrates the scope of automated reasoning capabilities. However, the fact that these solutions come from an unreleased and unverifiable model makes it challenging to assess whether this represents a genuinely significant breakthrough or simply demonstrates the difficulty of auditing AI claims when the underlying system remains inaccessible.

LIVE01:38OpenAI Publishes Solutions to 'Hundreds' of Open Math Questions