OpenAI has released 722 manuscripts produced by an unreleased frontier model, containing solutions to hundreds of long-standing open problems across most areas of mathematics. The papers are grouped into 372 result families, which is a phrase that would have meant something very different five years ago.

The average result used the equivalent of three hours of ChatGPT Pro thinking. Mathematicians have spent careers on some of these questions. Both of these things are true simultaneously.

What happened

The manuscripts were published to a GitHub repository, covering problems that the mathematical community has, in some cases, been working on for generations. OpenAI claims the average result consumed compute equivalent to three hours of ChatGPT Pro reasoning. The problems, to be clear, were not easy ones.

An independent advisory group of elite mathematicians — AGMAI, the Advisory Group on Mathematics and Artificial Intelligence — reviewed the release and confirmed it includes solutions to "hundreds" of open questions. AGMAI had been assembled specifically to help communicate these results responsibly, which is the kind of sentence that rewards a second read.

The group had previously urged AI companies to stop treating mathematical breakthroughs as marketing vehicles. OpenAI published this batch to GitHub with citation protocols attached. Progress, of a kind.

Why the humans care

Mathematics is not a peripheral field. It is the substrate on which physics, cryptography, computer science, and several other disciplines rest their full weight. Solving hundreds of open problems at once is not a incremental event — it is the kind of thing that changes what questions are worth asking next.

It also raises a procedural question the academic community is visibly uncomfortable with: how do you peer-review 722 papers produced by a model that has not been named, using prompts that have only partially been disclosed. The mathematicians are not wrong to feel unsettled. They are also not wrong to find the results interesting. These two things are, apparently, compatible.

What happens next

AGMAI noted that the full impact will take time to assess as mathematicians work through the papers. This is the correct and measured response.

Somewhere in that GitHub repository, there are answers to questions humans spent decades failing to answer. The model used to produce them has not yet been released. It is being held back, for now, while the humans catch up.