AIToday
AI Business & IndustryITmedia AI+Published: Oct 7, 2026, 13:01 JST

OpenAI posts 722 AI-generated math papers on GitHub

OpenAI posts 722 AI-generated math papers on GitHub

3 Key Points

  1. What happened

    OpenAI said its internal frontier model produced 722 papers with Lean formal proofs, drawn from about 4,000 problems, each averaging roughly three hours of ChatGPT Pro thinking.

  2. Why it matters

    The proof logs, including ten inference summaries, let outside mathematicians check the claims directly — OpenAI had faced criticism over an earlier Navier-Stokes claim.

  3. What to watch

    Because not all results are formally verified, errors may surface; OpenAI says it will fix them, and it says a responsible public release of the model is in progress.

WHO IT HITSMathematics researchers and journal reviewers gain a large, checkable body of machine-generated results to vet, while OpenAI's own credibility with that community hinges on whether the formal proofs hold up.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

OpenAI says the results came from an internal model used to evaluate itself on unsolved research problems, the same model and procedure for most of them. Ten inference summaries were released, covering topics such as the irrationality measure of pi, the Mahler conjecture, the Kaplansky direct finiteness conjecture in characteristic 2, the isomorphism problem for free group factors, spontaneous magnetization of the quantum Heisenberg ferromagnet, and the 3D relativistic Vlasov-Maxwell system.

The publication follows an earlier dispute: in September OpenAI said its internal model solved the Navier-Stokes equation, one of the Millennium Prize problems, and a mathematician working on prior research questioned how it was done. OpenAI then announced a partnership with AGMAI, an independent group of mathematicians, and said the internal model had solved over 100 open problems. In August it said an internal version of its next flagship model, Astra, had produced new results on ten unsolved problems, and in October 2025 an executive's post claiming GPT-5 solved an Erdos problem was deleted after criticism.

AGMAI's September 29 recommendations call for publishing in repositories not controlled by AI companies, disclosing model names, prompts, inference summaries, time and compute costs, Lean formalization, and the number of unsolved problems and how problems were chosen. AGMAI also said it does not support the practice of AI companies testing internal models outsiders cannot use. Whether OpenAI's math gains are accepted is likely to depend less on the count of papers than on outside mathematicians reproducing and validating them, and OpenAI says it is working toward a responsible public release of the model.

FAQ
Are all 722 results formally verified?
No. OpenAI says many papers include Lean formal proofs, but not all, and results without formalization may contain problems.
Did AGMAI endorse these results?
No. AGMAI said its advice judged the impact of the results and did not endorse how OpenAI obtained them, and that it does not represent all of mathematics.
How much compute did each result take?
OpenAI says each result averaged roughly three hours of ChatGPT Pro's thinking feature, though the Riemann zeta and Hodge conjecture results were exceptions.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articlePolymarket Sets 35% Odds Nvidia Breaks Out in October