OpenAI publishes 722 manuscripts of AI-generated mathematics results from unreleased model
OpenAI released 722 manuscripts covering 372 result families, produced by an unreleased model, in a GitHub repository. Mathematicians are sceptical of its claims about how they were produced, and of its disclosure compared with new advisory group guidelines.
1 / 2
The story, neutrally told
Left · 1OpenAI has published 722 manuscripts, grouped into 372 result families, containing solutions to long-standing mathematics problems produced by an unreleased frontier model. The VergeLC “in a batch of 722 manuscripts, covering 372 result families that group related papers” Read at The Verge ↗ Left · 1Scientific American reported that OpenAI posted the results to a GitHub repository at 6 P.M. EDT on 6 October, and that the company says each result resolves or makes substantial progress on a major open question in mathematics or theoretical computer science. Scientific AmericanLC “OpenAI revealed the results in a GitHub repository at 6 P.M. EDT.”“Each resolves or makes substantial progress on a major open question in mathematics or theoretical computer science, the company says.” Read at Scientific American ↗ Left · 1Among the claimed results are a solution to the four-dimensional Kakeya conjecture, improvements to important computer algorithms, and progress toward the Riemann hypothesis. Scientific AmericanLC “a solution to the four-dimensional Kakeya conjecture, improvements on some of the world’s most important computer algorithms” Read at Scientific American ↗
Left · 2The release follows an earlier OpenAI result on the Navier-Stokes problem, which Scientific American says came from a 10,000-agent swarm costing millions of dollars in computing; in September OpenAI said its model had resolved more than 100 long-standing open problems. Scientific AmericanLC “a 10,000-strong agentic swarm that cost millions of dollars in computing power” Read at Scientific American ↗ The VergeLC “resolved more than 100 long-standing open problems across most areas of mathematics.” Read at The Verge ↗ Left · 2An OpenAI spokesperson told Scientific American that the model produced almost every new result from a single prompt given to a single AI agent, though some may have taken multiple attempts; OpenAI separately says the average result used the equivalent of three hours of ChatGPT Pro thinking. Scientific AmericanLC “produced almost every one of the results in response to a single prompt handed to a single AI agent” Read at Scientific American ↗ The VergeLC “OpenAI claims the “average result” used the equivalent of three hours of ChatGPT Pro thinking.” Read at The Verge ↗ Left · 1MIT mathematician Andrew Sutherland said claims of single-agent results should be treated as unverified until the model is released and results can be replicated. Scientific AmericanLC “you should treat any claims about one-shotting problems with a single agent as unverified” Read at Scientific American ↗
Left · 1Many results have already been verified in Lean, a proof-checking language, according to Scientific American, but mathematicians will need months to judge whether the proofs contain novel and important ideas. Scientific AmericanLC “many of the results have already been verified in Lean”“The deluge will take mathematicians months to parse through and understand” Read at Scientific American ↗ Left · 2The Advisory Group on Mathematics and Artificial Intelligence (AGMAI), a newly formed independent group of mathematicians announced by OpenAI on 21 September, published recommendations in late September urging labs to release results promptly through established academic channels and to disclose the model, prompts and compute costs. The VergeLC “urging AI labs to release mathematical results promptly and through established academic channels where possible, while disclosing details such as the name of the model used, prompts, and compute costs” Read at The Verge ↗ Scientific AmericanLC “on September 21, the company announced it was assembling an independent advisory group of mathematicians” Read at Scientific American ↗ Left · 2The group also urged companies to avoid treating releases as marketing vehicles and criticised the use of proprietary internal models that mathematicians cannot access. The VergeLC “refrain from treating the release of mathematical results as marketing vehicles to promote their models” Read at The Verge ↗ Scientific AmericanLC “The recommendations are particularly critical of AI companies using “proprietary internal models”” Read at Scientific American ↗
Left · 1Scientific American reports OpenAI is disclosing only average compute time and some statistics, with no prompts; the company says it takes the guidelines seriously and is not bound by them, and is working to release the model as quickly and responsibly as possible. Scientific AmericanLC “OpenAI is choosing only to reveal the average compute time for a problem, with some additional statistics—and no prompts.”“the company is not bound by these recommendations” Read at Scientific American ↗ Left · 1OpenAI said it is publishing in a GitHub repository with protocols for paper revisions and citations while exploring community-hosted alternatives that meet the committee's guidelines. The VergeLC “We’re continuing to explore other community-hosted alternatives for this release which meet the committee’s guidelines.” Read at The Verge ↗ Left · 1Mathematicians are divided: Daniel Litt of the University of Toronto sees no reason for results to be kept secret and expects the release to be good for mathematics, while Terence Tao has criticised frontier labs for the “insane” pace of AI-generated results. Scientific AmericanLC “I see no reason why we should ask the company to keep them secret from us”“has vocally criticized OpenAI and other “frontier” AI labs for the “insane” pace of their AI-generated results” Read at Scientific American ↗
Left · 1OpenAI's spokesperson said many of the new results are not yet understood by its own mathematicians, and the company says it will not slow down because such problems are a test of whether its AI is improving. Scientific AmericanLC “many of the company’s newly released results are not yet understood by its own mathematicians”“it can’t slow down because these math problems are an indispensable test” Read at Scientific American ↗ Left · 1The Verge notes that rivals such as Anthropic are also producing results, including on a Millennium Prize problem, and that the pace has provoked debate over research ethics and how companies credit the human mathematicians whose work they build on. The VergeLC “rival labs like Anthropic that the field is still processing, including results concerning a Millennium Prize problem”“how companies credit the human mathematicians whose work their systems build upon” Read at The Verge ↗
Every sentence links to the reporting it rests on. The pill in front of each says where its sources sit: Left, Centre or Right when one side supplies at least half of them, Mixed when they are evenly split. The number is how many outlets it cites.
Left2 outlets
- Framing
- Both outlets present the release as another step in a fast-moving, contested wave of AI mathematics. Scientific American leads on scale and sceptical reaction, stressing OpenAI's limited transparency. The Verge leads on the numbers and the advisory group's guidelines and research-ethics questions.
- Emphasis
- Transparency, verification, the advisory group's recommendations and divided mathematician opinion (Sutherland, Litt, Tao).
- Leaves out or plays down
- The Verge omits the single-prompt claim, named mathematician reactions and the Lean verification; Scientific American omits the 722-manuscript count and the AGMAI marketing warning.
- Charged language
- “unleashes”“The floodgates are open”“drops another batch”“a field already in shock”
- For example
-
“The floodgates are open.” — Scientific American
“has provoked fierce debate over research practices, ethics” — The Verge
Centre0 outlets
No centre outlet in our sources has covered this story yet.
Right0 outlets
No right outlet in our sources has covered this story yet.
What every side reports
- OpenAI released a large batch of AI-generated mathematics results from an unreleased model on GitHub.
- The results will take mathematicians time to assess.
- An independent mathematicians' advisory group issued guidelines on releasing AI results, and the release is controversial in the field.
Where accounts differ
-
Whether OpenAI's release meets the advisory group's guidelines
- Left
- Scientific American says OpenAI discloses only average compute time and no prompts, despite the recommendations; The Verge quotes OpenAI saying it is exploring alternatives that meet the committee's guidelines.
-
Whether the results were produced by a single agent from a single prompt
- Left
- Scientific American relays OpenAI's claim but quotes MIT's Andrew Sutherland calling it unverified until the model is released; The Verge does not address it.
OpenAI organisation
Says its unreleased model produced the results, mostly from a single prompt to one agent, and is publishing on GitHub. It says it takes the advisory group's guidelines seriously but is not bound by them, plans to release the model responsibly, and will not slow down.
“the team is taking the advisory group’s guidelines seriously and doing its best to comply” — Scientific American
“we’re publishing the results in a GitHub repository, with protocols for paper revisions and citations.” — The Verge
Left2 articles
-
OpenAI unleashes hundreds more math results upon a field already in shock
Critical Breaking-news piece stressing the flood of results, OpenAI's limited disclosure and mathematicians' scepticism.

-
OpenAI drops another batch of mathematical breakthroughs
Neutral Concise report on the release's scale, set against the advisory group's guidelines and the ethics debate.

Centre0 articles
No coverage yet.
Right0 articles
No coverage yet.
- 6 Oct 23:19 First Scientific AmericanLC OpenAI unleashes hundreds more math results upon a field already in shock
- 7 Oct 00:26 +1h 8m The VergeLC OpenAI drops another batch of mathematical breakthroughs
Times are when each article was published, or when we first saw it if the outlet gave no time.