Skip to content
Bramble

Technology

OpenAI publishes 722 manuscripts of AI-generated mathematics results from unreleased model

OpenAI released 722 manuscripts covering 372 result families, produced by an unreleased model, in a GitHub repository. Mathematicians are sceptical of its claims about how they were produced, and of its disclosure compared with new advisory group guidelines.

2 outlets · 2L · 0C · 0R First reported Account updated
Image: The Verge
Image: Scientific American

1 / 2

The story, neutrally told

Left · 1OpenAI has published 722 manuscripts, grouped into 372 result families, containing solutions to long-standing mathematics problems produced by an unreleased frontier model. Left · 1Scientific American reported that OpenAI posted the results to a GitHub repository at 6 P.M. EDT on 6 October, and that the company says each result resolves or makes substantial progress on a major open question in mathematics or theoretical computer science. Left · 1Among the claimed results are a solution to the four-dimensional Kakeya conjecture, improvements to important computer algorithms, and progress toward the Riemann hypothesis.

Left · 2The release follows an earlier OpenAI result on the Navier-Stokes problem, which Scientific American says came from a 10,000-agent swarm costing millions of dollars in computing; in September OpenAI said its model had resolved more than 100 long-standing open problems. Left · 2An OpenAI spokesperson told Scientific American that the model produced almost every new result from a single prompt given to a single AI agent, though some may have taken multiple attempts; OpenAI separately says the average result used the equivalent of three hours of ChatGPT Pro thinking. Left · 1MIT mathematician Andrew Sutherland said claims of single-agent results should be treated as unverified until the model is released and results can be replicated.

Left · 1Many results have already been verified in Lean, a proof-checking language, according to Scientific American, but mathematicians will need months to judge whether the proofs contain novel and important ideas. Left · 2The Advisory Group on Mathematics and Artificial Intelligence (AGMAI), a newly formed independent group of mathematicians announced by OpenAI on 21 September, published recommendations in late September urging labs to release results promptly through established academic channels and to disclose the model, prompts and compute costs. Left · 2The group also urged companies to avoid treating releases as marketing vehicles and criticised the use of proprietary internal models that mathematicians cannot access.

Left · 1Scientific American reports OpenAI is disclosing only average compute time and some statistics, with no prompts; the company says it takes the guidelines seriously and is not bound by them, and is working to release the model as quickly and responsibly as possible. Left · 1OpenAI said it is publishing in a GitHub repository with protocols for paper revisions and citations while exploring community-hosted alternatives that meet the committee's guidelines. Left · 1Mathematicians are divided: Daniel Litt of the University of Toronto sees no reason for results to be kept secret and expects the release to be good for mathematics, while Terence Tao has criticised frontier labs for the “insane” pace of AI-generated results.

Left · 1OpenAI's spokesperson said many of the new results are not yet understood by its own mathematicians, and the company says it will not slow down because such problems are a test of whether its AI is improving. Left · 1The Verge notes that rivals such as Anthropic are also producing results, including on a Millennium Prize problem, and that the pace has provoked debate over research ethics and how companies credit the human mathematicians whose work they build on.

Every sentence links to the reporting it rests on. The pill in front of each says where its sources sit: Left, Centre or Right when one side supplies at least half of them, Mixed when they are evenly split. The number is how many outlets it cites.

Left2 outlets

Framing
Both outlets present the release as another step in a fast-moving, contested wave of AI mathematics. Scientific American leads on scale and sceptical reaction, stressing OpenAI's limited transparency. The Verge leads on the numbers and the advisory group's guidelines and research-ethics questions.
Emphasis
Transparency, verification, the advisory group's recommendations and divided mathematician opinion (Sutherland, Litt, Tao).
Leaves out or plays down
The Verge omits the single-prompt claim, named mathematician reactions and the Lean verification; Scientific American omits the 722-manuscript count and the AGMAI marketing warning.
Charged language
“unleashes”“The floodgates are open”“drops another batch”“a field already in shock”
For example
“The floodgates are open.” — Scientific American
“has provoked fierce debate over research practices, ethics” — The Verge

Centre0 outlets

No centre outlet in our sources has covered this story yet.

Right0 outlets

No right outlet in our sources has covered this story yet.