Radio
Now Playing
Quickyla Radio — Click to play
Open →
3 min left
Back to News

OpenAI’s math solutions aren’t meeting the field’s standards yet

When OpenAI released hundreds of claimed solutions to some of the world’s hardest math problems this week, the frontier lab said that it had consulted an advisory group of elite mathematicians to avo…

OpenAI’s math solutions aren’t meeting the field’s standards yet
TechCrunch — 8 October 2026
Text:
8 0 0

When OpenAI released hundreds of claimed solutions to some of the world’s hardest math problems this week, the frontier lab said that it had consulted an advisory group of elite mathematicians to avoid the controversy that came with the last time one of its models solved a long-standing problem in the field.

But OpenAI fell short of those standards, particularly where the mathematicians emphasized the need for human understanding of a mathematical result. That’s especially concerning after a new paper highlighted gaps between the natural language and formally expressed solution to a million-dollar problem ostensibly solved by OpenAI’s models.

The Advisory Group on Mathematics and Artificial Intelligence (AGMAI), hosted by Princeton University’s Institute for Advanced Studies, is made up of nine prominent researchers at institutions around the world.

The organization released guidelines for frontier labs solving math problems at the end of September. In a statement on the latest set of proofs, the AGMAI said that “it is ultimately up to the mathematical community to assess the extent to which our recommendations were followed successfully.”

However, the organization’s first request was “to stop testing advanced mathematical problems on proprietary models.” OpenAI’s release explicitly says that it is evaluating its proprietary models using open research problems in mathematics.

The advisory group did not respond when asked by TechCrunch for a more thorough evaluation of OpenAI’s latest proof release. The lab clearly followed some of its principles, including releasing results as soon as possible and including information about how the models reached their conclusions. But not for all of them: Just 10 of the 719 manuscripts included releases of the model’s chain of thought.

For papers that people don’t understand, the mathematicians suggested the proofs should be formalized — but just 42% of the proofs released by OpenAI had not undergone this process.

Ultimately, it’s still not clear that OpenAI is taking “responsibility for ensuring that human understanding will follow” when releasing its proofs, in accordance to the AGMAI principles. AGMAI suggested that OpenAI should help fund the work of human mathematicians who will be required to make the lab’s solutions meaningful in any real way.

Read Full Story at TechCrunch →
Advertisement
React:
Sources
Sponsored

More to Read

CNN poll shows 75% of Americans fear AI's impact on jobs, p…
💻 Technology
CNN poll shows 75% of Americans fear AI's impact on jobs, privacy.
The Hill · 14 days ago
Google's hidden Android features boost user loyalty against…
💻 Technology
Google's hidden Android features boost user loyalty against Apple competitors
Android Authority · 15 days ago
How to get started with Shortcuts on your MacBook
💻 Technology
How to get started with Shortcuts on your MacBook
Engadget · 10 days ago
PrismML launches tiny LLMs for Qualcomm-powered smart glass…
💻 Technology
PrismML launches tiny LLMs for Qualcomm-powered smart glasses, enhancing offline AI
TechCrunch · 14 days ago
China posts weakest industrial profit growth this year, exp…
📈 Markets & Finance
China posts weakest industrial profit growth this year, expanding 4.2% in August
CNBC Economy · 11 days ago
UNGA81: Why has Africa’s Security Council reform push remai…
🌍 World News
UNGA81: Why has Africa’s Security Council reform push remained unresolved?
Al Jazeera · 13 days ago
Full view