Doctorcrypto About RSS Subscribe
Doctorcrypto
HomeBusiness › OpenAI says 10,000 AI agents solved a $1 million math problem. Now mathematicians are fighting
Business

OpenAI says 10,000 AI agents solved a $1 million math problem. Now mathematicians are fighting

By Diego Whitfield · · 2 min read

OpenAI is at the center of a fierce academic dispute after claiming that a coordinated swarm of roughly 10,000 AI agents produced a proposed solution to one of mathematics' most famous unsolved challenges — a Millennium Prize Problem carrying a $1 million reward. While the company frames the result as a landmark for machine reasoning, mathematicians are questioning just how much of the work the AI actually did on its own.

The Claimed Breakthrough

According to OpenAI, an internal system said to be more capable than its publicly referenced GPT-6 Astra model was tasked with attacking one of the seven Millennium Prize Problems, a set of notoriously difficult questions catalogued by the Clay Mathematics Institute. Each carries a $1 million bounty for a verified solution, and only one has been resolved in more than two decades.

The approach reportedly relied on massive parallelism, with thousands of AI agents working simultaneously to explore different avenues of the proof before converging on a proposed answer. OpenAI has positioned the outcome as evidence that large-scale agentic collaboration can tackle problems long considered beyond the reach of automated systems.

A machine may have written the proof, but whether it truly discovered it is now the real question.

Mathematicians Push Back

Skepticism has spread quickly through the mathematics community, where the standard for a Millennium Prize solution is exacting. Critics argue that a proposed proof means little until it survives rigorous, independent peer review — a process that can take years and has sunk many confident claims in the past.

The sharper controversy concerns originality. Some researchers question how independently the AI arrived at its result, raising the possibility that the system leaned heavily on existing human work, published partial results, or guidance embedded in its training data rather than reasoning its way to a genuinely novel breakthrough.

Key points fueling the debate include:

  • Whether the proposed proof holds up under formal verification
  • How much the AI drew on prior human mathematical work
  • What "solved" should mean when thousands of agents operate in parallel

For now, the claim remains unverified by the broader field, and no prize has been awarded. The episode underscores a growing tension in AI research: as models produce increasingly ambitious outputs, the burden of proving those results are real, correct, and truly original is only becoming heavier.

Was this useful?👍 Yes👎 No