OpenAI AI Agents Claim to Solve Millennium Prize Problem

By bonuz NewsroomPublished September 10, 2026
OpenAI AI Agents Claim to Solve Millennium Prize Problem

OpenAI says a group of 10,000 AI agents produced a proposed solution to one of mathematics' seven Millennium Prize Problems, each worth $1 million (USD). Mathematicians are now disputing how the system reached its answer. This matters because it tests whether AI can make real breakthroughs, not just repeat known work.

What actually happened

According to CoinDesk, OpenAI said an internal model, more powerful than its GPT-6 Astra system, produced a proposed solution to a major math problem. The company used a coordinated group of 10,000 AI agents to work on the task. The target was one of the seven Millennium Prize Problems, each carrying a $1 million (USD) prize. The report does not specify which of the seven problems was targeted, or share the full technical writeup. Mathematicians reviewing the claim are questioning how independently the AI system reached its answer. The dispute centers on whether the agents solved the problem on their own, or relied heavily on existing human proof techniques built into the system.

How we got here

OpenAI has been racing to show its models can handle advanced reasoning, not just language tasks. GPT-6 Astra, mentioned as the baseline this new system beats, marks OpenAI's most recent named model before this internal system. AI labs have increasingly turned to large numbers of coordinated agents, rather than a single model, to tackle complex problems. This claim extends that approach to one of mathematics' seven Millennium Prize Problems, a category widely viewed as among the hardest in the field, according to CoinDesk.

Why this matters for you

For AI builders, the claim raises the bar for what agent based systems must prove before a result counts as verified. For mathematicians, it adds pressure to build faster ways to audit AI generated proofs. For everyday users and holders of AI related assets, the dispute is a reminder that bold capability claims need independent verification before they move markets or reputations. If confirmed, the result could push more research toward agent swarms. If disputed successfully, it could slow trust in single company announcements about breakthrough math results.

The bigger question

If an AI system cannot show its full reasoning path, how should anyone judge whether a mathematical breakthrough is genuine or borrowed? This question extends beyond OpenAI. As AI agents take on harder scientific problems, the field needs a way to separate real independent discovery from work that leans on hidden human input. Who sets that standard, and who checks it?

What to watch

Independent mathematicians are expected to review the proposed solution before it counts as verified. No formal peer review timeline has been announced. Watch for OpenAI to release more technical detail on the internal model and the specific Millennium Prize Problem involved. Further scrutiny from the mathematics community could confirm, revise, or reject the claim in the coming weeks.

Keep reading