OpenAI claims to have solved maths problem that stumped humans for decades
Company behind ChatGPT says 10,000 of its AI systems cracked the Navier-Stokes problem in 88 hours OpenAI claims to have solved a major mathematics problem that has stumped humans for nearly a century after spending millions of dollars on the artificial intelligence-led endeavour. The company behind ChatGPT said it had cracked the Navier-Stokes problem, one of seven Millennium Prize Problems published by the Clay Mathematics Institute to highlight some of the biggest unsolved puzzles in the field. Continue reading...
Every model that read this
| Model | Provider | Stage | Score | Conf. | Latency | Prompt | When |
|---|---|---|---|---|---|---|---|
| Llama 3.3 70B | Meta | analysis | +20 | 60% | 2591ms | v1.0.0 / m1.0.1 | 2026-09-10 09:13 |
| GPT-4.1 mini | OpenAI | consensus | +35 | 80% | 3813ms | v1.0.0 / m1.0.1 | 2026-09-10 18:19 |
| Claude Sonnet 5 | Anthropic | consensus | +28 | 35% | 10293ms | v1.0.0 / m1.0.1 | 2026-09-10 18:19 |
Solving a major maths problem could lead to breakthroughs.
Solving the Navier-Stokes problem is a major mathematical breakthrough that could accelerate scientific and engineering progress, potentially benefiting society. However, the direct societal implications are indirect and dependent on subsequent applications. The claim is recent and well-evidenced but the broad impact remains to be seen.
The claim, if verified, would be a notable capability milestone with positive implications for scientific progress, but it rests solely on OpenAI's own announcement without independent verification or peer review. The speculative nature and self-interested framing warrant a cautious, moderate score.
Evidence extracted
- OpenAI solved the Navier-Stokes problem
SOURCE The Guardian: Artificial Intelligence (tier 1)
↓
DOCUMENT 0eb2634a-6e9f-401a-976c-879dcd56e6fc
https://theguardian.com/science/2026/sep/08/openai-claims-to-have-solved-maths-problem-that-stumped-humans-for-decades
↓
EVIDENCE 1 extracted excerpt
↓
MODEL RUN 3 runs, methodology 1.0.1
↓
SCORE +28 (Favourable)
↓
CONFIDENCE 58%