OpenAI publishes a batch of new maths results from an internal model, with Lean proofs
OpenAI released a set of new mathematical results produced by an unnamed internal frontier model, with Lean formalisations of many proofs, summaries of the model's reasoning and statistics on the problems it attempted, all on GitHub. It said the average result used compute roughly equal to three hours of ChatGPT Pro thinking, and that it is working with the Institute for Advanced Study's advisory group on maths and AI on best practice. The post does not describe independent verification of the results.
Why it matters
AI labs are now releasing research results in bulk with machine-checkable proofs, which shifts the debate from whether AI can do new maths to how credit and checking should work.
Line of Thought
Follow this story
Pick any item to keep going. Your path builds up above as a line you can share.
Directly linked
Connections our researchers recorded
- DevelopmentOpenAI claims Navier-Stokes blow-up proof; mathematicians dispute credit8 Sep 2026 · Research · USFollows from: Comes after the disputed Navier-Stokes claim, with more transparency about method
- DevelopmentClaude completes first end-to-end Lean formal proof of Fermat's Last Theorem4 Sep 2026 · Research · US, GBContrasts with: Both labs now publish formal Lean proofs as evidence
What led here
Earlier developments on the same thread
- DevelopmentOpenAI ties large reasoning-distillation campaign to people linked to Moonshot AI30 Sep 2026 · Incident · US, CN
- DevelopmentGPT-6 Astra jumps to 62.7% on ARC-AGI-3, six months after models scored 0.5%3 Sep 2026 · Report · US
- DevelopmentOpenAI model disproves Erdős's 1946 unit-distance conjecture20 May 2026 · Research · US
- DevelopmentGemini Deep Think earns officially graded gold-medal score at IMO 202521 Jul 2025 · Research · GB, US
Same story elsewhere
What other countries and bodies did on this
- DevelopmentAI systems from Huawei and Xiaohongshu reported to score 42/42 at IMO 202623 Jul 2026 · Research · CN, INTL
- DevelopmentIndia hosts AI Impact Summit in New Delhi, first in the series in the Global South16 Feb 2026 · Statement · IN, INTL
- DevelopmentEU publishes General-Purpose AI Code of Practice ahead of AI Act model duties10 Jul 2025 · Rule change · EU
- DevelopmentDeepSeek releases R1 reasoning model with open weights under MIT licence20 Jan 2025 · Model release · CN
Threads by topic: Mathematics Frontier models