Lines of Thought / Topic

Mathematics: the line so far

8 developments, in date order. Built automatically from everything tagged with this topic.

  1. ResearchScienceGB US
    Google DeepMind's AlphaEvolve agent improves algorithms and open maths bounds

    It showed LLM-driven search producing verifiable new results in mathematics and engineering, a template later used against Erdős problems.

    Also on When AI started doing real mathematics

  2. ResearchScienceGB US
    Gemini Deep Think earns officially graded gold-medal score at IMO 2025

    Officially certified olympiad-level proof writing reset expectations for how quickly general models could do rigorous mathematics.

    Also on When AI started doing real mathematics

  3. ResearchScienceGB US
    DeepMind study: most AI 'solutions' to open Erdős problems were already in the literature

    It is a lab's own corrective on AI maths claims: novelty and attribution need verification, not just correctness.

    Also on When AI started doing real mathematics

  4. ResearchScienceUS
    OpenAI model disproves Erdős's 1946 unit-distance conjecture

    OpenAI called it the first time a prominent open problem central to a field of mathematics was solved autonomously by AI, and Gowers called it a milestone, shifting AI from solving obscure problems to famous ones.

    Also on When AI started doing real mathematics

  5. ResearchScienceCN INTL
    AI systems from Huawei and Xiaohongshu reported to score 42/42 at IMO 2026

    Olympiad maths is now saturated as an AI benchmark, and Chinese labs reached the top alongside US ones.

    Also on When AI started doing real mathematics

  6. ResearchScienceUS GB
    Claude completes first end-to-end Lean formal proof of Fermat's Last Theorem

    Large-scale autoformalisation makes machine-checked verification of major mathematics practical, which also matters for trusting AI-generated proofs.

    Also on When AI started doing real mathematics

  7. ResearchScienceUS
    OpenAI claims Navier-Stokes blow-up proof; mathematicians dispute credit

    It is the biggest AI maths claim yet and a test of credit, data use and verification norms when frontier labs do research.

    Also on When AI started doing real mathematics

  8. ResearchScienceUS
    OpenAI publishes a batch of new maths results from an internal model, with Lean proofs

    AI labs are now releasing research results in bulk with machine-checkable proofs, which shifts the debate from whether AI can do new maths to how credit and checking should work.