OpenAI's Unreleased Astra Model Solves Ten Decades-Old Math Problems, Publishes Machine-Checked Proofs
An internal version of Astra, the model family OpenAI has said will follow GPT-5.6, produced Lean 4-verified solutions to ten long-standing open problems in mathematics and theoretical computer science, published alongside a 249-page manuscript.
OpenAI said on August 1 that an internal, unreleased version of Astra — the model family it has previously described as coming after the GPT-5.6 line — generated solutions to ten open problems in mathematics and theoretical computer science, several of which had stood for decades. The company published every result as a machine-checkable Lean 4 certificate on GitHub under an Apache 2.0 license, alongside a 249-page technical manuscript and a separate account of how the model arrived at each argument.
What the model solved
The ten results span high-dimensional sphere packing, binary and spherical coding theory, arithmetic circuit complexity, group theory, operator algebras, quantum complexity, lattice-based cryptography and extremal combinatorics. The headline result is an explicit construction of a non-sofic group, resolving a question left open since Mikhail Gromov introduced the concept of soficity in 1999. Other results include an improved general bound on high-dimensional sphere-packing density — the first such improvement since 1978 — and disproofs or partial resolutions of several problems from Paul Erdős's catalogue of combinatorics questions, including Erdős problem 183 on multicolored Ramsey numbers. OpenAI said the full set of results was produced using roughly $2,000 of API compute.
Why the Lean verification matters
What distinguishes the announcement from earlier claims of AI-assisted mathematical progress is that each proof compiles in Lean 4, a formal proof assistant whose kernel returns a strict pass/fail verdict rather than a plausibility judgment — OpenAI reported a "sorry" count of zero across all ten certificates, meaning no step was left unproven. That removes the need to trust the model's own explanation of its reasoning, since outside mathematicians can independently verify the certificates compile. Coverage of the release noted that mathematicians reviewing the results, including at least one Fields Medalist, described some of the proofs as strong enough to submit to a top journal.
Why it matters
OpenAI has not set a release date for Astra or said whether it will ship as GPT-6 or as a variant within the existing GPT-5 line, but has described the family as designed to let multiple agents work on a single hard problem for extended stretches. Publishing genuine, independently verifiable mathematical advances — rather than benchmark scores — gives outside researchers a harder data point to evaluate frontier progress by, and raises the bar other labs will be measured against as they make their own claims about AI-assisted research.
Sources
- Ten advances in mathematics and theoretical computer science — OpenAI
- OpenAI's Astra Solves Ten Decade-Old Math Problems With Machine-Checkable Lean Proofs — Tech Times
- OpenAI's Astra solves 10 long-open math problems and publishes the proofs — SiliconANGLE
- OpenAI announces its 'next major model' Astra by dropping ten previously unsolved math solutions — The Decoder
AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.
More ai news
EU Set to Propose Barring Under-15s From AI Chatbots and Social Media in 'Kids Act'
The European Commission is preparing to unveil an EU Kids Act that would bar unsupervised access to AI chatbots, social media, video platforms and online games for under-15s, with tiered rules and mandatory age verification for 13-14 year-olds.
Amodei's 'Pace the Frontier' Plan Draws Same-Day Backing From OpenAI, DeepMind and xAI
Anthropic CEO Dario Amodei published an essay arguing frontier AI labs should deliberately slow capability gains, and within hours Sam Altman, Demis Hassabis and Elon Musk publicly endorsed the idea, with Microsoft's Satya Nadella following a day later.
Positron Raises $875M to Build an HBM-Free AI Inference Chip
Chip startup Positron closed an $875 million Series C at a $5 billion post-money valuation to fund its Asimov inference accelerator, which pairs its compute architecture with up to 2,304GB of commodity LPDDR5X memory instead of scarce high-bandwidth memory.
DeepSeek Releases V4.1 Flash, Cuts API Prices and Sets End Date for V4 Pro
DeepSeek officially released V4.1 Flash, a cheaper and faster multimodal successor to V4 Pro with a 1-million-token context window, and said it will reroute all V4 Pro API traffic to the new model from September 14.