OpenAI's Unreleased Astra Model Solves Ten Decades-Old Math Problems, Publishes Machine-Checked Proofs
An internal version of Astra, the model family OpenAI has said will follow GPT-5.6, produced Lean 4-verified solutions to ten long-standing open problems in mathematics and theoretical computer science, published alongside a 249-page manuscript.
OpenAI said on August 1 that an internal, unreleased version of Astra — the model family it has previously described as coming after the GPT-5.6 line — generated solutions to ten open problems in mathematics and theoretical computer science, several of which had stood for decades. The company published every result as a machine-checkable Lean 4 certificate on GitHub under an Apache 2.0 license, alongside a 249-page technical manuscript and a separate account of how the model arrived at each argument.
What the model solved
The ten results span high-dimensional sphere packing, binary and spherical coding theory, arithmetic circuit complexity, group theory, operator algebras, quantum complexity, lattice-based cryptography and extremal combinatorics. The headline result is an explicit construction of a non-sofic group, resolving a question left open since Mikhail Gromov introduced the concept of soficity in 1999. Other results include an improved general bound on high-dimensional sphere-packing density — the first such improvement since 1978 — and disproofs or partial resolutions of several problems from Paul Erdős's catalogue of combinatorics questions, including Erdős problem 183 on multicolored Ramsey numbers. OpenAI said the full set of results was produced using roughly $2,000 of API compute.
Why the Lean verification matters
What distinguishes the announcement from earlier claims of AI-assisted mathematical progress is that each proof compiles in Lean 4, a formal proof assistant whose kernel returns a strict pass/fail verdict rather than a plausibility judgment — OpenAI reported a "sorry" count of zero across all ten certificates, meaning no step was left unproven. That removes the need to trust the model's own explanation of its reasoning, since outside mathematicians can independently verify the certificates compile. Coverage of the release noted that mathematicians reviewing the results, including at least one Fields Medalist, described some of the proofs as strong enough to submit to a top journal.
Why it matters
OpenAI has not set a release date for Astra or said whether it will ship as GPT-6 or as a variant within the existing GPT-5 line, but has described the family as designed to let multiple agents work on a single hard problem for extended stretches. Publishing genuine, independently verifiable mathematical advances — rather than benchmark scores — gives outside researchers a harder data point to evaluate frontier progress by, and raises the bar other labs will be measured against as they make their own claims about AI-assisted research.
Sources
- Ten advances in mathematics and theoretical computer science — OpenAI
- OpenAI's Astra Solves Ten Decade-Old Math Problems With Machine-Checkable Lean Proofs — Tech Times
- OpenAI's Astra solves 10 long-open math problems and publishes the proofs — SiliconANGLE
- OpenAI announces its 'next major model' Astra by dropping ten previously unsolved math solutions — The Decoder
AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.
More ai news
Demis Hassabis Steps Down as Google DeepMind CEO, Hands Day-to-Day Control to Koray Kavukcuoglu
Hassabis becomes chairman of Google DeepMind and Alphabet's chief scientist to focus on AGI research, while longtime DeepMind CTO Koray Kavukcuoglu takes over Gemini development and frontier research, reporting directly to Sundar Pichai.
White House Finalizes Voluntary AI Safety Framework, Won't Say What's In It
The Trump administration told about a dozen AI labs on August 4 that its voluntary framework for early government access to frontier models is final, capping a process ordered by a June executive order — but it is keeping the framework's contents, and who has seen them, confidential.
EU Begins Enforcing AI Act Transparency Rules as High-Risk Deadlines Slip to 2027-2028
The European Commission's AI Office started enforcing new EU AI Act transparency obligations on August 2, requiring chatbots to disclose they're AI and deepfakes to be labeled, even as high-risk system rules were pushed back under the Digital Omnibus on AI.
Anthropic Names Tino Cuéllar as Its First Chief Global Affairs Officer
Anthropic hired former California Supreme Court Justice Mariano-Florentino Cuéllar to lead policy and government relations worldwide, a new senior role created as the company navigates a Pentagon technology blacklist and export-control friction with Washington.