OpenAI's Astra Model Solves Ten Major Open Math Problems
TECH

OpenAI's Astra Model Solves Ten Major Open Math Problems

27+
Signals

Strategic Overview

  • 01.
    OpenAI announced that an internal, unreleased version of its next major model family, Astra, produced new results on ten open problems in mathematics and theoretical computer science, each unsolved for at least a decade.
  • 02.
    OpenAI published a 249-page technical manuscript and a 62-page account of how the reasoning came together, alongside machine-checkable Lean 4 certificates for all ten results, on August 1, 2026.
  • 03.
    The headline result is the first explicit construction of a non-sofic group, a question open since Mikhail Gromov introduced the concept of soficity in 1999; Astra also reportedly disproved Connes's rigidity conjecture.
  • 04.
    OpenAI says the total token cost to produce all ten solutions was roughly $2,000 at its internal 'Sol' API rates.

Deep Analysis

The Verification Trick: Why Machines Trusted Astra's Proofs

OpenAI didn't ask anyone to take Astra's word for it. Every one of the ten proofs was formalized in Lean 4, producing a machine-checkable certificate whose correctness a compiler can verify without trusting the model's own natural-language reasoning at all [1]. That's the real news inside the announcement: math and formal logic are among the only domains where an AI system's output can be graded by software instead of a committee of experts, which is exactly why OpenAI chose it to preview an unreleased model family - Astra - without shipping a product, alongside a 249-page technical manuscript and a 62-page account of how the arguments came together, published August 1, 2026 [2].

The headline result is the first explicit construction of a non-sofic group, closing a question the mathematician Mikhail Gromov opened when he introduced the concept of soficity in 1999 - it had sat unresolved for 27 years [3]. Alongside it, Astra reportedly disproved Connes's rigidity conjecture by building infinitely many non-isomorphic groups with property (T) that share the same von Neumann algebra, and produced a hardness result on the closest vector problem that bears directly on lattice-based cryptography, the mathematical foundation most post-quantum encryption schemes rely on [4]. It also delivered the first improvement to the general upper bound on high-dimensional sphere-packing density since 1978, matching a threshold set decades ago by Cohn and Elkies [5]. These aren't benchmark scores - they're closed, named problems a working mathematician would recognize on sight.

The $2,000 Number Everyone's Fighting Over

OpenAI says the total token cost to generate all ten solutions was roughly $2,000 at its internal 'Sol' API rates [6]. It's a striking number, and it's also the most contested part of the announcement. Commentators pointed out that a token-cost figure only counts the run that succeeded - not the cost of every failed attempt along the way, and not the human time spent choosing which problems to point Astra at or curating the resulting certificates - which makes it closer to a 'cost of publication' than a genuine 'cost of discovery' [7].

The uniqueness of the achievement was also challenged directly. Levent Alpöge, a mathematician at rival lab Anthropic, said a public Claude model reproduced roughly half of Astra's results within 24 hours using nothing more than a generic prompt, no internet access, and full autonomy [7]. If that holds up, it reframes the story: not 'a closed model achieved something no other system could,' but 'a model achieved something, and a competitor's publicly available model got most of the way there almost immediately once it knew where to look.' The announcement also arrives a year after both OpenAI and Google DeepMind claimed gold-medal-level performance at the 2025 International Mathematical Olympiad [8]- Astra reads less like a bolt from nowhere and more like the next beat in a running competitive-signaling contest between frontier labs, timed to preview a model family without OpenAI having to ship or price it yet.

Real Advances or Just Counterexamples? The Substance Debate

Outside validation for at least part of the work is real. Fields Medalist Tim Gowers reviewed one of the proofs and said he would have recommended it for publication in a top mathematics journal without hesitation [9]. Thomas Bloom, who maintains the catalogue of open problems left by Paul Erdős, called the results 'big news' and rated them more significant than an earlier, narrower OpenAI math result - notably, three of the ten problems Astra solved come directly from Erdős's list [9].

But the harshest critique challenges what kind of achievement this actually is. AI researcher Gary Marcus called the announcement 'amazing but vastly oversold,' noting that neither the 249-page manuscript nor the reasoning walkthrough disclose any information about how the results were produced, and warning against what he calls a fallacy of composition - assuming narrow, verifiable success in a domain like formal mathematics says anything about broader general capability [7]. That skepticism echoes a technical split visible in the math community's own reaction: much of the discussion centered on whether Astra mostly found counterexamples that disprove existing conjectures rather than building genuinely new conceptual frameworks, with some readers also criticizing the reasoning-walkthrough documents themselves as hard to follow. Both critiques point at the same gap: verification (does the proof compile) is not the same question as understanding (did anyone - human or machine - learn something new about why it's true).

Mathematicians Are Not All Cheering

Not every reaction was awe. Mathematicians reportedly issued what's being called the 'Leiden Declaration,' endorsed by the International Mathematical Union, warning that AI companies are using published mathematical research without consent and bypassing peer review by announcing results through press release rather than journal submission [3]. That's a structural complaint, not a technical one: the objection isn't that the proofs are wrong, it's that a field's public output is being used as training and evaluation material without the field's consent, and that being 'agent-reviewed' is being allowed to stand in for peer review in public messaging.

There's also a more personal undercurrent running through the reaction - unease, not just professional critique. Software engineer Fernando Borretti captured it starkly: 'We will live in a demon-haunted world, full of marvelous devices whose operation we will not understand' [4]. Layered under the institutional objection is a quieter anxiety about what unequal access to a system like Astra does to who gets to make the next big discovery - and whether being first to a hard problem still means as much once the model doing the finding isn't available to everyone.

Historical Context

1999
Introduced the concept of soficity in group theory, opening the question of whether non-sofic groups exist - a question Astra's construction resolved 27 years later.
1978
The general upper bound on high-dimensional sphere-packing density had not been improved since this year until Astra's new result.
2025
Both companies' AI models achieved gold-medal-standard performance at the International Mathematical Olympiad, setting up context for the Astra math announcement a year later.
2026-08-01
Published the 249-page manuscript, 62-page reasoning walkthrough, and Lean 4 certificates disclosing Astra's ten math results.

Power Map

Key Players
Subject

OpenAI's Astra Model Solves Ten Major Open Math Problems

OP

OpenAI

Developer of Astra; published the results, a 249-page manuscript, a 62-page reasoning walkthrough, and Lean certificates under Apache 2.0 as an implicit preview of its next major model family without a full product release.

AN

Anthropic (via mathematician Levent Alpöge)

Rival lab; Alpöge said a public Claude model reproduced roughly half of Astra's results within 24 hours using a generic prompt, no internet access, and full autonomy - directly challenging how singular OpenAI's achievement is.

TI

Tim Gowers (Fields Medalist)

Independent reviewer whose endorsement lends outside credibility; said he would have recommended at least one proof for publication in a top math journal without hesitation.

TH

Thomas Bloom (maintainer, Erdős Problems database)

Assessed the significance of the Erdős-catalogue results, calling them 'big news' and rating them more significant than an earlier OpenAI math result.

GA

Gary Marcus (AI critic)

Publicly critiqued the announcement as 'amazing but vastly oversold,' arguing OpenAI disclosed no methodology and warning that success in narrow, verifiable math domains doesn't generalize to broader capability claims.

IN

International Mathematical Union / signatories of the 'Leiden Declaration'

Mathematics community body cited as warning that AI companies are using published mathematical research without consent and announcing results via press release rather than peer review.

Fact Check

9 cited
  1. [1] OpenAI Announces Its Next Major Model Astra by Dropping Ten Previously Unsolved Math Solutions
  2. [2] OpenAI's Astra Solves Ten Decade-Old Math Problems With Machine-Checkable Lean Proofs
  3. [3] OpenAI's Astra Model Produces Ten Math Proofs, Including Non-Sofic Groups
  4. [4] OpenAI's Astra Solves 10 Long-Open Math Problems, Publishes Proofs
  5. [5] OpenAI's Astra Model Solves 10 Major Mathematic Problems
  6. [6] OpenAI's Unreleased Model Astra Solves Ten Open Math Problems
  7. [7] OpenAI's Amazing But Vastly Oversold
  8. [8] OpenAI Claims IMO Gold Medal
  9. [9] OpenAI's Astra Solved 10 Decades-Old Math Problems For Just $2,000

Source Articles

Top 5

THE SIGNAL.

Analysts

Argues the announcement is impressive but oversold: 'Neither of those nor the 249-page math article that went with them give any information about how this was accomplished,' and warns against assuming narrow math prowess implies broader general intelligence.

Gary Marcus
AI researcher and critic

Endorsed the mathematical quality of at least one proof as publication-worthy in a top journal without hesitation.

Tim Gowers
Fields Medalist

Rated the August results as more significant than OpenAI's earlier unit-distance math result, calling them 'big news.'

Thomas Bloom
Maintainer, Erdős Problems database

Reacted with unease about a future shaped by AI systems whose internal operation is not understood by humans: 'We will live in a demon-haunted world, full of marvelous devices whose operation we will not understand.'

Fernando Borretti
Software engineer

Suggested that AI's edge is concentrated in domains with cheap, automatic verification, implying the achievement doesn't generalize to unverifiable domains: 'If you're in a verifiable domain, pivot to an unverifiable one.'

Jon Stokes
Commentator
The Crowd

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol API rates.

@@OpenAI13490

Terence Tao got voluntold into running a workshop because he sounded too excited in a meeting. Four years later, OpenAI published ten proofs to the standard that room was built around. ... Saturday, OpenAI published ten results from an unreleased model called Astra. A construction proving non-sofic groups exist, open since 1999. The first improvement to the general sphere packing bound since 1978. A disproof of Connes's rigidity conjecture. Three Erdos problems. Hardness for the closest vector problem, the lattice question post-quantum cryptography sits on. Every one had gone at least a decade without progress. ... Every result shipped with a Lean certificate on GitHub. Not a claim. A file that compiles or doesn't, on any laptop, for anyone. ... OpenAI cited the [Leiden] declaration on Saturday and declined to claim human authorship for any of the ten proofs. In May the same lab disproved the Erdos unit distance conjecture. ... The cost of a new theorem just fell to a line on an invoice. The cost of noticing did not move.

@@r0ck3t2328

SITUATION EXPLAINED: OpenAI's unreleased Astra model solved ten open math problems for under $2,000 in compute. ... The headline result is the first explicit construction of a non-sofic group, a question open since Gromov introduced soficity in 1999. Astra also disproved Connes's rigidity conjecture and proved Ehrhart's volume conjecture. The release is a 249-page manuscript collection, a separate reasoning walkthrough document, and Lean 4 certificates on GitHub. The Lean repo reports zero unproven steps, but OpenAI labels the review status "agent-reviewed," formal peer review hasn't happened.

@@MTSlive39

OpenAI: Ten advances in mathematics and theoretical computer science

@u/Salt_Attorney899
Broadcast
OpenAI's Astra JUST solved math...

OpenAI's Astra JUST solved math...

GPT 6 Astra Could Be OpenAI's Biggest Model Yet!

GPT 6 Astra Could Be OpenAI's Biggest Model Yet!

OpenAI's Hidden Model Astra Just Broke Mathematics

OpenAI's Hidden Model Astra Just Broke Mathematics