OpenAI's Astra Model Solves 10 Open Math Problems
TECH

OpenAI's Astra Model Solves 10 Open Math Problems

42+
Signals

Strategic Overview

  • 01.
    An internal, unreleased version of OpenAI's next model family, code-named Astra, solved ten previously open problems in mathematics and theoretical computer science, headlined by the first-ever explicit construction of a non-sofic group - a question open since 1999.
  • 02.
    OpenAI backed the claims with a 249-page manuscript and machine-checkable Lean 4 proof certificates for all ten results, published on GitHub, at an estimated compute cost of about $2,000.
  • 03.
    Reception is split: independent mathematician Thomas Bloom called the results significant while pushing back on the idea that AI replaces mathematicians, whereas critics like Tasmin Chu argue the framing erases the human research the proofs build on and undercuts peer review, echoing the mathematics community's Leiden Declaration.
  • 04.
    The announcement lands as Sam Altman demos Astra to U.S. senators in Washington, with the model expected to be among the first submitted under a new federal framework giving the government up to 30 days of pre-release access to frontier models.

Deep Analysis

What Astra Actually Proved - And What It Didn't

The headline result is the first explicit construction of a non-sofic group, closing a question that has sat open since Mikhail Gromov introduced the concept of soficity in 1999 - 27 years of failed attempts by human mathematicians [1]. Beyond that flagship result, the same internal Astra model disproved Connes's rigidity conjecture on von Neumann algebras, proved Ehrhart's volume conjecture, resolved three Erdos problems including one on multicoloured Ramsey numbers, and delivered the first improvement to the general upper bound on high-dimensional sphere-packing density since 1978 - a 48-year-old ceiling [1]. It also produced a parallel repetition theorem for two-player quantum games and new lower bounds on the circuit complexity of computing the permanent [2]. What gets lost in the 'ten problems solved' framing is that these are not ten equivalent achievements - some are full proofs of standing conjectures, others are disproofs via explicit counterexample (a problem type that plays to an LLM's strengths in construction and verification), and others are incremental improvements to existing bounds rather than final resolutions.

How The Online Math Community Actually Reacted

On X, the results weren't just an OpenAI press release - they were corroborated firsthand by the company's own researchers, with Sebastien Bubeck, Noam Brown, and Lijie Chen each confirming technical details of the results in public posts, which lends first-party credibility beyond the official announcement even if it isn't independent verification. Reaction outside the company was more mixed. On Reddit, a self-identified Ramsey theory PhD student singled out the multicoloured Ramsey number result as particularly strong, calling it a contender for top combinatorics paper of the year - notable praise from inside the relevant subfield. Other mathematicians on Reddit were harsher about the presentation: several described the prose in Astra's reasoning walkthroughs as incomprehensible or resembling crank writing, pointing to undefined jargon that made the arguments hard to follow independently of the Lean certificates underneath them. And the $2,000 cost figure did not survive online scrutiny unchallenged - multiple Reddit commenters, not just written critics, independently flagged that the number almost certainly counts only successful runs and excludes whatever compute went into problems Astra attempted and failed on. Layered on top of that is a framing critique circulating in online technical commentary: 'solved' is doing a lot of work in the announcement, since the ten results are not equivalent achievements - some are full proofs, some are disproofs by explicit counterexample, and some are incremental improvements to existing bounds rather than final resolutions.

The $2,000 Number Doesn't Mean What People Think

The most viral detail in the announcement is the price tag: OpenAI says the ten solutions together cost roughly $2,000 in API compute [1][4]. That figure has become the story's flashpoint, and for good reason - it invites a direct comparison to years of unpaid, uncredited human labor. But the number is also less clean than it sounds. OpenAI has not disclosed how many problems Astra attempted and failed on before landing these ten, only that it tried and failed on other major problems along the way, according to OpenAI researcher Noam Brown, who framed the result as a major step for scientific reasoning while conceding the failures [2]. A $2,000 figure that counts only successful runs, with no accounting for compute spent on dead ends, is a very different claim than '$2,000 to solve ten problems from scratch.' That framing gets sharper still when you look at what the proofs build on: critics have pointed out that the non-sofic group construction leans heavily on prior published work by mathematicians Andreas Thom and Gabor Kun, whose foundational contributions risk being erased from coverage that credits the achievement to Astra alone [4]. Mathematician Tasmin Chu's reaction captures the emotional core of the backlash: a price tag that low attached to work she says represents roughly 15 years of cumulative community effort [4].

Lean-Checked Is Not The Same As Peer-Reviewed

OpenAI's response to skepticism was procedural: alongside the 249-page manuscript, it published machine-checkable Lean 4 certificates for all ten results, built on Lean 4.32.0 and mathlib [1][3]. That is a real upgrade in verifiability compared to OpenAI's earlier unit-distance conjecture disproof in May, which external mathematicians checked and endorsed [8]. Yet a Lean certificate only proves that a formal statement follows from formal axioms - it does not establish that the formal statement is the theorem mathematicians actually care about, or that the surrounding argument is intelligible to the field it claims to advance [3]. That gap matters here specifically because the May result was validated by named external mathematicians including Fields Medalist Tim Gowers, who said he would recommend it for the Annals of Mathematics 'without hesitation' [1]- a level of named, independent endorsement that Astra's August results have not yet received in the same public way.

Why Astra Is Surfacing In Washington, Not Just In Math Journals

The timing is not incidental. Sam Altman has been personally demoing Astra to U.S. senators and administration officials in Washington around the same window as the math announcement, using the results as the headline proving ground for the model's reasoning capability [5][7]. That lobbying push lines up with a June 2 executive order directing federal agencies to build a voluntary process giving the government up to 30 days of access to frontier models before public release [6]. Astra is expected to be among the first models submitted under that framework, which makes the math results less a pure research announcement and more an opening argument for how capable - and how safe to release quickly - OpenAI wants regulators to believe its next flagship model is.

Historical Context

1999
Introduced the concept of soficity in group theory, opening the question of whether non-sofic groups exist - a question that stood unresolved for 27 years until Astra's construction.
1978
The prior general upper bound on high-dimensional sphere-packing density dated to 1978; Astra produced the first improvement to that bound in 48 years.
2026-05-20
An earlier internal OpenAI reasoning model disproved Paul Erdos's 1946 unit distance conjecture, checked and endorsed by external mathematicians, setting the precedent for AI-generated math results and the controversy that followed.
2026-06-02
Signed an executive order directing federal agencies to build a voluntary process giving government up to 30 days of access to frontier AI models before public release, a framework Astra is expected to go through first.
2026-06
Issued the Leiden Declaration warning that AI companies were using published mathematical research without consent, bypassing peer review, and threatening the integrity of proof and attribution norms.
2026-08-01
Published the Astra ten-proofs announcement, manuscript, and Lean certificates, while Sam Altman previewed Astra to U.S. senators in Washington around the same time.

Power Map

Key Players
Subject

OpenAI's Astra Model Solves 10 Open Math Problems

OP

OpenAI

Developer of Astra; published the ten results, the manuscript, and the Lean certificates, using the achievement to position Astra as its next flagship model ahead of a government review process.

SA

Sam Altman

OpenAI CEO; personally demoed Astra to U.S. senators and administration officials in Washington around the announcement, tying the math results to Astra's public debut and the federal review timeline.

SE

Sebastien Bubeck

OpenAI Head of Mathematics Research; confirmed and characterized the results publicly, emphasizing that each proof ships with a Lean certificate and chain-of-thought walkthrough.

TH

Thomas Bloom

Royal Society Research Fellow and curator of erdosproblems.com; independent mathematician whose public assessment lent outside credibility to the results while cautioning against overclaiming AI autonomy.

TA

Tasmin Chu

Mathematician and Substack author; published the most detailed critical reaction, arguing the field should refuse AI-company collaborations and resist using LLMs to prove new theorems absent stronger norms.

TR

Trump administration / federal frontier-model review framework

Building a voluntary 30-day pre-release review process for frontier AI models via a June 2 executive order; Astra is expected to be among the first models submitted under it, shaping the political context of the release.

Fact Check

8 cited
  1. [1] OpenAI's Astra Model Solves Ten Math Proofs, Including First Non-Sofic Group Construction
  2. [2] OpenAI Announces Its Next Major Model, Astra, By Dropping Ten Previously Unsolved Math Solutions
  3. [3] OpenAI's Astra: 10 Math Problems Solved With Lean Proofs
  4. [4] Mathematicians Need To Act
  5. [5] OpenAI Says It Has Solved 10 Open Math Problems Using Astra, Its New Model
  6. [6] OpenAI Senators: Astra And The 30-Day Review Framework
  7. [7] Altman Demos Astra To US Senators Amid Frontier Model Review Push
  8. [8] OpenAI Model Disproves Erdos Unit Distance Conjecture

Source Articles

Top 5

THE SIGNAL.

Analysts

Called the results more significant in some respects than OpenAI's earlier unit-distance conjecture disproof, while pushing back on framing that suggests AI replaces mathematicians, noting the system builds on over a century of accumulated mathematical theory.

Thomas Bloom
Royal Society University Research Fellow, University of Manchester / erdosproblems.com

Praised the technical and aesthetic quality of the results and stressed their verifiability through Lean certificates and chain-of-thought walkthroughs published alongside them.

Sebastien Bubeck
OpenAI Head of Mathematics Research

Framed the result as a major step for scientific reasoning capability, while acknowledging Astra failed on other attempted problems and suggesting further test-time compute could extend the gains.

Noam Brown
OpenAI researcher

Argued the $2,000 cost figure trivializes roughly 15 years of cumulative community effort, that credit to human mathematicians whose work underpins the proofs risks being erased in coverage, and that the profession could see early-career researchers leave the field over feeling devalued.

Tasmin Chu
Mathematician / Substack author
The Crowd

yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra proofs, complete with lean certificates and CoT walkthroughs for each of them. The results are wide-ranging, from von Neumann algebras (disproof of Connes' Rigidity Conjecture) to better bounds for high dimensional sphere packing, for circuit complexity, for monochromatic triangles in multicolored graphs, and more. More thoughts here: openai.com/index/ten-adva...

@@SebastienBubeck6236

An internal version of Astra, @OpenAI's next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science. We believe it will be a major step for scientific reasoning. openai.com/index/ten-adva

@@polynoamial15975

10 proofs from our next major model Astra on long-standing open problems in mathematics and theoretical computer science (also including new circuit lower bounds for computing the permanent!) GPT-5.6 has already enabled so much exciting work in math and science. Can't wait to see what comes next!

@@wjmzbmr11888

OpenAI: Ten advances in mathematics and theoretical computer science

@u/Salt_Attorney822
Broadcast
OpenAI's Astra JUST solved math...

OpenAI's Astra JUST solved math...

OpenAI's Hidden Model Astra Just Broke Mathematics

OpenAI's Hidden Model Astra Just Broke Mathematics

OpenAI Astra and the Truth Behind 10 Solved AI Mathematical Proofs

OpenAI Astra and the Truth Behind 10 Solved AI Mathematical Proofs

OpenAI's Astra Model Solves 10 Open Math Problems — AI News | Agentic Brew