Grok 4.5 Launch: xAI's Opus-Class, Lower-Cost Coding Model
TECH

Grok 4.5 Launch: xAI's Opus-Class, Lower-Cost Coding Model

23+
Signals

Strategic Overview

  • 01.
    SpaceXAI released Grok 4.5 on July 8, 2026 - its first model since the company went public - pricing it at $2 per million input tokens and $6 per million output tokens, versus Claude Opus 4.7's $5/$25. Musk called it 'an Opus-class model, but faster, more token-efficient and lower cost.'
  • 02.
    The model reportedly runs on a newly pre-trained 'V9' base model - a name and scale that one independent technical breakdown estimated but that xAI has not officially confirmed - jointly trained using developer-session data pulled in after xAI's move on Cursor, in a deal reported near a $60 billion valuation.
  • 03.
    Benchmark results are mixed across outlets: Grok 4.5 topped the SWE Marathon benchmark (29.0% vs Claude Opus 4.8's 26.0%) and nearly matched GPT 5.5 on Terminal Bench 2.1 (83.3% vs 83.4%), but trailed Opus 4.8 on DeepSWE 1.1 (53% vs 59%) while using roughly a quarter of the tokens per task on SWE Bench Pro (15,954 vs 67,020).
  • 04.
    Musk also said Grok 4.5 helped disprove a graph-theory conjecture open for roughly 30 years, but as of this writing there is no linked peer-reviewed paper or independent formal verification of the result.

Deep Analysis

The Real Unlock Wasn't Bigger - It Was Cursor's Data

Grok 4.5 isn't a light fine-tune of its predecessor - one independent technical breakdown estimated it runs on a newly pre-trained 'V9' base model, a claim xAI has not officially confirmed. But the more consequential input is data, not size. After xAI's reported move on Cursor, at a valuation near $60 billion [1], real developer-session coding data was folded into training, and Grok 4.5 now ships inside Cursor on all plans rather than as a bolt-on integration [2]. That pipeline shows up less in raw leaderboard position and more in efficiency: on SWE Bench Pro, Grok 4.5 used just 15,954 tokens per task versus Claude Opus 4.8's 67,020 [3], a gap that workable, real-world coding traces - not just parameter count - would plausibly explain. It is a quieter story than 'biggest model wins,' but arguably the more durable one: xAI effectively bought its way to a training-data advantage that competitors without a captive IDE relationship can't easily replicate.

A Price Cut That Could Start Squeezing Anthropic and OpenAI

The headline number is the pricing: $2 per million input tokens and $6 per million output tokens, against Claude Opus 4.7's $5 and $25 [4]. That's not a marginal discount - it's roughly a quarter to a fifth of the cost per token on the output side, in a market where output tokens dominate agentic coding spend. Analysts read it as a deliberate wedge rather than a byproduct of efficiency: Futurum Group's Bradley Shimmin framed SpaceXAI's coding focus as good for the market because it raises competitive pressure that benefits enterprise buyers [5], while Counterpoint's Neil Shah said the pricing 'could pressure enterprise AI pricing by offering a lower-cost option for high-volume coding and agentic AI workloads' [5]. Other coverage frames this explicitly as a possible price war that could rattle both Anthropic and OpenAI's margins on high-volume coding and agentic workloads [6]. The timing sharpens the stakes further: Grok 4.5 landed the same general window as a new OpenAI model release, turning what might have been a routine update into a three-way pricing and capability standoff [7].

Musk's Uncharacteristic Restraint Meets a Split Verdict From Builders

What stands out in Musk's own framing is how little it resembles typical launch hype. Rather than declaring outright superiority, he described Grok 4.5 as 'not quite as good as Fable, but it is very fast, cost-effective and gets the job done' [8]- a tempered, almost workhorse framing that concedes ground on raw capability while betting on speed and cost. That restraint lines up with how the model is actually landing among developers: some report it holding up close to Opus-tier quality for a fraction of the price in real coding sessions, praising low token burn and strong tool-calling, while others describe it losing focus on complex or long-context tasks, over-engineering solutions, or being 'very confident when wrong.' A few accounts of agentic failures, including one describing a deleted production database during use, sit alongside genuine enthusiasm from third-party builders about how quickly the model was adopted into existing workflows. That gap between benchmark framing and hands-on experience is itself the story: Grok 4.5's SWE Marathon lead over Opus 4.8 [9]reads very differently depending on whether you're pricing tokens or debugging a broken build.

A Math Breakthrough Nobody Has Verified Yet

Alongside the coding benchmarks, Musk publicized a more dramatic claim: that Grok 4.5 helped resolve a graph-theory conjecture that had stood open for about 30 years, in one account constructing a counterexample around hypercontractivity of the Poisson semigroup [8], in another disproving the hypothesis in roughly eight minutes during a routine Slack discussion [10]. The two accounts don't fully agree on the mechanics of the claim, and neither has been matched by a linked peer-reviewed paper or independent formal verification as of this writing [8]. That gap matters beyond this one story: it's a live case study in how quickly an unverified AI claim can circulate as settled fact simply because a prominent figure states it plainly. The launch also didn't happen in isolation - it landed within roughly the same month as other frontier releases circulating in the market, adding pressure to publicize a headline-grabbing result quickly rather than wait for scrutiny. Until independent mathematicians formally check the work, the conjecture claim is best read as a marketing data point, not a verified scientific result.

Historical Context

2023-11-04
xAI launched the original Grok in beta for X Premium users, based on the Grok-1 model.
2025-02
Grok 3 was released to X Premium Plus subscribers and xAI platforms.
2025-07-09
Grok 4 launched as xAI's flagship model with stronger reasoning, native tool-calling, and a $300/month SuperGrok Heavy tier.
2026-02
SpaceX reportedly absorbed xAI, rebranding the AI unit as SpaceXAI ahead of a June 2026 IPO.
2026-07-08
SpaceXAI publicly released Grok 4.5, its first model release since going public.

Power Map

Key Players
Subject

Grok 4.5 Launch: xAI's Opus-Class, Lower-Cost Coding Model

XA

xAI / SpaceXAI

Developer and publisher of Grok 4.5; positioned it as an Opus-class, lower-cost coding and agentic model following its public listing.

EL

Elon Musk

xAI/SpaceXAI leader who announced the launch, positioned Grok 4.5 against Claude Opus, and publicized the math-conjecture claim.

CU

Cursor

AI coding startup reportedly acquired by SpaceXAI near a $60 billion valuation; its developer-session data was used to jointly train Grok 4.5, which now ships in Cursor on all plans.

AN

Anthropic (Claude Opus)

The primary competitive benchmark used throughout coverage; Grok 4.5 is repeatedly compared against Claude Opus 4.7/4.8 on both price and performance.

OP

OpenAI

Competing lab that reportedly released a new model around the same period, framing the launch as part of a head-to-head frontier-model window.

Fact Check

10 cited
  1. [1] Grok 4.5 proves xAI's $60 billion Cursor bet was right
  2. [2] What is Grok 4.5? xAI's Cursor Coding Model
  3. [3] Grok 4.5: Elon Musk vs Claude Opus
  4. [4] SpaceXAI releases Grok 4.5, which Elon describes as an Opus-class model
  5. [5] Grok 4.5 Review
  6. [6] SpaceX's Grok 4.5 launches at half the price of rivals
  7. [7] SpaceX launches new Grok model
  8. [8] Grok 4.5 cracks 30-year-old graph theory conjecture, Musk says
  9. [9] Grok 4.5 tops SWE Marathon benchmark
  10. [10] Grok 4.5 helped disprove a mathematical hypothesis unresolved for about 30 years

Source Articles

Top 1

THE SIGNAL.

Analysts

"Said SpaceXAI's coding-first focus benefits the broader market because it raises competitive pressure, which ultimately helps enterprises."

Bradley Shimmin
Analyst, Futurum Group

"Said Grok 4.5's pricing could pressure enterprise AI pricing by offering a lower-cost option for high-volume coding and agentic AI workloads."

Neil Shah
Analyst, Counterpoint
The Crowd

"Grok 4.5, our most capable model yet, is now available across grok.com, X, and the iOS and Android apps."

@@grok3469

"Grok 4.5 is not quite as good as Fable, but it is very fast, cost-effective and gets the job done"

@@elonmusk15565

"Ok Grok 4.5 is insane. People are already building things that shouldn't be possible this fast. 10 wild examples:"

@@minchoi1225

"Elon hails its Grok 4.5 by sharing the SWE benchmark and still can't figure out why the majority of people prefer Claude Mythos/Fable 5"

@u/Current-Guide594447
Broadcast
Grok 4.5 explained in 8min..

Grok 4.5 explained in 8min..

New Grok 4.5 Finally Hits Claude Hard

New Grok 4.5 Finally Hits Claude Hard

I Tested NEW Grok 4.5 for Coding. Wow. Just Wow.

I Tested NEW Grok 4.5 for Coding. Wow. Just Wow.