Friday, Sep 25 | --:--
Back to home

SpaceXAI Ships Grok 4.7 at Grok 4.6 Prices, $2 and $6 per Million Tokens

On September 21, 2026, SpaceXAI released Grok 4.7 for coding and knowledge work, served at the same $2 / $6 token prices as Grok 4.6, with a fast variant at twice the speed and twice the price. On the lab’s table it leads EEBench at 64.0% and Harvey’s legal-agent bench at 19.6%. A day-later independent vending test put its average take at $10,537, behind GPT-6 Sol.

Tech Insights Reporter 5 min read San Francisco, CA
Cover illustration for SpaceXAI Ships Grok 4.7 at Grok 4.6 Prices, $2 and $6 per Million Tokens

TLDR

SpaceXAI, September 21, 2026: Grok 4.7, which the lab calls its most capable model for coding and knowledge work. It uses a larger base model than Grok 4.6, a longer reinforcement-learning run weighted toward tasks that take many hours, and training to understand the Grok Bot harness natively.

Price: starting at $2 per million input tokens and $6 per million output tokens — the same sticker as Grok 4.6. A fast variant serves at twice the output speed and twice the price. The hero line says “twice as fast, at half the price of comparable models.” The body says it is served at the same price and speed as Grok 4.6. Read the half-price claim against peers (GPT-5.6 Sol at $4 / $20, Fable 5.1 at $10 / $50 on the same table), not as a cut versus 4.6.

Where: Cursor, Grok Build, the Grok API, third-party coding harnesses, and model routers and cloud platforms, the day of the post.

Lab table (SpaceXAI, Grok 4.7 vs Grok 4.6 / GPT-5.6 Sol / Fable 5.1):

Bench Grok 4.7 Grok 4.6 GPT-5.6 Sol Fable 5.1
CursorBench 4.0 46.3% 40.4% 41.7% 51.8%
DeepSWE v1.1 71.0% (high) 65.2% 72.7% 70.0%
EEBench 64.0% 53.0% 39.4% 56.4%
AA Briefcase v1.1 1,657 1,546 1,487 1,678
Terminal-Bench 4.0 37.6% 20.3% 37.3% 57.9%
Harvey Legal Agent 19.6% 15.8% 2.5% 6.7%
HealthBench Professional 56.7% 48.5% 60.5% 62.1%

Safety, lab-reported: a new safeguard stack. SpaceXAI says it tops LatchBio’s biosafety benchmark at 62.4%, and on HackerBench v0.3 lets 3.3% of risky dual-use cyber prompts through while rarely blocking legitimate security work. Select cybersecurity partners get invite-only red-team access.

Independent checks, not the lab table:

  • Andon Labs, posted September 24, 2026, ran Vending-Bench 2 (six runs, $500 and one simulated year). Grok 4.7 averaged $10,537. GPT-6 Sol averaged $14,428. Claude Opus 5.5 averaged $9,235. Andon says this is the first time a Grok model beat the newest Claude Opus on that bench, and that Opus 5.5 made less than Opus 5. In the four-game arena among the three, Sol finished well ahead and the other two ended within $200.
  • A public Arena Code Elo snapshot checked September 25 still listed Claude Opus 5 first and Grok 4.6, and did not show Grok 4.7. No Elo is claimed here.
  • Artificial Analysis comparison pages loaded in this pass scored Opus 5.5 and GPT-6 Sol, not a confirmed Grok 4.7 index row. Secondary write-ups citing AA put Grok 4.7 near 46. That figure is not repeated as a loaded board value.

Product-line de-dupe: not Voice Transcribe 2.0 (September 18). Not the September 22 post on using Grok Bot for customer support, which is an internal ops note and is not filed separately.

Why this story matters

The price did not drop against Grok 4.6. The movement is on long coding, electrical engineering, and a legal-agent bench where the lab’s own numbers put it ahead of Sol and Fable. The independent cash test a few days later put it in the middle, closer to Opus 5.5 than to Sol.

Sources

Prior Coverage

Earlier Times of AI reporting on this thread.

Scroll to continue reading