Grok 4.7 vs SuperGrok Heavy: The New Model Is Not the Same Product

Grok 4.7 is a $2 and $6 API model launched 21 Sep 2026. SuperGrok Heavy is a separate $300 consumer tier. Here is which one you are buying.

Grok 4.7 and SuperGrok Heavy share a brand and almost nothing else. One is a text model that launched on 21 September 2026 at $2 per million input tokens and $6 per million output tokens. The other is a consumer subscription, listed around $300 a month, built around Grok 4 Heavy’s multi-agent mode. Buying the subscription does not hand you the new API model, and paying the API does not turn on Heavy’s video allowance. If a launch thread blurred those, this is the correction.

Side by side cards: Grok 4.7 the model versus SuperGrok Heavy the subscription
One is a token price. The other is a monthly plan built around Grok 4 Heavy and video. They are not substitutes.

Two products, two bills

Grok 4.7 SuperGrok Heavy
What it is A model: API, Cursor, Grok Build, Grok app A consumer plan above SuperGrok and Plus
List price $2 / $6 per million tokens About $300 a month, confirm on x.ai/pricing
Signature capability Larger base model, 500,000 token context, self-check on long tasks Grok 4 Heavy, up to 8 agents, about 500 video renders a day
Context 500,000 tokens, May 2026 cutoff Heavy’s published window is 256K, some sources say up to 428K
Images and video Reads images, writes text only Imagine image and video generation live on the plan
Confirmed on 21 Sep Yes, in Cursor, Build, the app, and the API No. Kingy still saw Grok 4.6 named on the consumer plan page

Kingy’s launch-day check is the line to remember. The consumer plan page was still naming Grok 4.6. A broad marketing sentence is not proof that web, mobile, or X chat flipped to 4.7. Open the model label in the session you are actually using.

What the new model scored

SpaceXAI compared Grok 4.7 at xHigh effort with Grok 4.6 at High, GPT-5.6 Sol at Max, and Claude Fable 5.1 at Max. DeepSWE for 4.7 is the exception, scored at high effort. It improved on 4.6 in every row OfficeChai transcribed.

EEBench called out at 64 percent, with Terminal-Bench still trailing at 38 and 26
These scores describe Grok 4.7 the model. A Heavy subscription badge does not include them.

The wins worth keeping: EEBench at 64.0 percent versus Fable’s 56.4 percent and Sol’s 39.4 percent, and a near-doubling on SpaceXAI’s own Terminal-Bench, 38.0 percent versus 20.3 percent for Grok 4.6. The losses worth keeping: Fable still leads CursorBench (51.8 to 46.3), HealthBench Professional (62.1 to 56.7), and that same terminal test by almost 20 points (57.9 to 38.0).

Artificial Analysis is less kind, and you should see both. Its Intelligence Index v4.3.2 puts Grok 4.7 at 46, with Fable 5.1 and GPT-6 at 53. Its Terminal-Bench 4.0 run is 26 percent, not 38, against 55 percent for Fable and 60 percent for GPT-6 Astra. High and xhigh effort tied on the index. I would not pick a subscription tier off either terminal number alone. I would pick it off the fact that both evaluators still rank the frontier Claude and GPT models ahead on agentic coding.

Office work is where 4.7 looks expensive to ignore

GDPval Elo is 1,695 for Grok 4.7, 1,735 for Fable 5.1, 1,605 for Grok 4.6, and 1,542 for GPT-6 Astra. AA-Briefcase, the multi-hour office task, is 1,657 against Fable’s 1,678. That is a 21-point Elo gap after a 111-point jump over Grok 4.6. For memos, decks, and long briefs, 4.7 is in the conversation with the leading Claude model and nowhere near “a cheap toy model.”

GDPval as two blocks, 1695 and 1735
A photo finish on office work. Neither number comes bundled with a Heavy login.

Harvey’s legal-agent score of 19.6 percent beats Fable’s 6.7 percent and Sol’s 2.5 percent. Quote it only with the absolute number attached. Nobody in that table is ready to practice law.

The price chart is the whole strategy

OfficeChai’s reading of the launch prices puts Sol at $4 input and $20 output, and Fable 5.1 at $10 and $50, per million tokens. Grok 4.7 stays at $2 and $6. Fast serving, inside Cursor and Grok Build only, is about twice that for short context ($4 and $12) and is not on the public API. Long prompts of at least 200,000 tokens rebill every token at $4 input and $12 output on the standard card.

A 13 cent API call beside a 300 dollar monthly Heavy plan
The API is a meter. Heavy is a ceiling. One does not pay for the other.

A $300 monthly subscription and a $2 token price answer different fears. The subscription is a ceiling: you stop thinking about tokens until you hit an unpublished chat cap or the video counter. The API is a meter: a 50,000-in, 5,000-out call is about $0.13, and an agent that repeats it ten times is $1.30 before tools. Heavy users of the API can spend more than $300 without ever seeing a plan named Heavy. Light users of Heavy can spend $300 and never touch the model that launched this week.

When Heavy is still the right buy

Keep Heavy in mind for a narrow job. You want Grok 4 Heavy’s multi-agent pass, up to eight collaborating reasoners, on a problem where a single confident wrong answer is expensive. You want the larger weekly usage pool and priority at peak. You want Imagine, on the order of 500 video renders a day, which 4.7 the text model cannot do at all. You want one login instead of an API key and a cost dashboard.

Skip Heavy if the job is code in Cursor, a long document, or anything you can already point at grok-4.7. The 500,000 token window on 4.7 is larger than Heavy’s older 256K figure, the token price is the cheapest serious card in this comparison, and the launch benchmarks above are about this model, not about the subscription badge.

xAI has changed consumer prices more than once in 2026. Before you pay list, read x.ai/pricing and the model name inside the product. If the session still says 4.6, you are not testing 4.7, however new the homepage looks.

A month is enough to tell them apart

Run the same three tasks on both surfaces if you have them: a hardware or EE-style problem, a multi-hour brief, and a terminal chore that has to call tools. Score pass or fail, edits required, and dollars spent. If the text model does the work, stop paying for a video-and-multi-agent plan you are not using. If the multi-agent pass is the thing that actually unsticks the bug, the API model was never the substitute.

For a bounded test of the consumer tier, this site lists a SuperGrok Heavy 1 month plan at $120 against the $300 reference price, with a 3 day warranty. It is an independent activation, not an xAI sale. It does not include API credit, and it is not a promise that the session will show Grok 4.7. Check the label.

Frequently asked questions

Does SuperGrok Heavy include Grok 4.7?

Not as a documented fact on launch day. Grok 4.7 shipped in Cursor, Grok Build, the Grok app, and the API. Kingy checked the consumer plan page and still saw Grok 4.6 named. Look at the model in your session rather than assuming the top subscription moved.

Is Grok 4.7 the same thing as Grok 4 Heavy?

No. Grok 4 Heavy is the multi-agent model associated with the Heavy plan, with up to eight collaborating agents. Grok 4.7 is a newer, larger general model with a 500,000 token window and a public token price. The version numbers sit next to each other and describe different products.

Which is smarter, 4.7 or Fable 5.1?

Fable 5.1 leads the combined Artificial Analysis index, 53 to 46, and leads the vendor coding boards that reward long terminal work. Grok 4.7 leads EEBench in the vendor table and sits 21 Elo behind Fable on AA-Briefcase. Cheaper, close on office work, behind on agentic coding.

Can I use a Heavy subscription as an API key?

No. Consumer subscriptions and the developer API are billed separately. A $300 plan does not load API credit, and API traffic does not draw down the subscription allowance.

What should I confirm before paying?

The live price on x.ai/pricing, the model name in the product, and whether you need video generation. If you need text and code, price grok-4.7 first. If you need multi-agent reasoning plus Imagine, you are shopping for Heavy, and the 4.7 launch does not retire that plan.

Sources