# Grok 4.7 holds its price while three rivals cut theirs — the bill is per task, not per token

> xAI released Grok 4.7 on 21 September 2026 at Grok 4.6's unchanged $2/$6 per million tokens.

*xAI's new model beats its predecessor on every benchmark it published, at the same $2/$6 sticker. Artificial Analysis says finishing an average task now takes more than twice the tokens.*

By Behzad Hosseini · WireRead
Canonical: https://wireread.com/news/grok-4-7-price-per-task-analysis

xAI released Grok 4.7 on 21 September 2026, and the number that did not move is the interesting one. The model is priced exactly where Grok 4.6 was five weeks ago — $2 per million input tokens, $6 per million output — while the company describes it as a genuinely larger, longer-trained system. Cursor, Grok Build and the Grok API had it live within hours; GitHub began rolling it out across every paid Copilot tier the same day. On the surface, that is a straightforward good-news release: more capability, no price rise. The more useful question, and one xAI's launch page does not answer for you, is what a finished task now costs once every token it actually uses is counted.

## What xAI changed under the hood

xAI says Grok 4.7 runs on a new, larger base model than Grok 4.6 and was trained with a longer reinforcement-learning run on a harder mix of tasks, weighted towards problems that take many hours to complete. The company frames the gain as endurance and self-checking rather than raw cleverness — the same framing it used for Grok 4.6 in August.

> Grok 4.7 uses a new, larger base model compared to Grok 4.6.
> — [xAI](https://x.ai/news/grok-4-7), 2026-09-21

xAI also says it trained the model to natively understand the harness behind Grok Bot, its own agent product, which it credits with better conversational and general-knowledge performance — a sign the lab is now tuning models around its own tooling rather than treating the raw API as the only surface that matters.

> It was trained with a longer reinforcement learning run on a harder mix of tasks, weighted toward problems that take many hours to complete.
> — [xAI](https://x.ai/news/grok-4-7), 2026-09-21

## The benchmark table, and the index version that changed underneath it

On xAI's own seven-benchmark comparison, Grok 4.7 improves on Grok 4.6 across the board — CursorBench 4.0, DeepSWE v1.1, EEBench, AA Briefcase v1.1, Terminal-Bench 4.0, the Harvey Legal Agent Benchmark and HealthBench Professional. Against Claude Fable 5.1 Max, the picture is split: Fable still leads on CursorBench 4.0 (51.8% to 46.3%), Terminal-Bench 4.0 (57.9% to 37.6%), GDPval (1,735 to 1,695 Elo) and HealthBench Professional, while Grok 4.7 leads Fable on EEBench and the Harvey benchmark.

| Model | Input / output per M tokens | CursorBench 4.0 | Terminal-Bench 4.0 |
| --- | --- | --- | --- |
| Claude Fable 5.1 Max | $10 / $50 | 51.8% | 57.9% |
| **Grok 4.7 xHigh** | **$2 / $6** | **46.3%** | **37.6%** |
| GPT-5.6 Sol Max | $4 / $20 | 41.7% | 37.3% |
| Grok 4.6 High | $2 / $6 | 40.4% | 20.3% |

Independent testing complicates one of those rows. VentureBeat reported that Artificial Analysis, running its own harness rather than xAI's, measured Grok 4.7 at roughly 26% on Terminal-Bench 4.0 at its xHigh setting — well below xAI's own reported 37.6%, and also behind GPT-6 Astra's independently measured 59.6% on the same test. The gap is not necessarily dishonesty; different harnesses, prompts and reasoning-effort settings routinely produce different numbers for the same model. It is, though, a reminder that a vendor's own comparison table is a marketing document with real numbers in it, not a substitute for testing on your own workload.

On Artificial Analysis's Intelligence Index — a ten-evaluation composite now at version 4.3.2 — Grok 4.7 scored 46, two points ahead of Grok 4.6's 44 on the same version. That is a genuine improvement, but it is worth being precise about what it is not: our coverage of Grok 4.6's launch in August quoted it at 61 on the Intelligence Index. That was a different, earlier version of the same index. Artificial Analysis periodically rebuilds the composite's evaluations and reweights them, and scores from different versions are not comparable — a model's number can fall even as its real capability rises, simply because the ruler changed. Always check the version before comparing two Intelligence Index scores.

## The number that actually decides the bill

Here is the more consequential finding. Artificial Analysis measured Grok 4.7, at its highest reasoning setting, using roughly 81,000 output tokens to complete an average Intelligence Index task — against about 36,000 for Grok 4.6 at a comparable setting. The rate card did not move. The number of tokens spent reaching an answer did, by well over double.

> A model charging less per token can still be more expensive on a finished workload if it needs substantially more reasoning tokens to get there.
> — [VentureBeat](https://venturebeat.com/technology/grok-4-7-pairs-coding-gains-with-the-same-affordable-pricing-but-high-token-consumption-threatens-real-world-roi), 2026-09-21

Put in dollars, Artificial Analysis's own cost-per-task figures make the point cleanly: Grok 4.7 costs roughly $3.74 to finish an average Index task at its highest setting, and about $2.73 at a lower one — while GPT-5.6 Sol Max, the model xAI's own launch chart benchmarks against at $4/$20 list pricing, comes out at around $1.99 per task on the same measure. The model with the lower sticker price is not, on this evidence, the model with the lower bill. That is the trap in reading a 'price per million tokens' figure as if it were the price of the work.

> **Key:** Cost per completed task, not price per token, is now the honest unit for agentic and coding work — because reasoning effort, tool calls and retries all sit between the rate card and the invoice. A model that halves the sticker price can still double the bill if it needs several times the tokens to finish the same job.

The timing sharpens the point. A day after Grok 4.7 shipped, Anthropic released Claude Opus 5.5 at $4 input / $20 output per million tokens — 20% below Opus 5 — and OpenAI released GPT-6 Sol at $2/$10 and GPT-6 Luna at $0.10/$0.50, each roughly half the prior generation's rate. Of the three labs that moved inside 48 hours, xAI was the only one that held its price rather than cutting it, betting that a bigger model at the same rate is itself the competitive move. Whether that reads as confidence or as a company with less room to cut depends on how the token-efficiency story lands with buyers running Grok 4.7 at volume.

The practical read for anyone routing production traffic: xAI's rate card is real and unchanged, and the model is a genuine improvement on Grok 4.6 by its own seven-benchmark table. But 'same price, better model' is only half the sentence. The other half is that finishing a job now appears to cost meaningfully more in tokens than it did five weeks ago, at exactly the moment two well-funded rivals cut their own rates. Model the cost of a completed task on your own workload before treating the $2/$6 sticker as the number that matters.

## Key takeaways

- xAI released Grok 4.7 on 21 September 2026 at Grok 4.6's unchanged $2 input / $6 output per million tokens, with GitHub beginning a phased Copilot rollout the same day.
- It beats Grok 4.6 on all seven benchmarks in xAI's own comparison table, though Claude Fable 5.1 Max still leads on CursorBench 4.0, Terminal-Bench 4.0, GDPval and HealthBench Professional.
- Artificial Analysis scored it 46 on the Intelligence Index v4.3.2 against Grok 4.6's 44 on the same version — a two-point gain that is not comparable to the 61 our Grok 4.6 coverage quoted on an earlier index version.
- The same testing found Grok 4.7 needs roughly 81,000 output tokens to finish an average Index task, against about 36,000 for Grok 4.6 — more than double the token bill behind an unchanged rate card.
- The launch landed the day before Anthropic's Claude Opus 5.5 and OpenAI's GPT-6 Sol and Luna, each repriced downward, leaving Grok 4.7 as the only one of the three that held its price flat.

## FAQ

### How much does Grok 4.7 cost?
$2 per million input tokens and $6 per million output under 200,000-token prompts, doubling to $4/$12 above that — unchanged from Grok 4.6. A Cursor/Grok Build-only fast tier costs twice the standard rate.

### Is Grok 4.7 better than Grok 4.6?
On xAI's own seven-benchmark table, yes, on all seven. Independent testing from Artificial Analysis broadly agrees but puts its Terminal-Bench 4.0 score well below xAI's own figure.

### Why does Grok 4.7 score lower than Grok 4.6 did on the Artificial Analysis Intelligence Index?
It doesn't — Grok 4.6 was quoted at 61 on an earlier version of the index. Grok 4.7 scores 46 against Grok 4.6's 44 on the current version, v4.3.2. Scores from different index versions are not comparable.

### Does Grok 4.7 cost more to run than Grok 4.6 despite the same price?
Per finished task, likely yes: Artificial Analysis measured roughly 81,000 output tokens per task for Grok 4.7 against about 36,000 for Grok 4.6, more than doubling the real cost of a task at an unchanged token price.

### How does Grok 4.7 compare on price to Claude Opus 5.5 and GPT-6 Sol?
It is far cheaper on the rate card — $2/$6 against Opus 5.5's $4/$20 and GPT-6 Sol's $2/$10 — though Artificial Analysis's cost-per-task figures narrow that gap once token usage is counted.

## Sources

- [Introducing Grok 4.7](https://x.ai/news/grok-4-7) — xAI, 2026-09-21
- [Grok 4.7 is now available in GitHub Copilot](https://github.blog/changelog/2026-09-21-grok-4-7-is-now-available-in-github-copilot/) — GitHub, 2026-09-21
- [Grok 4.7 scores 46 on the Artificial Analysis Intelligence Index to bring SpaceXAI into the top 4 AI labs](https://artificialanalysis.ai/articles/benchmarking-grok-4-7) — Artificial Analysis, 2026-09-21
- [Grok 4.7 pairs coding gains with the same affordable pricing — but high token consumption threatens real-world ROI](https://venturebeat.com/technology/grok-4-7-pairs-coding-gains-with-the-same-affordable-pricing-but-high-token-consumption-threatens-real-world-roi) — VentureBeat, 2026-09-21
- [Grok 4.7 Brings Big Coding Upgrades to Challenge Claude AI at Unchanged Pricing](https://www.androidheadlines.com/2026/09/grok-4-7-ai-launch-coding-upgrades-pricing.html) — Android Headlines, 2026-09-21
- [Grok 4.7 pricing: API token costs, tiers, and the Fast tax](https://www.eesel.ai/blog/grok-4-7-pricing) — eesel AI, 2026-09-23
- [xAI launches Grok 4.7 at $2 per million input tokens](https://tbreak.com/xai-grok-4-7-launch-price-benchmarks/) — tbreak, 2026-09-22
