SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6
By Michal Sutter | September 21, 2026
SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built on a larger base model and a longer reinforcement learning run. It still ships at the same price and speed as Grok 4.6.
Is it deployable? Yesβas a hosted model. You can call grok-4.7 today through the xAI API, Cursor, Grok Build, OpenRouter, Vercel, and Cloudflare.
What Changed Under the Hood
SpaceXAI lists four changes over Grok 4.6:
- A new, larger base model: Grok 4.7 does not reuse the Grok 4.6 base.
- A longer RL run on harder tasks: The task mix is weighted toward problems that take many hours to complete.
- Better self-verification and long-context handling: The company says the model checks its own work more carefully.
- Native Grok Bot harness support: It was trained to understand the Grok Bot harness for conversational and knowledge work.
The developer docs list the API specs:
| Property | Value |
|---|---|
| Model name | grok-4.7 |
| Context window | 500,000 tokens |
| Knowledge cutoff | May 2026 |
| Modalities | Text and image input, text output |
| Reasoning effort | low, medium, high (default), xhigh |
| APIs | Responses API, Chat Completions |
| Tools | Function calling, web search, X search, code execution |
Benchmarks
The launch table compares Grok 4.7 at xHigh effort with Grok 4.6 High, GPT-5.6 Sol Max, and Fable 5.1 Max. The Grok 4.7 DeepSWE score was run at high effort. All scores are vendor-reported.
| Benchmark | Grok 4.7 xHigh | Grok 4.6 High | GPT-5.6 Sol Max | Fable 5.1 Max |
|---|---|---|---|---|
| Benchmark scores | Vendor-reported | Vendor-reported | Vendor-reported | Vendor-reported |
[Table content truncated in sourceβfull benchmark results available in the official launch post.]
via MarkTechPost
