AI/TLDR

Grok 4.7

SpaceXAI's September 2026 Grok flagship for coding and knowledge work, held at the same price as Grok 4.6.

Grok (flagship)API onlyAvailable
Released
21 Sep 2026
Context
500K
Input
$2.00 / 1M tokens
License
Proprietary
Coverage
1 story

Overview

Grok 4.7 is SpaceXAI's Grok flagship, announced on September 21, 2026 and described by the company as its most powerful model for coding and knowledge work. SpaceXAI says it runs on a larger base model than Grok 4.6, with extended reinforcement learning on harder and longer-running tasks, and that the result works longer on difficult problems, verifies its own output more reliably, and manages extended context better. The company also gave it native understanding of the Grok Bot interface.

The model keeps the 500,000-token context window of Grok 4.6, accepts text and image input, returns text, and has a knowledge cutoff of May 2026. Its API model id is grok-4.7. Reasoning effort is configurable as low, medium, high (the default), or xhigh, and the model supports function calling, web search, X search, and code execution through both the Responses API and Chat Completions.

Pricing did not move with the upgrade: $2.00 per 1M input tokens and $6.00 per 1M output tokens for prompts under 200K tokens, the same rates SpaceXAI charged for Grok 4.6, with a fast variant served at twice the speed for twice the cost. Grok 4.7 is available in Cursor, Grok Build, the SpaceXAI API, third-party coding harnesses, model routers and cloud platforms, and GitHub added it to Copilot on launch day for Pro, Pro+, Max, Business, and Enterprise plans.

Released2026-09-21
LicenseProprietary
WeightsAPI only
Context500K
Knowledge cutoffMay 2026
ModalitiesText, Vision
StatusAvailable

Benchmarks

Grok 4.7 against Grok 4.6 and rival frontier models, as published by SpaceXAI at launch.

BenchmarkGrok 4.7Grok 4.6GPT-5.6 SolFable 5.1
CursorBench 4.046.3%40.4%41.7%51.8%
DeepSWE v1.171%65.2%72.7%70%
EEBench64%53%39.4%56.4%
AA Briefcase v1.11657154614871678
Terminal-Bench 4.038%20.3%37.3%57.9%
Harvey Legal Agent Benchmark19.6%15.8%2.5%6.7%
HealthBench Professional56.7%48.5%60.5%62.1%

Comparison source ↗

This model's scores

  1. DeepSWE v1.171%
  2. CursorBench 4.046.3%
  3. EEBench64%
  4. Terminal-Bench 4.038%
  5. HealthBench Professional56.7%
  6. Harvey Legal Agent Benchmark19.6%
  7. AA Intelligence Index46

Scores on a 0–100 scale (25-point gridlines); higher is better. Each benchmark links to its published source.

Pricing

Input$2.00 / 1M tokens
Cached input$0.50 / 1M tokens
Output$6.00 / 1M tokens

Rates apply to prompts under 200K tokens; prompts of 200K tokens or more are billed at $4.00 input, $1.00 cached input, and $12.00 output per 1M tokens. A fast variant is served at twice the speed for twice the cost.

Pricing source ↗

Strengths

  • Beats Grok 4.6 on every benchmark SpaceXAI published — CursorBench 4.0 46.3% vs 40.4%, DeepSWE v1.1 71.0% vs 65.2%, Terminal-Bench 4.0 38.0% vs 20.3%
  • Leads GPT-5.6 Sol and Fable 5.1 on EEBench electrical engineering at 64.0%
  • Strongest published score on the Harvey legal agent benchmark in its comparison set at 19.6%
  • 500,000-token context window with text and image input, unchanged from Grok 4.6
  • No price increase over Grok 4.6: $2.00 / $6.00 per 1M input/output tokens under 200K-token prompts

Best for

  • Reach for it for coding agents in Cursor, Grok Build or GitHub Copilot that must hold a task across many steps.
  • Reach for it for engineering-heavy knowledge work — it posts the highest published EEBench score in its comparison set.
  • Reach for it for legal and document-heavy agent runs, where SpaceXAI reports a large lead on the Harvey benchmark.
  • Reach for it when you are already on Grok 4.6 and want the gains without changing your per-token budget.

How to access

ProviderModel ID
SpaceXAI API ↗grok-4.7
GitHub Copilot ↗grok-4.7

Grok (flagship) — every version

The full lineage of the Grok (flagship) line, newest first. Every version has its own page — click any to compare specs, benchmarks and pricing.

VersionReleasedContextLicense
Grok 4.7current2026-09-21500KProprietary
Grok 4.62026-08-12500KProprietary
Grok 4.52026-07-09Proprietary
Grok 4.32026-04-301MProprietary
Grok 4.202026-03Proprietary
Grok 4.12025-11-17Proprietary
Grok 42025-07-09Proprietary
Grok 32025-02-17Proprietary
Grok 22024-08-20Open weights
Grok 1.52024-05-15Proprietary
Grok 12023-11-03Apache-2.0

FAQ

When was Grok 4.7 released?

SpaceXAI announced Grok 4.7 on September 21, 2026. It launched the same day in Cursor, Grok Build, the SpaceXAI API, third-party coding harnesses, model routers and cloud platforms, and GitHub added it to Copilot on launch day for the Pro, Pro+, Max, Business, and Enterprise plans.

How much does Grok 4.7 cost?

SpaceXAI prices Grok 4.7 at $2.00 per 1M input tokens, $0.50 per 1M cached input tokens, and $6.00 per 1M output tokens for prompts under 200K tokens. Prompts of 200K tokens or more are billed at double those rates: $4.00 input, $1.00 cached input, and $12.00 output. A fast variant costs twice as much for twice the speed.

What is the Grok 4.7 context window?

Grok 4.7 has a 500,000-token context window, the same size SpaceXAI served for Grok 4.6. It accepts text and image input and returns text, with no stated output limit. Its knowledge cutoff is May 2026 and its API model id is grok-4.7.

How does Grok 4.7 compare with Grok 4.6?

Grok 4.7 leads on every benchmark in SpaceXAI's published table: CursorBench 4.0 46.3% against 40.4%, DeepSWE v1.1 71.0% against 65.2%, EEBench 64.0% against 53.0%, Terminal-Bench 4.0 38.0% against 20.3%, and HealthBench Professional 56.7% against 48.5%. Both models share the 500K context window and the same per-token price.

How does Grok 4.7 compare with Fable 5.1 and GPT-5.6 Sol?

On SpaceXAI's comparison Fable 5.1 leads on coding — 51.8% against 46.3% on CursorBench 4.0 and 57.9% against 38.0% on Terminal-Bench 4.0 — while GPT-5.6 Sol edges Grok 4.7 on DeepSWE v1.1 at 72.7% against 71.0%. Grok 4.7 leads both on EEBench at 64.0% and on the Harvey legal agent benchmark at 19.6%.