AI/TLDR

xAI · 2026-09-21 · seismic

Grok 4.7 — xAI's most capable model for coding and knowledge work

Grok 4.7 is xAI's new Grok flagship, released 21 September 2026 for coding and knowledge work. It keeps Grok 4.6's 500K context and $2/$6 per million token price while lifting CursorBench 4.0 from 40.4% to 46.3%.

xAI announcement graphic for the Grok 4.7 model release

xAI's Grok 4.7 lifts coding and engineering scores over Grok 4.6 while holding the same $2/$6 price and 500K context.

Quick facts

MakerxAI
Model IDgrok-4.7
Context window500K tokens
Price (input)$2 / 1M tokens
Price (output)$6 / 1M tokens
Knowledge cutoffMay 2026
AvailabilityxAI API, Cursor, Grok Build, GitHub Copilot

Benchmarks

CursorBench 4.0
Grok 4.746.3%
Grok 4.640.4%
GPT-5.6 Sol41.7%
Fable 5.151.8%
source ↗
DeepSWE v1.1
Grok 4.771%
Grok 4.665.2%
GPT-5.6 Sol72.7%
Fable 5.170%
source ↗
EEBench
Grok 4.764%
Grok 4.653%
GPT-5.6 Sol39.4%
Fable 5.156.4%
source ↗
Terminal-Bench 4.0
Grok 4.738%
Grok 4.620.3%
GPT-5.6 Sol37.3%
Fable 5.157.9%
source ↗

Pricing

Input · Prompts under 200K tokens$2.00 / 1M tokens
Cached input · Prompts under 200K tokens$0.50 / 1M tokens
Output · Prompts under 200K tokens$6.00 / 1M tokens
Long context · Prompts of 200K tokens or more, input and output$4.00 / $12.00 / 1M tokens
source ↗

What is it?

Grok 4.7 replaces Grok 4.6 as xAI's top Grok model, released 21 September 2026. xAI built it on a larger base model than 4.6 and ran more reinforcement learning on harder, longer-running tasks. The company says the result is better at checking its own work and at managing long context. Its API model id is grok-4.7.

How does it work?

The extra reinforcement learning targets self-verification, so the model tests its own output before moving on. Reasoning effort is a dial — low, medium, high (the default), or xhigh — and Grok 4.7 supports function calling, web search, X search, and code execution. xAI also gave it native understanding of the Grok Bot interface, which the company says improves conversation. Knowledge cutoff is May 2026.

Why does it matter?

Price is what makes this release land: xAI kept the same $2 per million input tokens and $6 per million output tokens it charged for Grok 4.6, so teams already paying for 4.6 get the gains without a budget change. Grok 4.7 moves Terminal-Bench 4.0 from 20.3% to 38.0% and the Harvey legal agent benchmark from 15.8% to 19.6%. GitHub Copilot added it the same day.

Who is it for?

teams running coding agents

Frequently asked questions

How much does Grok 4.7 cost?
Grok 4.7 costs $2.00 per 1M input tokens, $0.50 per 1M cached input tokens, and $6.00 per 1M output tokens for prompts under 200K tokens. Prompts of 200K tokens or more are billed at double: $4.00 input, $1.00 cached input, and $12.00 output. A fast variant runs at twice the speed for twice the cost.
How does Grok 4.7 compare to Grok 4.6?
On xAI's published table Grok 4.7 beats Grok 4.6 on every benchmark listed: CursorBench 4.0 46.3% against 40.4%, DeepSWE v1.1 71.0% against 65.2%, EEBench 64.0% against 53.0%, and Terminal-Bench 4.0 38.0% against 20.3%. Both models share the same 500K context window and the same per-token price.
Is Grok 4.7 available in GitHub Copilot?
Yes. GitHub's 21 September 2026 changelog lists Grok 4.7 for Copilot Pro, Pro+, Max, Business, and Enterprise plans. The rollout is gradual, so not every user sees it at once. Business and Enterprise administrators control access through model policies in Copilot settings, where new models turn on automatically unless an admin disabled them.
Does Grok 4.7 beat Claude Fable 5.1 on coding?
Not on xAI's own numbers. Fable 5.1 scores 51.8% on CursorBench 4.0 against Grok 4.7's 46.3%, and 57.9% on Terminal-Bench 4.0 against 38.0%. Grok 4.7 leads on EEBench, an electrical engineering test, at 64.0% against 56.4%, and on the Harvey legal agent benchmark at 19.6% against 6.7%.

Try it

Set model grok-4.7 in the xAI API, or pick Grok 4.7 in Cursor or GitHub Copilot.

Sources · 4 outlets

Tags

  • grok-4-7
  • xai
  • llm
  • frontier-model
  • coding
  • agents
  • reasoning
  • api
  • cursor
  • github-copilot
  • benchmark

← All releases · Learn AI